Description
Caju is a **Brazilian technology company** that aims to add more flavor to professional life by transforming the relationship between companies and employees through more innovative and secure solutions such as the Multi-Benefits Card, Corporate Expense Management Solution, Rewards Programs, and Caju Cycles.
At Caju, we **constantly learn**, and continuously improve within a collaborative and fun environment!
Applications from Black/Black-identifying individuals, women, Indigenous peoples, LGBTQIA\+, or other marginalized groups are warmly welcomed.
Apply now and learn more about our team
**Responsibilities and Duties**
* Design and maintain robust data models using DBT, applying best development practices such as testing, documentation, and model versioning.
* Develop fact tables, dimension tables, and data marts aligned with business unit needs, following dimensional modeling standards (Star Schema/Snowflake Schema).
* Serve as the liaison between business teams and data engineering, gathering requirements, interpreting business rules, and translating analytical demands into reliable and scalable data solutions.
* Document developed models, including metric definitions, applied business rules, data glossaries, and lineage, ensuring traceability and organizational-wide understanding.
* Implement and maintain orchestration pipelines in Databricks and Airflow, ensuring data flow reliability and monitoring.
* Develop Python and SQL solutions for data transformation, processing, and aggregation within the Databricks environment.
* Implement data quality tests in DBT models and analytical layers, ensuring consistency, completeness, and accuracy of information delivered to business units.
* Establish monitoring and alerts for pipelines and data models, proactively identifying and resolving inconsistencies.
* Lead end-to-end complex modeling projects—from understanding business pain points to delivering analytical layers into production.
* Apply appropriate partitioning, clustering, and materialization strategies to ensure performance and cost efficiency in data processing.
* Collaborate with product and business squads to ensure data solutions support both strategic and operational decision-making across the company.
* Use GitHub for code versioning, pull request reviews, and maintaining engineering best practices within the team.
**Requirements and Qualifications** **Data Modeling and Architecture**
* Proficiency in analytical data modeling: Star Schema, Snowflake Schema, and OBT (One Big Table).
* Solid knowledge of layered architectures such as Medallion Architecture (Bronze, Silver, Gold).
* Ability to define and enforce data contracts across layers, ensuring consistency and predictability for consumers.
**Performance and Optimization**
* Experience with DBT materialization strategies (tables, views, incremental models, snapshots), selecting the most appropriate approach per layer and data volume.
* Knowledge of data partitioning by date columns or other high-cardinality keys to reduce query read volume and optimize cost and performance in Databricks.
* Familiarity with clustering and Z\-Ordering techniques in Databricks/Delta Lake for optimizing reads on large-volume tables.
* Ability to identify and resolve SQL query performance bottlenecks—e.g., avoiding full table scans, efficient use of joins, aggregations, and window functions.
* Ability to estimate and manage the impact of complex models on cloud processing costs, proposing solutions balancing performance and operational efficiency.
**Tools and Technologies**
* Proficiency in DBT (dbt Core or dbt Cloud), including model creation, testing, macros, and documentation.
* Experience with Databricks and Delta Lake, including features such as Time Travel, VACUUM, and OPTIMIZE.
* Advanced SQL skills for building complex queries and analytical modeling.
* Python knowledge for automation, data transformation, and developing supporting scripts for pipelines.
* Experience orchestrating pipelines using Apache Airflow and Databricks.
* Experience with code versioning using GitHub, including review workflows and team collaboration.
**Behavioral Skills**
* Critical thinking and autonomy to lead complex projects involving multiple stakeholders.
* Ability to communicate clearly with non-technical business areas, translating analytical needs into data solutions.
* Strong ownership mindset regarding the quality and reliability of delivered data.
**Nice-to-Haves**
* Experience with Unity Catalog or other data governance and cataloging tools.
* Knowledge of data observability tools such as Elementary or Monte Carlo.
* Familiarity with BI tools like Metabase, Explo, GoodData, Luzmo, or Power BI, understanding how analytical teams consume models.
* Knowledge of Terraform or Infrastructure-as-Code (IaC) for provisioning cloud data environments.
* Experience in fintech and benefits environments.
**Additional Information**
Caju Card—with greater freedom to use your benefits (Meal, Food, Mobility, Health, Home Office, Culture, and Education);
Health Plan with no copayment (Unimed, Sulamerica, or Alice);
Zenklub—online therapy and coaching sessions to support your mental health;
Wellhub;
We also encourage language learning through our partnership with Rosetta Stone;
Recharge Day—paid day off;
Conexa Saúde—online medical consultations;
Childcare Assistance;
Partnership with Alura;
Remote Work—work from anywhere in Brazil;
We provide work equipment;
Many growth opportunities—we have ambitious goals and strongly hope you’ll help us achieve them!
Caju is a Brazilian company and welcomes candidates from all regions of the country.
**We operate in a fully remote model** (if you’re interested in visiting or working at our São Paulo office, we’ll keep our doors open)
**Interested? \#JoinCaju**
Caju is a **technology company**, **born from Brazilian entrepreneurship**, created to transform the relationship between companies and employees through innovative and flexible solutions.
We’ve built a platform enabling companies to manage multiple solutions in one place—and for employees, a single card offering diverse possibilities.
**Our Products:**
**Multi-Benefits Card:** A solution for managing up to eight benefit categories on a single Caju Visa-branded card. It includes an easy-to-use app for employees and an intuitive HR platform for service management.
**Caju Expenses:** A faster advance workflow, corporate expense management (travel, office supplies, third-party services, and daily operational expenses), and linked Caju cards—all within one platform. Farewell reimbursements; hello pocket-friendly convenience!
**Caju Rewards:** A solution to reward employees with extra funds loaded onto their card in a dedicated wallet.
**Caju Cycles:** End-to-end employee journey tracking—from onboarding to data management—in a single unified platform.
**Caju Plus:** For companies aiming to improve employee health and wellbeing, in partnership with Gympass, Conexa Saúde, and Psicologia Viva.
**Caju Fair:** Exclusive offers and coupons at Caju’s partner stores for employees and HR teams.
We dream big here—and we’re proud of our **roots spread across many parts of Brazil**, and of the people who nurture and harvest fruits with a unique flavor, found only here.
Having trouble applying or encountering any issues with the Gupy platform? Access support here.
**How about building our cashew tree together?**
**\#JoinTheCashewTree**