Data Engineer
- Composer
- Airflow
- dbt
- BigQuery
- GCP
- PostgreSQL
- CI/CD
- GitHub Actions
- OpenTofu
- Terragrunt
- IaC
- Apache Airflow
- Data Modeling
- Python
- SQL
- Git
- Terraform
- Pub/Sub
- FinOps
Hello! We’reCashea 👋, and we’re on a mission to give Venezuelans back the opportunity to access credit through a BNPL business model. Since our launch in 2022, we’ve been dedicated to promoting financial inclusion. Today we have more than9 million active users, both consumers and merchants, and we’ve become a trusted brand in Venezuela, winning hearts and minds 💛.
About the role
We’re looking for a Senior Data Engineer who owns the platform that turns raw operational data into trusted, analytics-ready data the whole company relies on. You’ll keep our orchestration and transformation layer scaling as we grow, with deep hands-on expertise acrossCloud Composer (Airflow),dbt, andBigQuery onGCP, and you’ll be the person who makes good data engineering practice the norm across teams.
We run a medallion lakehouse on GCP:Datastream CDC pipelines replicate dozens of PostgreSQL databases intoBigQuery bronze datasets, anddbt on BigQuery transforms them through bronze, silver, and gold layers. Orchestration runs onCloud Composer / Airflow. This role is part deep building and part enablement: you’ll scale Airflow, harden the platform, optimize cost and performance, and leave every team more capable of shipping reliable data than it was before.
Overall you will…
Own the health, performance, and scalability of ourCloud Composer (Airflow) environments, including environment sizing, worker autoscaling, DAG parse performance, and orchestration reliability across staging and production.
Design and extend ourdbt on BigQuery transformation layer: incremental and CDC merge models, deduplication and SCD patterns, macros, tests, and source freshness across the bronze, silver, and gold layers.
Own ourDatastream CDC pipelines feeding BigQuery, keeping the PostgreSQL to BigQuery flow healthy across schemas, backfills, and replication.
Build and improve ourDataOps: CI/CD for DAGs and dbt withGitHub Actions and Workload Identity Federation, automated testing, deploy smoke tests, and safe promotion across environments.
Drive data quality and observability, freshness checks, lineage, dbt-artifact capture, and alerting so issues surface before consumers notice them.
Define and champion data engineering standards for model design, incrementality, idempotency, and safe change processes, and mentor engineers through reviews, pairing, and documentation.
Partner with platform, product, and analytics teams as the bridge that makes the data platform self-serve, working within ourOpenTofu andTerragrunt Infrastructure-as-Code.
Qualifications
Bachelor’s degree in Computer Science or a related field, or equivalent practical experience.
5+ years in data engineering or a closely related role, building and operating production data platforms at scale.
Strong hands-on experience withApache Airflow (ideallyCloud Composer): DAG design, operators, scheduling, and scaling orchestration as pipeline counts grow.
Strong hands-on experience withdbt, including incremental models, testing, macros, and managing a large model graph.
Solid hands-on experience withBigQuery: data modeling, partitioning and clustering, query optimization, and cost control.
Solid experience operating data services onGoogle Cloud Platform (GCP).
Proficiency inPython for pipeline and operator development, and strongSQL.
Comfort withGit-based workflows and CI/CD (GitHub Actions or similar), treating data pipelines with the same rigor as application code.
Strong communication and stakeholder skills, with a genuine teaching mindset and the ability to mentor others.
Excellent problem-solving and analytical thinking, with genuine curiosity about new technology.
Desirable skills
Experience withDatastream or other CDC tooling feeding analytics warehouses, including backfills and replication troubleshooting.
Familiarity withInfrastructure-as-Code (Terraform or OpenTofu, ideally Terragrunt) for managing data resources.
Experience withDataplex or other data quality, governance, and lineage tooling, and with PII handling in a regulated context.
Familiarity withastronomer-cosmos for running dbt within Airflow.
Knowledge of GCP data and messaging services such as Pub/Sub and Cloud Storage.
A track record of cost and performance optimization in a data warehouse, including FinOps practices.
Experience building in fintech, payments, or credit, where data correctness matters.
Spanish and English proficiency.
Why you’ll love working at Cashea
At Cashea, we have a work culture based on trust and purpose. If you need a clue as to why we are a good choice, these are our core values:
We don’t work on autopilot. Everything we do is intentional. We love to develop ideas with full awareness of the impact they can have on our users.
Your creativity and curiosity are our most important assets.
Your voice matters. We listen and make space for ideas and feedback. Everyone belongs, and what’s important to you is important to us.
We value transparency. Clarity keeps us connected and grounded.
Last but not least, we focus on real impact.
If you want to work with us, fill out the application. We’d love to meet you!
Data Engineer · Cashea