Jobless Developer
C

Posted 11 days ago

Open

Data Engineer

Buenos AiresOn-site

AI Summary

A Senior Data Engineer who owns and scales the medallion lakehouse on GCP, ensuring reliable orchestration (Airflow/Cloud Composer), transformations (dbt), and CDC pipelines (Datastream) from PostgreSQL to BigQuery.

About this role

Hello! We’re Cashea 👋, and we’re on a mission to give Venezuelans back the opportunity to access credit through a BNPL business model. Since our launch in 2022, we’ve been dedicated to promoting financial inclusion. Today we have more than 9 million active users, both consumers and merchants, and we’ve become a trusted brand in Venezuela, winning hearts and minds 💛.

About the role

We’re looking for a Senior Data Engineer who owns the platform that turns raw operational data into trusted, analytics-ready data the whole company relies on. You’ll keep our orchestration and transformation layer scaling as we grow, with deep hands-on expertise across Cloud Composer (Airflow), dbt, and BigQuery on GCP, and you’ll be the person who makes good data engineering practice the norm across teams.

We run a medallion lakehouse on GCP: Datastream CDC pipelines replicate dozens of PostgreSQL databases into BigQuery bronze datasets, and dbt on BigQuery transforms them through bronze, silver, and gold layers. Orchestration runs on Cloud Composer / Airflow. This role is part deep building and part enablement: you’ll scale Airflow, harden the platform, optimize cost and performance, and leave every team more capable of shipping reliable data than it was before.

Overall you will…

  • Own the health, performance, and scalability of our Cloud Composer (Airflow) environments, including environment sizing, worker autoscaling, DAG parse performance, and orchestration reliability across staging and production.

  • Design and extend our dbt on BigQuery transformation layer: incremental and CDC merge models, deduplication and SCD patterns, macros, tests, and source freshness across the bronze, silver, and gold layers.

  • Own our Datastream CDC pipelines feeding BigQuery, keeping the PostgreSQL to BigQuery flow healthy across schemas, backfills, and replication.

  • Build and improve our DataOps: CI/CD for DAGs and dbt with GitHub Actions and Workload Identity Federation, automated testing, deploy smoke tests, and safe promotion across environments.

  • Drive data quality and observability, freshness checks, lineage, dbt-artifact capture, and alerting so issues surface before consumers notice them.

  • Define and champion data engineering standards for model design, incrementality, idempotency, and safe change processes, and mentor engineers through reviews, pairing, and documentation.

  • Partner with platform, product, and analytics teams as the bridge that makes the data platform self-serve, working within our OpenTofu and Terragrunt Infrastructure-as-Code.

Qualifications

  • Bachelor’s degree in Computer Science or a related field, or equivalent practical experience.

  • 5+ years in data engineering or a closely related role, building and operating production data platforms at scale.

  • Strong hands-on experience with Apache Airflow (ideally Cloud Composer): DAG design, operators, scheduling, and scaling orchestration as pipeline counts grow.

  • Strong hands-on experience with dbt, including incremental models, testing, macros, and managing a large model graph.

  • Solid hands-on experience with BigQuery: data modeling, partitioning and clustering, query optimization, and cost control.

  • Solid experience operating data services on Google Cloud Platform (GCP).

  • Proficiency in Python for pipeline and operator development, and strong SQL.

  • Comfort with Git-based workflows and CI/CD (GitHub Actions or similar), treating data pipelines with the same rigor as application code.

  • Strong communication and stakeholder skills, with a genuine teaching mindset and the ability to mentor others.

  • Excellent problem-solving and analytical thinking, with genuine curiosity about new technology.

Desirable skills

  • Experience with Datastream or other CDC tooling feeding analytics warehouses, including backfills and replication troubleshooting.

  • Familiarity with Infrastructure-as-Code (Terraform or OpenTofu, ideally Terragrunt) for managing data resources.

  • Experience with Dataplex or other data quality, governance, and lineage tooling, and with PII handling in a regulated context.

  • Familiarity with astronomer-cosmos for running dbt within Airflow.

  • Knowledge of GCP data and messaging services such as Pub/Sub and Cloud Storage.

  • A track record of cost and performance optimization in a data warehouse, including FinOps practices.

  • Experience building in fintech, payments, or credit, where data correctness matters.

  • Spanish and English proficiency.

Why you’ll love working at Cashea

At Cashea, we have a work culture based on trust and purpose. If you need a clue as to why we are a good choice, these are our core values:

  • We don’t work on autopilot. Everything we do is intentional. We love to develop ideas with full awareness of the impact they can have on our users.

  • Your creativity and curiosity are our most important assets.

  • Your voice matters. We listen and make space for ideas and feedback. Everyone belongs, and what’s important to you is important to us.

  • We value transparency. Clarity keeps us connected and grounded.

  • Last but not least, we focus on real impact.

If you want to work with us, fill out the application. We’d love to meet you!

Skills

Apache AirflowBigQueryCI/CDCloud ComposerDatastreamDbtGCPOpenTofuPostgreSQLPythonSQLTerragrunt

Explore related jobs

Browse these categories