Likeremote

Subscribe to the latest remote jobs:

  • Likeremote jobs on https://LinkedIn.com/
  • Likeremote jobs on https://telegram.org/
  • Likeremote jobs on Reddit.com
ES

Senior Data Software Engineer

EPAM Systems
๐Ÿ‡ง๐Ÿ‡ท Brazil | ๐Ÿ‡ฒ๐Ÿ‡ฝ Mexico | ๐Ÿ‡ฆ๐Ÿ‡ท Argentina | ๐Ÿ‡จ๐Ÿ‡ฑ Chile | ๐Ÿ‡จ๐Ÿ‡ด Colombia
Remote
Senior
2 days ago
  • BigQuery
  • Snowflake
  • Delta Lake
  • Databricks
  • Pandas
  • RBAC
  • Python
  • GCP
  • Data Modeling
  • Apache Iceberg
  • Apache Spark
  • Unity Catalog
  • AI
  • Claude Code
  • Cursor
Not scoredNo CV on file. Upload one and this job gets a score out of 100.Upload CV

We are seeking aSenior Data Software Engineer to build reusable, governed data-sharing adapters that connect a cloud lakehouse and analytics warehouse to external data platforms while meeting strict access controls and SLAs. You will design lakehouse patterns, implement integrations, and strengthen lineage and governance.

Responsibilities

  • Design a UniForm write layer with one physical dataset and dual-format metadata for Delta and Iceberg consumers
  • Build and validate GCS-to-BigQuery ingestion pipeline patterns for structured operational data
  • Implement Kafka-based CDC patterns for real-time and near-real-time movement into the lakehouse
  • Develop dependency-aware bookkeeping and data lineage tracking patterns across pipelines
  • Engineer modular, version-controlled adapter code designed for reuse across new integrations
  • Configure Iceberg external table definitions in Snowflake using Horizon Catalog governance features
  • Validate zero-copy read access from Snowflake to Iceberg and Delta tables without data movement
  • Implement tenant-scoped access controls aligned with Snowflake metadata governance requirements
  • Implement and certify a Delta Sharing adapter for live, zero-copy sharing from Delta Lake to Databricks consumers
  • Configure Delta Sharing endpoints and manage sharing agreements for Databricks access
  • Validate Databricks read access via Delta Sharing for Spark, Pandas, and compatible consumers
  • Test end-to-end freshness and sharing latency to meet agreed SLA targets
  • Register connector types and implement RBAC plus tenant-scoped authorization for all data-out paths
  • Implement metering hooks compatible with billing requirements for governed data-out flows

Requirements

  • 3+ years of data engineering experience with Python and cloud data platforms
  • Experience with Google Cloud BigQuery in advanced analytics and data modeling
  • Experience with lakehouse table formats including Apache Iceberg and Delta Lake
  • Strong leadership skills to drive integration designs, technical decisions, and delivery ownership
  • Proven project execution skills delivering data pipelines and adapters that meet SLAs
  • Advanced hard skills in Kafka/CDC patterns for real-time and near-real-time data movement
  • Strong architecture skills in data lake and lakehouse ingestion patterns on object storage
  • Excellent collaboration skills to work across platform, governance, and downstream consumer needs
  • Upper-Intermediate English proficiency (B2) for technical discussions and documentation

Nice to have

  • Apache Spark experience for validation and consumption testing
  • Databricks Unity Catalog knowledge for governed access patterns
  • Delta Lake expertise including Delta Sharing configuration and troubleshooting
  • Gen AI Assisted Development proficiency with tools such as Claude Code, GitHub Copilot, or Cursor
  • Snowflake Horizon Catalog experience for metadata governance and external table management

Senior Data Software Engineer ยท EPAM Systems

Auto apply with Likeremote