Databricks_Hyderabad
- Databricks
- Apache Spark
- Delta Lake
- ETL
- ELT
- PySpark
- SQL
- Scala
- Java
- AWS
- Azure
- GCP
-
Design and build scalable data pipelines usingDatabricks,Apache Spark, andDelta Lake.
-
Develop and optimizeETL/ELT processes to ingest data from multiple sources (structured, semi-structured, unstructured).
-
Write and maintain clean, reusable code inPySpark,SQL, and optionallyScala orJava.
-
Implement data quality checks, validation frameworks, and pipeline monitoring.
-
Collaborate with data scientists, analysts, and business stakeholders to understand data needs.
-
Integrate Databricks workflows with cloud platforms (AWS, Azure, or GCP).
-
Tune and optimize Spark jobs for performance and cost-efficiency.
-
Maintain documentation of data flows, schema designs, and system architecture
Databricks_Hyderabad · Diverse Lynx India