
Data Engineer
- PySpark
- Scala
- Python
- Apache Spark
- GCP
- Azure
- AWS
- ETL
- Git
- EMR
- BigQuery
- Airflow
- Composer
- CI/CD
Job Title: Data Engineer (PySpark / Scala / Python)
Location:Remote
Job Description:
We are hiring aData Engineer with strong hands-on experience inPySpark,Scala, andPython. You must have solid expertise inApache Spark, as it will be the core technology used for building and managing large-scale data processing pipelines.
Experience with cloud platforms likeGoogle Cloud Platform (GCP),Microsoft Azure, orAWS is a plus.
Required Skills:
Strong hands-on experience withApache Spark
Proficient inPySpark
Experience inScala andPython
Knowledge ofETL processes and data pipeline design
Understanding of distributed data processing
Familiarity with version control tools likeGit
Basic knowledge ofcloud platforms (GCP, AWS, or Azure)
Nice to Have:
Experience with cloud-native data tools (e.g., Dataproc, Glue, EMR, BigQuery)
Familiarity with workflow/orchestration tools likeAirflow orCloud Composer
Experience with CI/CD for data engineering
Exposure to both structured and unstructured data
Data Engineer ยท Pontoonglobal