Likeremote

Subscribe to the latest remote jobs:

  • Likeremote jobs on https://LinkedIn.com/
  • Likeremote jobs on https://telegram.org/
  • Likeremote jobs on Reddit.com
DL

Pyspark Developer

Diverse Lynx India
๐Ÿ‡ฎ๐Ÿ‡ณ India
On-site
2 months ago
  • PySpark
  • Python
  • SQL
  • ETL
  • AWS
  • ELT
  • Data Modeling
  • Git
  • GitHub
  • AWS S3
  • EMR
  • Databricks
  • Airflow
  • Snowflake
  • dbt
  • Redshift
  • CI/CD
  • Agile
  • Scrum
Not scoredNo CV on file. Upload one and this job gets a score out of 100.Upload CV
Job Title: PySpark Developer
Experience:
5+ Years
Location:
Pune / Bangalore / Hyderabad / Chennai / Mumbai
Job Description:
We are looking for a skilledPySpark Developer with 5+ years of experience in designing, developing, and optimizing large-scale data processing applications. The ideal candidate should have strong hands-on expertise inPySpark, Python, SQL, and Data Engineering concepts, with experience building scalable ETL pipelines and processing large volumes of data. Relevant skills from enterprise JD templates include PySpark development, SQL, data warehousing, and cloud-based data engineering environments.[TCS_JD_Tem...Databricks | Word],[TCS_JD_Tem...Developer | Word],[TCS_JD_Tem...pr aws TRP | Word]
Must-Have Skills:
  • Strong hands-on experience inPySpark
  • Proficiency inPython Programming
  • Strong SQL skills
  • Experience in ETL/ELT development
  • Knowledge of Data Warehousing concepts
  • Experience with data modeling and large-scale data processing
  • Understanding of Spark architecture and performance tuning
  • Git/GitHub version control
  • Strong debugging and troubleshooting skills
Good-to-Have Skills:
  • AWS (S3, EMR, Glue)
  • Databricks
  • Airflow
  • Snowflake
  • dbt
  • Redshift
  • CI/CD pipelines
  • Agile/Scrum methodology
Roles & Responsibilities:
  • Develop and maintain scalable data pipelines using PySpark and Python.
  • Design and implement ETL/ELT workflows for enterprise data platforms.
  • Process and analyze large structured and unstructured datasets.
  • Optimize Spark jobs for performance, scalability, and efficiency.
  • Develop reusable frameworks and components for data processing.
  • Collaborate with data architects, business analysts, and application teams.
  • Perform data validation, reconciliation, and quality checks.
  • Troubleshoot production issues and provide timely resolutions.
  • Participate in code reviews and follow coding best practices.
Relevant Experience:
  • 5+ years of overall IT experience.
  • Minimum 3 years of hands-on PySpark development experience.
  • Experience in data engineering and data warehouse projects.
  • Exposure to cloud-based data platforms is preferred.[TCS_JD_Tem...Databricks | Word],[TCS_JD_Tem...Developer | Word]
Education:
  • BE / B.Tech / MCA / M.Tech or equivalent.

Pyspark Developer ยท Diverse Lynx India

Auto apply with Likeremote