Likeremote

Subscribe to the latest remote jobs:

  • Likeremote jobs on https://LinkedIn.com/
  • Likeremote jobs on https://telegram.org/
  • Likeremote jobs on Reddit.com
DL

Python Pyspark

Diverse Lynx India
Location not stated
2 months ago
  • Python
  • SQL
  • PySpark
  • Pandas
  • NumPy
  • Java
  • Linux
  • AWS
  • EC2
  • Snowflake
  • dbt
  • Talend
  • Angular
  • ETL
  • pytest
  • GitHub
  • Jenkins
  • CI/CD
  • Excel
Not scoredNo CV on file. Upload one and this job gets a score out of 100.Upload CV
Skill: Python
Location: Chennai
Experience: 6+
Desired Competencies (Technical/Behavioral Competency)
Must-Have  Ability to analyze, design, build, and deploy data pipelines in different environments.
 Analytical ability to solve medium to complex problems
 Writing complex SQL scripts across extensive data volumes
 Ability to transform SQL script to PySpark/Snowpark code
 Hands on the use of data science libraries like pandas and NumPy.
 Design and develop code considering performance
 Ability to analyze data from different sources and extract it at regular intervals considering performance
Must have worked as an individual contributor and interacted with different teams in the micro architecture application to get the work done.
 Ability to understand and adapt to existing applications designed on Java/SQL and make code changes.
 Must have hands-on experience with Linux commands and scripts.
Good-to-Have  AWS development (S3, EC2) exposure
Snowflake experience
 Hands-on with DBT
 Hands-on with Autosys
Talend exposure
 Orchestration exposure
 Knowledge of CDC
 Hands-on with Java
 Hands-on with angular
 Understanding of API design

SN Responsibility of / Expectations from the Role
1 Hands-on experience designing, building, deploying, testing, maintaining, monitoring, and owning scalable, resilient, and distributed data pipelines/ETL processes.
2 Experienced in Testing and debugging with Python test framework tools like Pytest, unit test. Should be able to mock and write unit tests for pyspark/pandas/Snowpark code.
3 Proficiency in SQL and Python for applied large-scale data processing
4 Expertise with RDBMS and OLAP databases, and big data technologies-Snowflakes, including Datawarehouse and Data Lake.
5 Have used GitHub, Jenkins, Source Tree, VSS Code, CI/CD pipeline
6 Ability to excel in a short timeframe under short sprints
7 Strong problem solving, analytical, communication and documentation skills.

Python Pyspark · Diverse Lynx India

Auto apply with Likeremote