KH
Data Engineer (Data Pipelines & Modeling)
Katalyst Healthcares & Life Sciences
๐บ๐ธ United States
On-site
5 months ago
- Python
- SQL
- Apache Spark
- ETL
- ELT
- CI/CD
- Jenkins
- GitHub Actions
- AWS
- Azure
- GCP
5 months ago
- Design and implement robust data ingestion pipelines from multiple sources (APIs, databases, files, streaming systems).
- Support C4C offline database migration, ensuring data accuracy and consistency.
- Integrate data from enterprise systems into centralized data platforms.
- Design and implement data models for Workforce planning.
- Service operations forecasting.
- Develop optimized schemas for reporting and analytics.
- Ensure data quality, integrity, and consistency across models.
- Strong experience in data engineering and pipeline development.
- Proficiency in Python / SQL.
- Hands-on experience with Apache Spark or similar big data tools.
- Strong understanding of ETL/ELT concepts and data warehousing.
- Ability to work independently and in cross-functional teams.
- Bachelor's / Master's in Computer Science, IT, or related field.
- Exposure to CI/CD tools like Jenkins or GitHub Actions.
- Knowledge of cloud platforms (AWS / Azure / GCP).
- Experience in healthcare or regulated environments.
Data Engineer (Data Pipelines & Modeling) ยท Katalyst Healthcares & Life Sciences