Senior Data Engineer with Pyspark
Why This Role Stands Out
This hybrid Senior Data Engineer role offers a fantastic opportunity to build and optimize scalable data pipelines within the healthcare sector, leveraging your PySpark and SSIS expertise. You'll thrive here if you enjoy both hands-on development and operational oversight, contributing to critical data solutions in a collaborative environment. Apply now to grow your skills and make a significant impact!
Quick Overview
Job Description
Job Title: Senior Data Engineer (PySpark, ETL SSIS) -W2 only
Location: Rocky Hill,CT
Job Description:
We are looking for an experienced and motivated Data Engineer with expertise in PySpark to join our dynamic team. As a key member of our data engineering team, you will play a crucial role in designing, building, and maintaining scalable data pipelines that enable efficient data processing and analytics within the healthcare domain.
This role will combine both development and administrative activities, making it essential that the candidate has experience not only in building robust data pipelines but also in overseeing their operational aspects to ensure performance, reliability, and optimization.
Key Responsibilities:
- Design, develop, and maintain scalable ETL pipelines using PySpark to process large datasets.
- Collaborate with cross-functional teams (data scientists, analysts, business stakeholders) to understand data requirements and deliver high-quality solutions.
- Work on administrative tasks, including monitoring, troubleshooting, and optimizing data pipelines and infrastructure.
- Manage data integration across healthcare systems, ensuring compliance with relevant standards.
- Leverage SSIS for ETL development and ensure smooth data movement across different environments.
- Integrate and transform data from multiple sources, ensuring data quality and consistency.
- Handle and resolve data processing issues, ensuring minimal disruption to operations.
- Document best practices, processes, and workflows to maintain pipeline efficiency and scalability.
- Work with both relational and non-relational databases, ensuring smooth data flow and optimized performance.
Required Skills & Qualifications:
- Strong experience in Data Engineering, with expertise in designing, building, and maintaining ETL pipelines.
- Strong proficiency in PySpark for large-scale data processing and transformation.
- Experience with ETL tools, particularly SSIS (SQL Server Integration Services).
- Solid understanding of data modeling, relational databases, and data warehousing principles.
- Experience working with cloud-based data storage and processing technologies (AWS, GCP, or Azure).
- Familiarity with healthcare data standards, such as HL7 and FHIR, is highly desirable.
- Proven ability in data pipeline monitoring, troubleshooting, and performance tuning.
- Strong communication skills and the ability to work collaboratively with cross-functional teams.
Infowave Systems is an equal opportunity employer that is committed to diversity and inclusion in the workplace.
Skills
Similar jobs
Data Engineer - contract to hire - MUST HAVE HEALTH PAYER / HEALTH PLAN EXPERIENCE
Svam International, Inc. · United States
2 minutes agoSenior Data Engineer
eTeam, Inc. · San Jose, United States
20 minutes agoData Engineer
CVS Health · United States
20 minutes ago$64.9k - $173.0k/yrData Engineer
Strategic Staffing Solutions · United States
23 minutes agoAssociate Data Engineer 2027 - AI & Data Analytics
IBM · San Francisco, United States
23 minutes agoAssociate Data Engineer 2027 - AI & Data Analytics
IBM · Chicago, United States
24 minutes ago