Lead Databricks Engineer with Python (Need Only local to MN)
Why This Role Stands Out
You can lead a critical healthcare analytics pipeline modernization, migrating legacy SAS workflows to Python/PySpark on Databricks for a reputable company. This hybrid role offers a fantastic opportunity for a mid-senior engineer with strong Python and Databricks skills to drive significant impact and grow their expertise in a collaborative environment. Explore this contract-to-hire position to leverage your skills and advance your career.
Quick Overview
Job Description
Title: Senior / Lead Data Engineer
Location: Eagan, MN (Hybrid – 2 days/week onsite)
Duration: 3–6 Months Contract-to-Hire
Start Date: ASAP
Job Summary
We are seeking a Senior/Lead Data Engineer to modernize and own enterprise data pipelines by migrating legacy SAS-based workflows to Python/PySpark on Databricks. The engineer will partner with SMEs during the transition period and eventually take ownership of a critical healthcare analytics pipeline supporting HEDIS reporting.
Key Responsibilities
- Migrate legacy SAS pipelines to Python/PySpark on Databricks.
- Design, build, and maintain scalable ETL/ELT pipelines.
- Develop distributed data processing solutions using Databricks.
- Create and optimize complex SQL queries.
- Schedule, automate, and monitor data pipelines.
- Work with AWS services including S3, Lambda, Glue, and EC2.
- Manage Databricks notebooks, workflows, and clusters.
- Collaborate with business stakeholders and SMEs to transition pipeline ownership.
- Follow Agile methodologies and Git-based development practices.
Required Skills
- 8+ years of Data Engineering experience.
- Strong Python programming and scripting skills.
- Hands-on PySpark development.
- Extensive Databricks experience (Notebooks, Workflows, Cluster Management).
- Experience processing large datasets in distributed environments.
- Strong SQL skills.
- AWS experience with:
- S3
- Glue
- Lambda
- EC2
- Experience building and automating ETL/ELT pipelines.
- Git/version control experience.
- Agile/Scrum experience.
Preferred Skills
- Experience with SAS-to-Python migration.
- AI/Automation experience.
- Healthcare or HEDIS domain experience.
- Lead/Principal Data Engineering experience.
Must-Have Skills
- Python
- PySpark
- Databricks
- SQL
- AWS (S3, Glue, Lambda, EC2)
- ETL/ELT Pipeline Development
- Distributed Data Processing
- Git
- Agile/Scrum
Nice-to-Have Skills
- SAS Modernization
- AI/Generative AI
- Healthcare/HEDIS
- Airflow or other workflow orchestration tools
- Leadership/Pipeline Ownership experience
Skills
Similar jobs
Sr Lead AI Data Engineer - 100% onsite work and onsite interview
Unisoft Technology Inc · Woodlawn, United States
5 minutes agoSenior Principal AI Data Engineer
General Dynamics · Arlington, United States
7 minutes ago$174.3k - $235.8k/yrLead Software Engineer - Data Engineer and Applied AI
JPMorgan Chase & Co. · Jersey City, United States
7 minutes agoSenior Data Engineer with Iceberg
TekLeaders, Inc · Boston, United States
8 minutes agoData Engineer - Sr. Consultant level
Visa Inc. · Bellevue, United States
8 minutes ago$152.2k - $243.7k/yrAssociate Data Engineer - Data Services - 2027
IBM · Baton Rouge, United States
9 minutes ago