PySpark lead Developer - Irving, TX
Quick Overview
Job Description
Job Title: PySpark developer
Location: Irving, TX- Hybrid
Role Overview
Experience with big data processing and distributed computing systems like Spark.
Implement ETL pipelines and data transformation processes.
Ensure data quality and integrity in all data processing workflows.
Troubleshoot and resolve issues related to PySpark applications and workflows.
Understand source, dependencies and data flow from converted PySpark code.
Strong programming skills in Python and SQL.
Experience with big data technologies like Hadoop, Hive, and Kafka.
Understanding of data warehousing concepts and relational databases like SQL.
Demonstrate and document code lineage.
Integrate PySpark code with frameworks such as Ingestion Framework, DataLens, etc.,
Ensure compliance with data security, privacy regulations, and organizational standards.
Knowledge of CI/CD pipelines and DevOps practices.
Strong problem-solving and analytical skills.
Excellent communication and leadership abilities.
Qualifications:
6+ years of experience in big data development, Hadoop , Hive & Spark framework.
Good to have experience in SAS.
Strong Python, PySpark Development and SQL knowledge.
Certification in big data or cloud technologies is preferred.
Skills
Similar jobs
.NET Developer - HBITS-08-14939
GreyCell Labs, Inc · Albany, United States
11 minutes agoSenior AI Developer (Full Stack) - Hybrid
Genesis10 · Charlotte, United States
11 minutes ago$60 - $68/hr.Net Developer with AWS
SRI Tech Solutions · Jersey City, United States
11 minutes agoTop Secret Power Platform Developer with Security Clearance
Zachary Piper Solutions, LLC · Washington, United States
13 minutes ago$120k - $165k/yrSharePoint Lead Developer
Swanktek Inc · United States
14 minutes agoFull-Stack Software Engineer
Booz Allen Hamilton · Chantilly, United States
16 minutes ago$112.8k - $257k/yr