Haystack
← Back to Jobs
Technology

DataBrick Data Engineer

HR PunditsUnited States🇺🇸United StatesPosted 30 Jul 2026

Quick Overview

Work Type
Hybrid
Level
Mid Senior

Job Description

Job Title: DataBricks Data Engineer

Remote

Full Time Only

Job Summary

The Databricks Data Engineer owns the end-to-end migration strategy, target architecture design, and technical execution of moving legacy ETL workloads to the Databricks Lakehouse. He will establish migration standards, optimize PySpark pipelines, orchestrate complex data workflows, and deploy proprietary automation tools to ensure a seamless, high-performing transition from legacy systems.

Key Responsibilities

Architectural Strategy & Governance

  • Design the target Databricks Lakehouse architecture utilizing Delta Lake, Photon, and Unity Catalog.
  • Establish global code refactoring standards, optimization benchmarks, and PySpark best practices.
  • Resolve highly complex dependency mappings and architect seamless, zero-downtime dual-run strategies.
  • Lead the technical deployment and integration of specialized migration accelerators

Hands-on Engineering & Optimization

  • Review automated output from migration tools and manually refactor complex legacy logic into high-performing PySpark notebooks.
  • Eliminate legacy anti-patterns such as massive row-by-row processing and inefficient lookups.
  • Optimize PySpark code performance using advanced Spark features including Z-Ordering, partitioning, and caching.
  • Build robust Databricks Workflows and orchestrate complex DAGs based on comprehensive source lineage.

Technical Skills & Competencies

  • Core Platforms: Databricks, Delta Lake, Unity Catalog, Photon, DataStage.
  • Languages & Frameworks: PySpark, Python, SQL, Shell Scripting.
  • Cloud & DevOps: AWS alongside CI/CD deployment pipelines.
  • Orchestration: Apache Airflow, Databricks Workflows.

Skills

SQL
Shell
AWS
ETL
Airflow
Apache
Databricks
Python
Unity

Similar jobs