Senior Data Engineer (Apache Flink)
Why This Role Stands Out
You'll drive innovation by designing and operating a high-volume, real-time event-processing pipeline using Apache Flink, offering significant impact and technical growth. This hybrid role is perfect for experienced data engineers with a strong background in Flink and Kafka who thrive on solving complex distributed systems challenges. Join Whiztek Corp to build cutting-edge technology and advance your career.
Quick Overview
Job Description
About the Role
We're looking for an experienced data engineer to design, build, and operate our real-time event-processing pipeline. You'll own a system that ingests high-volume event data (100M+ events/day) from multiple producers, processes it through Apache Flink for enrichment, deduplication, and aggregation, and delivers results to downstream services and APIs with low latency and strong correctness guarantees.
Responsibilities
- Design and maintain Kafka-based ingestion pipelines, including topic strategy, partitioning, and producer/consumer contracts
- Build and optimize Apache Flink streaming jobs (Java or Scala) for real-time transformation, deduplication, and event correlation
- Implement event-time processing using watermarks to correctly handle out-of-order and late-arriving events
- Ensure exactly-once processing semantics and manage Flink checkpointing/state backends for fault tolerance
- Diagnose and resolve production issues: consumer lag, checkpoint failures, back pressure, schema mismatches
- Tune Flink jobs for latency, throughput, and resource efficiency
- Collaborate on system architecture from ingestion through storage to API exposure
- Participate in on-call rotation and incident root-cause analysis
Requirements
- Strong experience with Apache Flink in production (Java or Scala)
- Solid understanding of Apache Kafka: producers, consumers, partitioning, offset management
- Experience with event-time semantics, watermarks, and stateful stream processing
- Familiarity with checkpointing, state backends (RocksDB, etc.), and recovery strategies
- Experience debugging distributed systems issues (back pressure, lag, failures) using logs/metrics
- Understanding of exactly-once vs at-least-once delivery guarantees
- Experience designing systems handling 100M+ daily events is a strong plus
- Comfortable with live coding/system design during interviews
Nice to Have
- Migration experience from Spark to Flink
- Experience exposing streaming results via REST APIs
- Familiarity with storage systems for high-throughput write patterns (e.g., Cassandra, Druid, ClickHouse)
Skills
Similar jobs
Senior Data Engineer
eTeam, Inc. · San Jose, United States
1 minute agoData Engineer
CVS Health · United States
1 minute ago$64.9k - $173.0k/yrData Engineer
Strategic Staffing Solutions · United States
4 minutes agoAssociate Data Engineer 2027 - AI & Data Analytics
IBM · San Francisco, United States
4 minutes agoAssociate Data Engineer 2027 - AI & Data Analytics
IBM · Chicago, United States
5 minutes agoData Engineer
Absolute Business Solutions Corp · Springfield, United States
11 minutes ago$150k - $185k/yr