Haystack
← Back to Jobs
Technology
BD

Principal Data Engineer

Black Duck Software, Inc.Belfast🇬🇧United KingdomPosted 11 Sept 2026

Why This Role Stands Out

As a Principal Data Engineer at Black Duck Software, you will lead the design and build-out of innovative cross-product data services, shaping the future of how data is managed and utilized across multiple product lines. This leadership role is ideal for experienced data professionals who thrive on defining canonical data models, operationalizing data pipelines, and ensuring data reliability and accessibility for a variety of applications, including ML workflows and AI automation. If you're passionate about driving data strategy and impact within a recognized leader in application security, this is an exceptional opportunity to advance your career.

Quick Overview

Seniority
Leader
Location
Belfast, United Kingdom
Posted
5 days ago
SQLAWSGoogle CloudPython

Job Description

Black Duck Software, Inc. helps organizations build secure, high-quality software, minimizing risks while maximizing speed and productivity. Black Duck, a recognized pioneer in application security, provides SAST, SCA, and DAST solutions that enable teams to quickly find and fix vulnerabilities and defects in proprietary code, open source components, and application behavior. With a combination of industry-leading tools, services, and expertise, only Black Duck helps organizations maximize security and quality in DevSecOps and throughout the software development life cycle.

What you’ll do

  • Lead the design and build-out of cross-product data services for multiple product lines from one governed data plane.
  • Define the “customer data plane” model: canonical customer identifiers, shared dimensions, and consistent facts used across products.
  • Build and operationalize ingestion patterns for batch, streaming, and event data, with repeatable onboarding for new sources.
  • Own the operational playbook for data reliability: data contracts, quality checks, lineage, monitoring, and incident response.
  • Implement and run access methods that make data usable: curated datasets, secure query interfaces, and product-ready data APIs where needed.
  • Productize customer-facing data products (datasets, metrics, exports, and feeds) with versioning, documentation, and clear ownership.
  • Design data models that fit both operational systems (RDS) and analytics stores (columnar/OLAP), including performance and cost tuning.
  • Ensure data products also power ML workflows: trusted training datasets, feature-ready outputs, and consistent definitions for decision-making.
  • Enable AI automation by delivering reliable, low-latency, governed data products that can be used safely in automated workflows.
  • Partner closely with product, engineering, and security stakeholders to align data products to roadmap priorities and customer outcomes.
  • Raise the technical bar through architecture reviews, standards, and mentoring—while staying hands-on in key systems.

Required

  • Significant experience building and operating production data platforms at scale, including on-call and operational ownership.
  • Strong SQL skills and strong Python skills, used to build pipelines, services, and automation.
  • Hands-on experience running cloud systems on AWS and Google Cloud (IaaS level: compute, storage, networking, IAM).
  • Practical experience with both operational databases (RDS-style) and analytics stores (columnar/OLAP), including performance tuning.
  • Strong data modeling ability, including schema evolution, conformed dimensions, and “one source of truth” metric definitions.
  • Track record of delivering data products that other teams or customers depend on, with clear contracts and reliability expectations.
  • Ability to make sound engineering tradeoffs across latency, accuracy, cost, and security without creating brittle complexity.
  • Experience with lakehouse patterns and open table formats (or similar), including governance and table maintenance.
  • Experience with orchestration and streaming systems used in production (batch + real-time), and managing backfills safely.
  • Familiarity with ML data needs (training/serving splits, feature-ready datasets, evaluation datasets) and AI-adjacent workflows.

Preferred

  • Experience building self-service data platforms (catalog, discoverability, access controls) used by multiple teams.
  • Experience in regulated or security-sensitive environments, including retention, auditing, and data access controls.

Work model, location & travel

  • Location: Belfast, UK
  • Reports to: VP of Data Engineering
  • Work model: Hybrid (details TBD)
  • Collaboration hours: Flexible; overlap with UK and US time zones
  • Travel: Minimal

 

Black Duck is an equal opportunity employer. We consider all applicants for employment without regard to race, color, national origin, religion, sex, gender identity or expression, age, disability, sexual orientation, veteran or military service status, or any other characteristic protected by applicable law. Black Duck complies with all applicable laws prohibiting employment discrimination in every jurisdiction where it operates and provides reasonable accommodations to individuals with disabilities in accordance with applicable law.

Similar jobs