Quick Overview
Job Description
Job Title: Sr AI Data Engineer
Location: Remote
Duration: 6 Months + Extension
Bill Rate: $89/hour
Job Type: W-2 Contract
Client: To Be Discussed Later
Work Authorization: US-Citizen, H-1B, OPT-EAD, GC-EAD
Job Summary:
Roles and Responsibilities:
Knowledge Graph and Semantic Layer (primary focus)
- Lead the design and evolution of the knowledge graphs and ontologies powering our AI's reasoning, retrieval, and explainability.
- Align enterprise data (engineering handbooks, parts, service manuals, DMAIC records, user files) into a coherent, queryable graph with clear provenance across structured, semi-structured, and unstructured sources.
- Own the retrieval substrate - graph queries, vector indexes, and hybrid retrieval - and drive measurable improvements in grounding quality.
AI/ML Data Quality
- Curate grounding corpora, eval datasets, and retrieval benchmarks for LLM-based features.
- Instrument metrics for retrieval quality, grounding accuracy, and freshness; drive regressions down over time.
- Shape training and inference data contracts with AI engineers, including feedback loops from user signals.
Data Modeling and Pipelines on AWS
- Produce conceptual, logical, and physical data models for operational and analytical workloads; establish modeling standards, naming conventions, and reuse patterns.
- Build ingestion and transformation pipelines in Python and SQL using AWS services - Glue, Lambda, Step Functions, S3, Athena, OpenSearch, Neptune - and AI services such as Bedrock and Bedrock Knowledge Bases.
- Author infrastructure as code in CloudFormation (CDK welcome) and apply AWS best practices for IAM, security, cost, and observability.
- Profile sources, identify data quality gaps, and design automated validation, monitoring, metadata, and lineage.
Data Governance and Identity Integration
- Partner with security and platform teams to integrate data access with enterprise identity and access policies, as we look to modernize for AI.
- Define data contracts, attributes, and metadata that policy engines can reason over for attribute- and context-based access control.
- Contribute to the technical data dictionary, business glossary, and data catalog.
Technical Leadership
- Set the design direction for data and semantic modeling across the team.
- Mentor engineers and citizen developers on modeling, ontology design, and retrieval engineering.
- Communicate tradeoffs and value clearly to product, business, and executive stakeholders
Required Qualifications:
- Bachelor's degree in Computer Science, Engineering, or a STEM field
- A minimum of three years of data engineering experience
Eligibility Requirement:
- Legal authorization to work in the U.S. is required. We will not sponsor individuals for employment visas, now or in the future, for this job.
Desired Qualifications:
- 5+ years of hands-on data engineering with a track record of designing - not just implementing - data models and semantic layers.
- Production experience with knowledge graphs and ontologies (Neo4j, Neptune, TigerGraph, RDF/SPARQL, or similar) and graph query languages (Cypher, Gremlin, SPARQL).
- Strong AWS proficiency required: CloudFormation (or CDK), Glue, Lambda, Step Functions, S3, IAM, Bedrock, Bedrock Knowledge Bases; OpenSearch and Neptune a plus.
- Strong Python and SQL; comfort across relational, graph, vector, and document stores.
- Experience supporting AI/ML or LLM systems - RAG pipelines, embeddings, eval datasets, grounding corpora.
- Experience integrating data access with enterprise identity and policy systems
- Strong cross-functional collaboration and communication, including technical presentations to non-data audiences.
Leadership Skills:
- Ability to work effectively with multi-disciplinary teams (e.g., Digital Technology, GE Business teams) and understand the inter-dependencies between them.
- Ability to showcase teamwork skills to achieve common goals, provide resolutions and share ideas.
- Demonstrate the presentation and influencing skills
Equal Opportunity Employer: We are an equal opportunity employer. All aspects of employment including the decision to hire, promote, discipline, or discharge, will be based on merit, competence, performance, and business needs. We do not discriminate on the basis of race, color, religion, marital status, age, national origin, ancestry, physical or mental disability, medical condition, pregnancy, genetic information, gender, sexual orientation, gender identity or expression, national origin, citizenship/ immigration status, veteran status, or any other status protected under federal, state, or local law.
Similar jobs
- UL
Senior Data Engineer
NewUline, Inc.
Waukegan, IL🇺🇸$96k - $148k/yrOn-site14 hours agoSQLT-SQLETL+2Technology - DO
Director, Data Platform Engineering
Domino's
Ann Arbor, MI🇺🇸On-site3 days agoMachine LearningGenerative AIStakeholder ManagementTechnology - MI
Senior Python Data Scraping Engineer (Freelance)
NewMindrift
Austin, Texas🇺🇸$45/hrRemote2 days agoDockerAWSSelenium+4Engineering - MI
Senior Python Data Scraping Engineer (Freelance)
NewMindrift
San Antonio, Texas🇺🇸Remote2 days agoDockerAWSSelenium+4Engineering - MI
Senior Python Data Scraping Engineer (Freelance)
NewMindrift
Houston, Texas🇺🇸Remote2 days agoDockerAWSSelenium+4Engineering - MI
Senior Python Data Scraping Engineer (Freelance)
NewMindrift
Dallas, Texas🇺🇸Remote2 days agoDockerAWSSelenium+4Engineering