Why This Role Stands Out
This hybrid AI Engineer role offers a dynamic opportunity to leverage cutting-edge AI tools and cloud technologies, fostering significant career growth and skill development within a reputable company. You'll thrive here if you're a proactive problem-solver with a strong background in Python, infrastructure, and a passion for AI-assisted development, so be sure to apply!
Quick Overview
Seniority
Mid Senior
Work mode
Hybrid
Location
Austin, TX, United States
Posted
4 days ago
MLOpsDatadogGitHub ActionsGitLab CIGrafanaJenkinsLLMPhoenixPrometheusPythonTypeScript
Job Description
- 5+ years of experience in ML engineering, MLOps, platform engineering, or SRE, including 2+ years working hands-on with LLMs or LLM-powered applications in production.
- Demonstrated experience building evaluation systems for ML or LLM applications: test harnesses, benchmark datasets, automated scoring (including LLM-as-judge approaches), and regression detection.
- Strong software engineering skills in Python (and ideally TypeScript), with a track record of building reliable, well-tested internal platforms and tooling.
- Deep familiarity with CI/CD systems (e.g., GitHub Actions, GitLab CI, Jenkins, Buildkite) and experience embedding automated quality gates into deployment pipelines.
- Experience with observability and monitoring stacks (e.g., OpenTelemetry, Datadog, Grafana/Prometheus) and, ideally, LLM-specific observability tools (e.g., LangSmith, Langfuse, Arize Phoenix, Braintrust, W&B Weave).
- Proven ability to debug complex distributed systems under pressure, including production incident response, root-cause analysis, and blameless postmortems.
- Excellent cross-functional communication: able to translate evaluation results into clear findings and recommendations for both engineers and non-technical stakeholders.
- Comfort with ambiguity and a builder’s mindset: this role starts with a blank page and ends with the evaluation platform the whole organization relies on.
- Experience with agentic frameworks and orchestration patterns (e.g., multi-agent systems, tool use, RAG pipelines) and their distinct failure modes.
- Experience with prompt management, model routing, or fine-tuning workflows and evaluating changes across model versions and providers.
- Background in statistics or experimentation (A/B testing, significance testing, sampling strategies for human review).
- Design and build reusable AI agent skills, plugins and maintain internal marketplace infrastructure to extend and scale Data, AIML capabilities across the organization.
- Expertise in causal inference and measurement strategy including causal graphs, ontologies, and knowledge graphs to drive rigorous, decision grade data analysis.
- Experience operating in regulated or high-stakes domains where agent errors carry real business or customer impact.
Similar jobs
- AT
Senior Virtual Reality / AI Engineer
NewAneka Talent Solutions
United States🇺🇸Hybrid19 hours agoC#Computer VisionGenerative AI+2Technology - ST
AI Engineer
NewStefanini
Dallas, TX🇺🇸On-site19 hours agoTechnology - DE
AI Architect
NewDTEL Engineering & Consultants Inc
Miami, FL🇺🇸Hybrid19 hours agoDockerNLPAzure+2Technology - BL
Senior AI Engineer
Blend360
Columbia, MD🇺🇸On-site2 months agoTechnology - KI
AI Engineer/ Architect
NewKey Infotek LLC
Jersey City, NJ🇺🇸Hybrid19 hours agoDockerFastAPIFlask+11Technology - TE
AI Scientist
NewTechridge, Inc.
Dallas, TX🇺🇸Hybrid19 hours agoDeep LearningPython