AI Safety & Evaluation Engineer
Why This Role Stands Out
This hybrid role offers a unique opportunity to be at the forefront of responsible AI development, shaping the safety and trustworthiness of cutting-edge models. You'll thrive here if you possess a strong background in AI/ML or security and a passion for building robust evaluation frameworks and guardrails. Apply today to make a significant impact on the future of AI.
Quick Overview
Job Description
AI Safety & Evaluation Engineer
Summary Own responsible AI deployment — red-teaming, guardrails, hallucination detection, and evaluation. You''re the last line of defense ensuring our agents and models are safe, accurate, and trustworthy in production.
Role Expectations
- Design and run red-teaming and adversarial testing against models and agents (jailbreaks, prompt injection, misuse).
- Build guardrails: input/output filtering, policy enforcement, content moderation, and safe fallback behavior.
- Develop hallucination detection and grounding checks; measure and reduce factual error rates.
- Design evaluation frameworks and benchmarks for accuracy, safety, robustness, and bias — both offline and in production.
- Define and enforce responsible AI standards, documentation, and deployment gates.
- Partner with agent and ML teams to close safety gaps before and after release.
Required Skills
- 6+ years in ML/AI, security, or evaluation-focused engineering.
- Strong Python and deep familiarity with LLM behavior, failure modes, and prompt engineering.
- Hands-on experience building evaluations, benchmarks, or guardrail/moderation systems.
- Understanding of red-teaming, adversarial attacks, and hallucination mitigation techniques.
- Rigorous, data-driven approach to measuring model quality and risk.
Value adds:
- Experience with eval tooling and safety frameworks.
- Background in AI security, alignment research, or trust & safety.
- Familiarity with AI governance, regulatory, or compliance requirements.
Core Stack:
Python · Evaluation Frameworks · Guardrails & Moderation · Red- Teaming · Prompt Engineering · Hallucination Detection
Skills
Similar jobs
Technical Architect
Della Infotech · Seattle, United States
1 minute agoEden Maintenance Tech
Brooksource · United States
1 minute ago$17 - $21/hrHealth and Safety Coordinator - Onsite
Genesis10 · Edison, United States
1 minute ago$44 - $54/hrMechatronics & Robotics Technician (MRT) - Kansas City, KS
Genesis10 · Kansas City, United States
1 minute ago$29/hrAI Agentic Engineer - Hybrid
Genesis10 · Columbus, United States
1 minute ago$58 - $68/hrAccessibility QA Tester
Creative IT Inc · United States
17 minutes ago