Quick Overview
Seniority
Mid Senior
Location
San Francisco Bay Area, United States
Posted
4 months ago
Job Description
We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will ensure the safe deployment of AI systems by conducting adversarial testing, implementing protective measures, and aligning AI behavior with ethical principles.
Responsibilities:
- Conduct adversarial testing on LLMs and multimodal agents.
- Implement guardrails and real-time filtering for autonomous tool use.
- Develop constitutional AI principles and assist with RLHF alignment pipelines.
Qualifications:
- Background in cybersecurity, prompt engineering, or adversarial ML.
- Experience with jailbreak taxonomies and automated red-teaming frameworks.
- Strong analytical mindset for identifying edge cases.
Similar jobs
- TC
Ai Engineer
NewTEKsystems c/o Allegis Group
Palo Alto, CA🇺🇸$80 - $85/hrHybrid13 hours agoMachine LearningPhoenixPython+1Technology - FA
AI Engineer
NewFactspan Inc
United States🇺🇸Hybrid13 hours agoMicroservicesSQLAngular+7Technology - FB
M365 Systems Engineer III - AI and Collaboration (Remote)
First-Citizens Bank & Trust Company
Raleigh, NC🇺🇸Remote4 weeks agoActive DirectoryAgileAzure+3Technology - CT
AI / LLM Engineer
NewConquest Tech Solutions Inc
United States🇺🇸Remote13 hours agoDockerSQLAWS+15 - SG
Interim Head of Search & AI Discoverability (Global)
Software Guidance & Assistance
New York, NY🇺🇸Hybrid4 weeks agoSEOGenerative AILLM - FB
M365 Systems Engineer III - AI and Collaboration (Remote)
First-Citizens Bank & Trust Company
TX🇺🇸Remote7 weeks agoActive DirectoryAgileAzure+3Technology