AI Safety Specialist (AI Engineering)
Hyphen Connect Limited
Apply to this job San Francisco Bay Area, US Until 8/21/2026 First posted April 25, 2026 Last posted April 25, 2026
Job description
We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will ensure the safe deployment of AI systems by conducting adversarial testing, implementing protective measures, and aligning AI behavior with ethical principles.
Responsibilities:
- Conduct adversarial testing on LLMs and multimodal agents.
- Implement guardrails and real-time filtering for autonomous tool use.
- Develop constitutional AI principles and assist with RLHF alignment pipelines.
Qualifications:
- Background in cybersecurity, prompt engineering, or adversarial ML.
- Experience with jailbreak taxonomies and automated red-teaming frameworks.
- Strong analytical mindset for identifying edge cases.
About this role
Summary
Enhance AI security and robustness through testing, guardrails, and ethical alignment
Job title
AI Safety Specialist (AI Engineering)
Experience level
Industry
software
Location requirements
San Francisco Bay Area, USA; remote work not specified
Salary
Not specified
Management role
No
Skills & keywords
Required skills
cybersecurityprompt engineeringadversarial MLjailbreak taxonomiesred-teaming
Preferred skills
None specified
Specializations
AI safetyadversarial testingRLHFsecurity
Locations
Structured locations inferred from the posting.
Unknown location
Work arrangement unknown
Related searches