Skill is seeking an Applied AI Safety Engineer to design and execute adversarial evaluations for conversational and tool-using AI systems. The role involves developing threat models, creating Python evaluation pipelines, and validating LLM-based judges to ensure safety and reliability. You will collaborate with engineering and trust & safety teams to implement mitigations and integrate safety checks into the product lifecycle.
Search
Go
