Search
Go
Researcher, Alignment Interpretability is a AI Research Engineer role (full-time). with OpenAI. in SAN FRANCISCO, US. Compensation shown: $295K–$500K. Imported listing (source: zerogtalent.com). Apply on the employer's site (zerogtalent.com).
Imported listing (source: zerogtalent.com) · Apply on zerogtalent.com
The sections below reproduce the third-party job description for reference. AIEngineer.careers does not write or control this text.
Sourced from zerogtalent.com
OpenAI is seeking a Researcher focused on Alignment Interpretability to study internal representations of deep learning models and engineer safer, more understandable AI systems. You will develop and publish research on mechanistic interpretability, build scalable infrastructure for analyzing model internals, and collaborate across teams to advance long-term AI safety goals.

San Francisco · US · 7850+ employees
OpenAI is an AI research and deployment company that develops large-scale artificial intelligence models and safety systems. It focuses on advancing artificial general intelligence through research, product development, and collaborative partnerships.
Together AI