Software Engineer, API Multimodal is a AI Engineer role (full-time). with OpenAI. in SAN FRANCISCO, US. Compensation shown: $293K–$385K. Imported listing (source: zerogtalent.com). Apply on the employer's site (zerogtalent.com).
Imported listing (source: zerogtalent.com) · Apply on zerogtalent.com
The sections below reproduce the third-party job description for reference. AIEngineer.careers does not write or control this text.
Imported job description
Sourced from zerogtalent.com
OpenAI is hiring a software engineer for its API Multimodal team to build and operate the developer-facing products and infrastructure behind its image, audio, and real-time APIs. The role involves designing low-latency streaming and model integration systems, partnering with research and safety teams, and shipping frontier multimodal capabilities to developers. This hands-on backend engineering role combines systems depth with product judgment.
Responsibilities
Design, build, and ship developer-facing APIs and backend services that serve frontier models.
Architect low-latency streaming, request, session, and model integration systems that make complex multimodal interactions reliable and intuitive at scale.
Work directly with Research to bring new model capabilities into production, shape the systems around them, and incorporate feedback from real-world developers and customers.
Own the availability, latency, scalability, and cost efficiency of the services you build.
Own projects from technical design and implementation through launch and ongoing iteration, while raising the team's engineering standards.
Requirements
7+ years of professional experience, excluding internships, in backend, infrastructure, platform, or product engineering roles.
OpenAI is an AI research and deployment company that develops large-scale artificial intelligence models and safety systems. It focuses on advancing artificial general intelligence through research, product development, and collaborative partnerships.
A track record of designing, building, and operating production backend services, developer-facing APIs, or distributed systems.
Strong software engineering and systems fundamentals, with experience leading technically complex projects from ambiguous ideas to production.
Proficiency in one or more general-purpose backend languages, such as Python, Go, Rust, or TypeScript.
Experience building reliable, scalable systems and reasoning about distributed architecture, concurrency, latency, observability, and operational tradeoffs.
Product judgment and developer empathy, including an ability to turn complex model or infrastructure capabilities into clear, intuitive APIs.
Clear communication and a collaborative approach to working with researchers, product managers, designers, infrastructure engineers, and customers.
A strong sense of ownership, comfort with ambiguity, and a bias toward learning directly from users while continuously improving engineering quality.
Nice to Have
Experience with real-time streaming, audio processing, speech systems, image generation, computer vision, or multimodal AI applications is helpful, but not required.