Search
Go
This job is no longer accepting applications
This listing was closed on Jul 19, 2026.
Browse Open JobsSenior ML Engineer (Inference Serving) is a AI Infrastructure Engineer role (full-time). with OHO US. in SAN FRANCISCO, US. Compensation shown: $200K–$275K. Imported listing (source: oho.us). Apply on the employer's site (oho.us).
Imported listing (source: oho.us) · Apply on oho.us
The sections below reproduce the third-party job description for reference. AIEngineer.careers does not write or control this text.
Sourced from oho.us
Join a stealth AI hardware startup building a custom AI SoC and inference serving stack. As a Senior ML Engineer, you will architect high-performance multi-node inference stacks and implement optimizations in frameworks like vLLM and PyTorch. You will work on advanced cluster scheduling and contribute to open-source AI infrastructure.
Boston · US · 65+ employees
There are two distinct companies known as OHO: OHO US is a specialized recruitment firm for tech and engineering roles, while OHO is a digital marketing and strategy agency with over 25 years of experience.
TikTok