Search
Go
This job is no longer accepting applications
This listing was closed on Sep 18, 2026.
Browse Open JobsSr. Software Development Engineer, Inference Team - AWS Neuron - Seattle is a AI Infrastructure Engineer role (full-time). with Annapurna Labs (U.S.) Inc. in SEATTLE, US. Compensation shown: $168K–$227K. Imported listing (source: educativ.net). Apply on the employer's site (educativ.net).
Imported listing (source: educativ.net) · Apply on educativ.net
The sections below reproduce the third-party job description for reference. AIEngineer.careers does not write or control this text.
Sourced from educativ.net
AWS Neuron is the software stack for AWS Inferentia and Trainium accelerators for cloud-scale ML. As a senior engineer on the Machine Learning Inference Applications team, you will lead development and optimization of open-source inference frameworks like vLLM and SGLang to deliver high-performance LLM serving on AWS Neuron. You will collaborate with model, compiler, runtime, and performance engineers to ensure end-to-end performance, scalability, and production readiness.
Cupertino · US · 234+ employees
Annapurna Labs is a fabless semiconductor subsidiary of Amazon Web Services (AWS) that designs custom silicon and software infrastructure for cloud computing. Its innovations include the AWS Nitro System, Graviton processors, and machine learning accelerators like Trainium and Inferentia.
TikTok