Search
Go
This job is no longer accepting applications
This listing was closed on Jul 18, 2026.
Browse Open JobsHippocratic AI is seeking an LLM Inference Engineer to optimize large language model serving infrastructure. The role involves designing multi-node serving architectures, applying quantization techniques, and implementing speculative decoding. Candidates need expertise in Python, C++, CUDA, and GPU optimization.

Palo Alto · US · 200+ employees
Hippocratic AI is a healthcare technology company focused on developing safe, patient-facing large language models (LLMs) to improve healthcare accessibility. The company prioritizes safety and ethical AI development to address global healthcare worker shortages.
NVIDIA Corporation