Senior Principal AI Solutions Engineer is a AI Solutions Engineer role (full-time). with SambaNova. in REMOTE, WORLDWIDE. Compensation shown: $234K–$286K. Imported listing (source: aihiringboard.com). Apply on the employer's site (aihiringboard.com).
Imported listing (source: aihiringboard.com) · Apply on aihiringboard.com
The sections below reproduce the third-party job description for reference. AIEngineer.careers does not write or control this text.
Imported job description
Sourced from aihiringboard.com
SambaNova is hiring a Senior Principal AI Solutions Engineer to lead hands-on engagements with strategic customers, building production-grade agentic, multimodal, and voice AI applications on its full-stack inference platform. This is a senior technical leadership role that combines architecture direction, customer advisory, and full-stack development across open-weight models, agent frameworks, and modern front-end tooling.
Responsibilities
Design and build single- and multi-agent systems with planning, tool/function calling, MCP, memory, and long-running workflows
Build self-improving systems with automated evals, LLM-as-judge, trace-driven prompt/program optimization, synthetic data, and fine-tuning
Set up eval harnesses, tracing, and guardrails for quality, cost, and latency observability
Select, adapt, and combine open-weight models (Llama, Qwen, DeepSeek, Gemma, GLM) including LoRA/PEFT fine-tuning and model routing
Benchmark end-to-end solutions and validate new models on SambaNova's platform stack
Build polished full-stack web applications (React/Next.js, TypeScript) for knowledge workers
Build vertical solutions for financial services, legal, healthcare, public sector, and research
SambaNova is an AI infrastructure company founded in 2017 that builds a full-stack platform spanning purpose-built RDU AI chips, systems, software, and foundation-model services. It helps developers, enterprises, governments, and data centers run secure, high-performance inference and other AI workloads in cloud, on-premises, hybrid, and air-gapped environments.
Set technical direction for the solutions portfolio and mentor senior engineers
Act as trusted technical advisor to customer CTOs and AI leaders
Lead technical discovery, workshops, hackathons, and co-builds with strategic customers
Publish reference architectures, starter kits, and open-source examples
Develop tooling and automation for SambaStack deployment and integration
Feed field learnings into model roadmap and platform priorities
Requirements
Bachelor's degree or higher in Computer Science, Electrical Engineering, Applied Mathematics, Physics, Statistics, or related field
8+ years (IC5) or 10+ years (IC6) of industry experience in software, ML, or solutions engineering
3+ years building LLM-based applications that reached production
Proven experience building agentic systems with tool/function calling and multi-agent orchestration (LangGraph, CrewAI, OpenAI Agents SDK, Claude Agent SDK, or equivalent)
Hands-on experience with open-weight models: serving, prompting, fine-tuning (LoRA/PEFT), and evaluation
Strong full-stack skills: expert Python plus modern front-end (TypeScript, React/Next.js), APIs, and cloud deployment
Experience designing evals and benchmarks for LLM applications covering quality, latency, and cost
Excellent customer communication and ability to lead workshops and explain architecture to executives and engineers
Track record of technical leadership across teams: setting architecture direction, mentoring senior engineers, and owning strategic account outcomes
Nice to Have
Experience building self-improving or learning systems: prompt/program optimization, RL from feedback, synthetic data pipelines, or continual fine-tuning
Experience building real-time voice agents (LiveKit, Pipecat, WebRTC) and working with speech models (ASR/TTS)
Experience with multimodal and vision-language models and document-understanding pipelines
Domain experience in financial services, legal, healthcare, public sector, or scientific research
Familiarity with inference frameworks (vLLM, SGLang, TensorRT-LLM) and hardware-aware performance tuning
Experience with MCP, RAG at enterprise scale, and agent security and guardrails
Open-source contributions, public demos, or technical content in the AI community
Benefits
Health insurancedental insurancevision insurancehsafsalife insurancedisability insuranceEquity / stock optionswellness programgym membershipmental health supporteap