Search
Go
On-Device AI Inference Engineer is a AI Infrastructure Engineer role (full-time). with Hark. in SAN JOSE, US. Compensation shown: $200K–$450K. Imported listing (source: zerogtalent.com). Apply on the employer's site (zerogtalent.com).
Imported listing (source: zerogtalent.com) · Apply on zerogtalent.com
The sections below reproduce the third-party job description for reference. AIEngineer.careers does not write or control this text.
Sourced from zerogtalent.com
Hark is seeking a Senior On-Device AI Inference Engineer to optimize transformer workloads for next-generation hardware. You will write low-level kernels and runtime paths to ensure models run efficiently within strict latency, power, and memory budgets on constrained silicon like DSPs and NPUs. This role bridges the gap between model architecture and physical device constraints.

San Jose · US · 70+ employees
Hark is an artificial intelligence company focused on building proactive, multimodal personal intelligence systems that integrate foundation models with bespoke, AI-native hardware. Founded by serial entrepreneur Brett Adcock, the company aims to create a universal, seamless interface between humans and machines that can see, listen, and reason with persistent memory.
Figure