Device LLM Inference Engineer Rust & GPU
Listed on 2026-09-30
-
Software Development
AI Engineer (Applied/Software), Machine Learning/ ML Engineer
Mirai Labs in San Francisco seeks engineers to join a senior team building the full on-device stack for real-time local intelligence. You will primarily work on uzu, our inference engine, and focus on supporting new modalities and a variety of features.
The ideal candidates should have a deep understanding of how computers and modern language models work, and experience in writing high-performance GPU kernels or Rust systems programming. We welcome applications from talented students and early-career engineers.
We invite applications for the On-Device LLM Inference Engineer — Real-Time, Rust & GPU position located in San Francisco, CA, United States.
For the On-Device LLM Inference Engineer — Real-Time, Rust & GPU position at Mirai Labs, we are reviewing applications now.
Step into the On-Device LLM Inference Engineer — Real-Time, Rust & GPU role at Mirai Labs in San Francisco, CA, United States and grow with us.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).