LLM Inference Runtime Architect AI Accelerator
Listed on 2026-09-30
-
Software Development
AI Engineer (Applied/Software), Machine Learning/ ML Engineer, AI Reliability/ Performance Engineer
United States Digital Space LLC is hiring a production-grade LLM inference runtime engineer to join the Hardware AI team. You will design and optimize the end-to-end runtime that maps frontier models onto our AI accelerator, balancing latency, throughput, and hardware utilization.
Work spans model architecture, systems software, and silicon interfaces, with emphasis on reliability, observability, and scalable performance in production environments.
Join us at United States Digital Space LLC as our next LLM Inference Runtime Architect for AI Accelerator in United States.
Take a moment to read everything above and see whether this role is right for you.
This posting is for the LLM Inference Runtime Architect for AI Accelerator role at United States Digital Space LLC, based in United States.
We are looking to fill the LLM Inference Runtime Architect for AI Accelerator position at United States Digital Space LLC in United States.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).