Senior LLM Inference Architect; On-Site Menlo Park
Listed on 2026-10-06
-
IT/Tech
AI Engineer (Applied/Software), Machine Learning/ ML Engineer
Hippocratic AI Inc. is seeking an experienced LLM Inference Engineer to own and optimize its inference serving stack. You will drive sub-100ms responses, cost efficiency, and reliable availability for millions of patient conversations across healthcare systems.
You will design multi-node serving architectures and advance techniques like disaggregated serving and speculative decoding. You will collaborate with systems engineers, ML researchers, and infrastructure experts to push the boundaries of
We are currently recruiting a Senior LLM Inference Architect (On-Site Menlo Park) for our team in Menlo Park, CA, United States.
Consider building your career as a Senior LLM Inference Architect (On-Site Menlo Park) at Hippocratic AI Inc.
The Senior LLM Inference Architect (On-Site Menlo Park) position in the IT & Technology field is open for applications.
We have an opening for a Senior LLM Inference Architect (On-Site Menlo Park) in Menlo Park, CA, United States within IT & Technology.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).