More jobs:
Member of Technical Staff - ML Systems & Inference
Job in
San Francisco, San Francisco County, California, 94199, USA
Listed on 2026-07-14
Listing for:
Acceler8 Talent
Full Time
position Listed on 2026-07-14
Job specializations:
-
Software Development
AI Engineer (Applied/Software), Machine Learning/ ML Engineer, AI Reliability/ Performance Engineer, Cloud Engineer - Software
Job Description & How to Apply Below
Member of Technical Staff - ML Systems & Inference
Join a well-funded AI infrastructure startup building the orchestration layer for next-generation AI workloads
This role sits at the intersection of ML systems, inference, distributed systems, and performance engineering. You'll build production inference systems, optimize scheduling and memory management, improve KV cache efficiency, and work closely with compiler, kernel, and distributed systems engineers to push AI infrastructure forward
We're looking for engineers with:
- Strong software engineering fundamentals
- Experience with ML inference or model serving
- Knowledge of distributed systems and performance optimization
- Python and/or C++
- Experience with vLLM, TensorRT-LLM, CUDA, or similar is a plus
This is an opportunity to help define how AI workloads are executed at scale alongside a world-class engineering team
Interested? Apply or message me directly
#J-18808-LjbffrTo View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×