LLM Inference Frameworks & Optimizations Engineer
Job in
San Francisco, San Francisco County, California, 94199, USA
Listed on 2026-09-09
Listing for:
Together AI
Full Time
position Listed on 2026-09-09
Job specializations:
-
Software Development
AI Engineer (Applied/Software), Machine Learning/ ML Engineer
Job Description & How to Apply Below
Together AI is building state-of-the-art LLM inference infrastructure to enable scalable, low-latency deployment for multimodal models. You will design and optimize distributed inference engines, collaborate with hardware and research teams, and push performance, scalability and cost-efficiency.
We seek an engineer with 3+ years in deep learning inference or HPC, proficient in Python and C++/CUDA, and hands-on experience with TensorRT, MoE and GPU optimization.
#J-18808-LjbffrTo View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×