More jobs:
Inference Performance Engineer: Latency & Cost Optimization
Job in
San Francisco, San Francisco County, California, 94199, USA
Listed on 2026-06-18
Listing for:
OpenAI
Full Time
position Listed on 2026-06-18
Job specializations:
-
Manufacturing / Production
Systems Engineer
Job Description & How to Apply Below
OpenAI is looking for a performance modeler in San Francisco who will analyze inference stack performance and build cost-to-serve estimates. In this role, candidates should have expertise in performance profiling and enjoy reasoning about distributed systems.
The position offers a compensation range of $295K to $555K and requires collaboration with engineering and research teams to enhance performance and address system bottlenecks.
#J-18808-LjbffrTo View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×