Remote Inference Engine Engineer - LLMs & Diffusion
Chicago, Cook County, Illinois, 60290, USA
Listed on 2026-09-30
-
Software Development
AI Engineer (Applied/Software)
Inferact is seeking an inference runtime engineer to push the boundaries of LLM and diffusion model serving. You will optimize how models execute across diverse hardware and architectures, contributing to the core of vLLM and related projects.
This fully remote role offers salary plus equity, visa sponsorship on a case-by-case basis, and comprehensive benefits. Regular overlap with Pacific Time is expected for critical syncs.
Consider building your career as a Remote Inference Engine Engineer - LLMs & Diffusion at Inferact.
Please review the full job details above before applying.
If your experience matches this role, we encourage you to apply.
All applications are reviewed carefully by our team.
This is a Full Time role.
The position is based in United States.
This opportunity is part of our work in IT & Technology.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).