SRE AI Infra & Multi-Cloud Platform
Listed on 2026-10-07
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer
Mithril is seeking an experienced Site Reliability/Infrastructure Engineer to scale its GPU orchestration platform across multiple clouds. You will own reliability, automation, observability, and tooling, collaborating with the founding team on capacity planning and SLO design.
You will build infrastructure as code, automate operations with Python or Go, and help manage GPU capacity while shaping how Mithril delivers fast, reliable access to its platform.
We have an opening for a SRE for AI Infra & Multi-Cloud Platform in Palo Alto, CA, United States within IT & Technology.
We aim to respond to suitable candidates as soon as possible.
Full responsibilities and requirements are described in the listing above.
Learn more about the SRE for AI Infra & Multi-Cloud Platform role in the description above.
We appreciate your interest in this position.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).