×
Register Here to Apply for Jobs or Post Jobs. X

Senior Cloud​/ML Ops Engineer

Job in San Francisco, San Francisco County, California, 94199, USA
Listing for: Icehouseventures
Full Time position
Listed on 2026-06-23
Job specializations:
  • IT/Tech
    Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, IT Infrastructure
Salary/Wage Range or Industry Benchmark: 125000 - 150000 USD Yearly USD 125000.00 150000.00 YEAR
Job Description & How to Apply Below

The Role:

Why, What and the Who

Infrastructure Engineers build the foundation for Ivo’s entire platform. Customers are cagey about their contracts, so each customer gets an isolated environment with containers, database, VPC, etc. Things break; regions go down; cloud and LLM providers have incidents. Customers still expect us to hit our SLAs.

Why

Infrastructure Engineers build the foundation for Ivo’s entire platform. Customers are cagey about their contracts, so each customer gets an isolated environment with containers, database, VPC, etc. Things break; regions go down; cloud and LLM providers have incidents. Customers still expect us to hit our SLAs.

What
  • Own and evolve our Kubernetes platform across AWS, GCP, and Azure.
  • Design and operate multi‑cluster, multi‑region architectures with failover and disaster recovery strategies adhering to secure cluster isolation boundaries.
  • Build internal tooling for cluster provisioning and lifecycle management, standardizing environments (dev → staging → prod).
  • Design strategies to isolate ML vs API workloads while optimizing for cost, performance, and reliability.
  • Implement security and compliance controls at the platform layer with RBAC, workload identity, secrets management while preserving data isolation aligned with residency requirements and auditability for enterprise customers.
  • Partner with SRE and ML teams to ensure realistic and enforceable SLOs and to deploy models in production environments.
Who
  • Deep, hands‑on experience with Kubernetes in production (debugging it at 2 a.m., not just deploying).
  • Strong experience with infrastructure as code (Pulumi, Terraform, etc.).
  • Strong understanding of cluster architecture, scheduling, networking, storage primitives, and failure modes in distributed systems.
  • Experience managing multi‑cluster or multi‑region setups with Git Hub CI/CD.
Ivo might be a good fit for you if you:
  • You love writing code, but you love having impact more: building a great product and making pragmatic choices.
  • You describe yourself as relentlessly resourceful.
  • You have a strong internal sense of urgency and a bias toward doing things today, not tomorrow.
  • Experience working in a startup environment is preferred but not required.
  • Are excited about the adventure of building a company!
#J-18808-Ljbffr
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary