×
Register Here to Apply for Jobs or Post Jobs. X

Senior Infrastructure Engineer - AI Ops

Job in Austin, Travis County, Texas, 78716, USA
Listing for: Commerce.com
Full Time position
Listed on 2026-08-28
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, IT Infrastructure
Salary/Wage Range or Industry Benchmark: 135960 - 203940 USD Yearly USD 135960.00 203940.00 YEAR
Job Description & How to Apply Below

Welcome to the Agentic Commerce Era At Commerce, our mission is to empower businesses to innovate, grow, and thrive with our open, AI-driven commerce ecosystem. As the parent company of Big Commerce, Feedonomics, and Makeswift, we connect the tools and systems that power growth, enabling businesses to unlock the full potential of their data, deliver seamless and personalized experiences across every channel, and adapt swiftly to an ever-changing market.

We believe in harnessing AI responsibly to unlock new possibilities, and we’re looking for individuals who use it intentionally to solve problems, accelerate outcomes, and expand what’s possible in their role. Our purpose is to help businesses confidently solve complex commerce challenges so they can build smarter, adapt faster, and grow on their own terms. If you want to be part of a team of bold builders, sharp thinkers, and technical trailblazers who shape the future of commerce, this is the place for you.

We’re hiring a Senior Infrastructure Engineer to join Commerce’s AI Operations team — a highly autonomous team focused on rapid development of AI tools. We operate like a startup within the organization, with a small headcount and full ownership from idea through production. You’ll own the infrastructure, delivery pipelines, and cloud architecture that enable every AI automation. This is a hands‑on infrastructure role.

What You’ll Do:
  • Design, provision, and manage cloud infrastructure on GCP (Cloud Run, GKE, Redis/Memory store) using Terraform and Puppet.
  • Build and maintain CI/CD pipelines in Cloud Build, Git Hub Actions, and CircleCI—including automated testing gates, canary and blue‑green deployments, and artifact promotion across environments.
  • Design and maintain the full edge‑to‑origin traffic stack: DNS, load balancing (NGINX, GCP Cloud Load Balancing), CDN, WAF, edge compute (Cloudflare, Cloud Armor), bot management, authentication, and rate limiting.
  • Operate and tune container orchestration for AI workloads—autoscaling policies, resource quotas, cold‑start optimization, and right‑sizing compute for inference cost efficiency.
  • Implement production observability end‑to‑end: structured logging (Cloud Logging, ELK/Kibana), metrics and dashboards (Prometheus, Grafana), distributed tracing, SLO/SLI definition, and alerting.
  • Own incident response coordination and blameless post‑mortems.
  • Implement strong security posture—least‑privilege IAM with Workload Identity, network segmentation, firewall rules, and VPC Service Controls.
  • Evaluate, build, and integrate infrastructure tooling—whether that’s custom automation in Python or a managed AI serving platform.
Who You Are:
  • 7+ years in infrastructure engineering, Dev Ops, SRE, or platform engineering with direct operational ownership of production systems.
  • Deep Linux systems administration experience (Debian, Ubuntu) at scale.
  • Strong hands‑on experience with GCP services (Cloud Run, GKE, Cloud Build, IAM, VPC).
  • Infrastructure as code as a core discipline (Terraform, Puppet).
  • Experience designing and maintaining non‑trivial CI/CD delivery pipelines. Cloud Build, Git Hub Actions, CircleCI.
  • Hands‑on experience with NGINX and edge infrastructure at scale, including DNS, load balancing, and CDN (Cloudflare or equivalent).
  • Hands‑on experience with distributed monitoring and logging stacks—Prometheus, Grafana, ELK/Kibana, Cloud Logging.
  • Proficiency in Python, Bash, and shell scripting for infrastructure automation and custom tooling.
  • You design infrastructure with least‑privilege IAM, Workload Identity, and compliance controls as defaults.
  • Nice to have: experience with ML infrastructure.
Compensation Transparency

#LI-LH1 #LI-Hybrid (Pay range transparency $135,960 - $203,940)

  • The national base salary range for this role is posted above in this job post.
  • Final compensation will be determined based on factors such as relevant experience, skills, qualifications and geographic location.
  • We also consider internal equity to help ensure fair and consistent pay practices across our teams.
  • Where applicable, this role may also be eligible for variable compensation (such as bonus or commission), equity, and benefits in accordance with…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary