×
Register Here to Apply for Jobs or Post Jobs. X

Senior Site Reliability Engineer

Job in Coeur d Alene, Kootenai County, Idaho, 83814, USA
Listing for: Autodesk, Inc.
Full Time position
Listed on 2026-08-16
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, AWS
Salary/Wage Range or Industry Benchmark: 117000 - 209000 USD Yearly USD 117000.00 209000.00 YEAR
Job Description & How to Apply Below

Job Requisition  # 26WD99276 Position Overview Want to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk Gov Cloud products. As part of a new SRE team supporting Autodesk Gov Cloud, you will have a unique opportunity to help shape how Autodesk deploys, runs, and improves production services in restricted cloud environments.

This is a foundational role where you will help establish the operating model, reliability practices, automation, and engineering standards needed to support critical customer-facing services. You will combine software engineering and production operations to deploy, run, monitor, improve, and automate Autodesk services in Gov Cloud. You will partner closely with product engineering, security, compliance, platform, and infrastructure teams to ensure services are reliable, scalable, secure, and ready for production.

The ideal candidate has deep experience operating production systems at scale, an automation‑first mindset, and the ability to improve reliability through engineering practices such as SLOs/SLIs, production readiness, incident management, observability, resilience testing, and toil reduction. Success in this role requires strong technical judgment, a customer-focused mindset, and a passion for using software engineering to solve operational problems  accordance with Gov Cloud Cloud Service Provider Security Requirements, this role must be performed by U.S. Citizens.

Employment is contingent upon meeting all applicable government security and eligibility requirements, including necessary background investigations and government issued security clearances.

Responsibilities

Serve as a primary owner for the reliability, availability, performance, operability, and capacity of one or more production services

Deploy, operate, maintain, and continuously improve production services running in Autodesk Gov Cloud environments

Partner with engineering teams to ensure services are designed with reliability, scalability, security, and operability in mind

Define and operate reliability practices such as SLOs/SLIs, error budgets, production readiness reviews, service reviews, and operational health reviews

Build automation to improve deployment safety, operational efficiency, incident response, and service recovery

Design, develop, and maintain software, automation, and tooling that improve the reliability, scalability, and efficiency of production systems

Implement and improve monitoring, alerting, logging, tracing, and observability capabilities across supported services

Lead and participate in incident response, troubleshooting, and post-incident reviews focused on learning and continuous improvement

Develop and maintain operational documentation, runbooks, and recovery procedures

Scale and enhance resilience testing and Gameday practices to validate system behavior, recovery capabilities, and operational readiness

Continuously identify and eliminate operational toil through software engineering, automation, and process improvement

Ensure supported services remain compliant with Autodesk security, privacy, and regulatory requirements, including FedRAMP and related controls where applicable

Participate in a 24x7 on‑call rotation for production services

Function effectively in a fast‑paced environment while helping establish and mature operational excellence practices for Autodesk Gov Cloud

Minimum Qualifications

B.S. or higher in Computer Science, Engineering, or a related technical discipline, or equivalent practical experience

7+ years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Cloud Infrastructure, or Production Operations

Experience operating and supporting customer‑facing production services in large‑scale cloud environments

Strong understanding of reliability engineering principles, including SLOs/SLIs, observability, incident management, capacity planning, production readiness, and automation

Experience with AWS, Azure, or other public cloud platforms

Experience developing automation using languages such as Python, Go,…

Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary