×
Register Here to Apply for Jobs or Post Jobs. X

Senior Site Reliability Engineer

Job in Pleasanton, Alameda County, California, 94566, USA
Listing for: Oracle
Full Time position
Listed on 2026-07-08
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations
Salary/Wage Range or Industry Benchmark: 81100 - 187000 USD Yearly USD 81100.00 187000.00 YEAR
Job Description & How to Apply Below

Overview

We are looking for a Site Reliability Engineer 3 to support mission‑critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection and resolution of issues.

The engineer will work closely with development, infrastructure, security, and operations teams to monitor service health, troubleshoot production issues, participate in incident response, improve observability, and implement reliability best practices. This role also includes analyzing recurring failures, building automation, supporting deployments, and contributing to capacity planning, disaster recovery, and operational readiness.

Responsibilities
  • Takes proactive steps to design and architect infrastructure and/or services according to reliability standards.
  • Forecasts demands for infrastructure and responds to capacity needs, ensuring systems can handle current and future workloads.
  • Collaborates with software development teams to build reliable and scalable infrastructures and features.
  • Identifies and drives prototyping opportunities, including testing new applications or infrastructure and assisting onboarding.
  • Performs data collection, triage, technical analysis, and redirection to maintain and optimize operations and infrastructure reliability.
  • Monitors services, keeps performance knowledge up‑to‑date, and documents their status.
  • Executes incident response, root‑cause analysis, and maintenance (e.g., software installs, upgrades, security updates, backup and recovery).
  • Provides health and performance reporting and acts on trends.
  • Provisioning services, applications, and infrastructure when required.
  • Performs standard and non‑standard decommissioning of unused resources.
  • Identifies automation opportunities and assesses benefits.
  • Develops automation tools or scripts to gather metrics, monitor, analyze, mitigate, or remediate infrastructure issues.
  • Tests automation to ensure correct performance and expected results.
  • Communicates scale, capacity, security, performance attributes, and requirements of services to current and external teams.
  • Explains potential impacts of infrastructure, feature, and tool changes on operations.
  • Provides operational support for technology, escalating incidents and other issues.
  • Participates in on‑call shifts to address issues.
  • Investigates and resolves technical issues across services to meet SLOs.
  • Documents incidents and performs root‑cause analyses following standard reporting methods.
  • Conducts post‑mortem procedures to prevent recurrence.
  • Experiments with new tools and technologies, assessing impact and ensuring security compliance.
  • Identifies and implements improvements for performance bottlenecks and efficient resource use.
  • Shares reliability trends and best practices with team members, management, and others.
  • Performs data analysis to support business development decisions.
  • Manages work, timelines, and deliverables to keep projects on track.
  • Collaborates across teams to align on expectations and achieve shared objectives.
  • Addresses standard and non‑standard issues, escalating complex problems as needed.
  • Engages in continuous learning, staying current with industry trends and best practices.
  • Recommends updates to improve processes, protocols, and workflows.
Technical Skills

IAC:
Terraform, Chef, Ansible

Languages:

Python, Java, Bash

Orchestration:
Kubernetes, Helm

CI/CD:
Jenkins

Observability:
Grafana, Prometheus

Qualifications

Candidate must comply with applicable U.S. immunization, occupational health, and drug‑testing requirements for U.S. based or customer‑facing roles.

Hiring Range (U.S.): USD $81,100 – $187,000 per year; may be eligible for bonus and equity.

Benefits
  • Medical, dental, and vision insurance, including expert medical opinion
  • Short‑term and long‑term disability
  • Life insurance and AD&D
  • Supplemental life insurance (Employee/Spouse/Child)
  • Health care and dependent care Flexible Spending Accounts
  • Pre‑tax commuter and parking benefits
  • 401(k) savings and investment plan with company match
  • Paid time off with flexible vacation accrual and paid sick leave
  • Paid parental leave
  • Adoption assistance
  • Employee stock purchase plan
  • Financial planning and group legal services
  • Voluntary benefits including auto, homeowner, and pet insurance
Equal Employment Opportunity

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

#J-18808-Ljbffr
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary