×
Register Here to Apply for Jobs or Post Jobs. X

Senior Staff Engineer

Job in Richardson, Dallas County, Texas, 75080, USA
Listing for: Talentify
Full Time position
Listed on 2026-10-03
Job specializations:
  • IT/Tech
    SRE/Site Reliability, IT Support, Cloud Computing: Infrastructure & Operations, Systems Engineer
Salary/Wage Range or Industry Benchmark: 120000 - 260000 USD Yearly USD 120000.00 260000.00 YEAR
Job Description & How to Apply Below
Why Join GEICO?

At GEICO, we offer a rewarding career where your ambitions are met with endless possibilities.

Every day we honor our iconic brand by offering quality coverage to millions of customers and being there when they need us most. We thrive on relentless innovation to exceed our customers' expectations while making a real impact on local communities nationwide.

Founded in 1936, GEICO is a member of the Berkshire Hathaway family of companies and one of the largest auto insurers in the United States. When you join our company, we want you to feel valued, supported, and proud to work here. That's why we offer the GEICO Pledge:
Great Company, Great Culture, Great Rewards, and Great Careers.

Position Summary

Technology Operations Center is at the core of GEICO’s application and platform resiliency. It assists GEICO’s engineering teams with maintaining high availability of our customer and internally facing services while driving down time to detect and recover from incidents. It is governing our incident management processes and builds platforms that allow GEICO to manage and recover from incidents.

GEICO is seeking an experienced SRE Software Engineer with a passion for building, operating and troubleshooting high-performance, low-maintenance, zero-downtime complex distributed platforms and applications. You will help drive our transformation to a tech organization with engineering excellence and site reliability as its mission, by defining, implementing and operating our incident management processes and in-house technology platforms that automate them.

This role focuses on improving Incident Management tooling and process across GEICO. It is a hands-on technical leadership role focused on better incident management, faster time to detect, troubleshoot and recover from incidents and fewer repeat incidents, by designing, developing and operating the the tools and processes that help all GEICO engineering teams manage their on-call, runbooks, troubleshooting and BCDR.

Success in this role requires strong technical depth and equally strong process leadership. The right candidate can go deep on incident analysis, system behavior, and architecture, while also improving how teams run COEs, learn from incidents, and turn those lessons into engineering improvements.

Why This Role Is Different
  • This role blends deep technical understanding, hands on execution and ability to build SW with real time incident leadership, platform and process improvements. You will be driving our incident management and response processes and the platforms that automate them.
  • You are shaping how GEICO service engineering teams detect, manage, resolve and learn from the incidents.
  • Your work directly impacts the availability of GEICO’s critical applications and platforms, experiences and satisfaction of millions of customers and tens of thousands of associates.
Position Responsibilities

As a Senior Staff Engineer, you will:

Own and Evolve Enterprise-Critical Platforms
  • Design, develop and operate automation, self-service tools, dashboards, and data pipelines that automate and scale our incident management, on-call, paging and troubleshooting processes.
  • Build shared services, APIs, data contracts, automation, and integrations that standardize incident response and reduce systemic operational risk.
  • Set and uphold engineering standards across design, implementation, deployment, testing, observability, security, operational support, and production readiness.
  • Champion safe deployment, CI/CD, infrastructure as code, automated testing, rollback patterns, and operational controls that support frequent and reliable delivery.
  • Evaluate, select, and implement modern technologies and tools that improve platform capability, compliance, visibility, reliability, and engineering effectiveness.
Serve as a Senior Technical Authority During Incidents
  • Act as a technical leader and escalation point during high-severity incidents, bringing architectural judgment, system-level problem solving, and calm execution under pressure.
  • Guide troubleshooting strategy, cross-team coordination, impact analysis, and risk-based decision making to restore service safely and efficiently.
  • Lead or heavily influence post-incident reviews, root cause analysis, corrective action planning, and systemic reliability improvements.
  • Develop and maintain incident response strategies, operational runbooks, readiness criteria, triage models, and resilience practices across multiple integration points.
Influence…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary