×
Register Here to Apply for Jobs or Post Jobs. X

Incident Response Engineer | REF#299191

Remote / Online - Candidates ideally in
Los Angeles, Los Angeles County, California, 90079, USA
Listing for: BairesDev
Remote/Work from Home position
Listed on 2026-08-08
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Systems Engineer
Job Description & How to Apply Below

At Baires Dev®, we've been leading the way in technology projects for over 15 years. We deliver cutting-edge solutions to giants like Google and the most innovative startups in Silicon Valley.

Our diverse 4,000+ team, composed of the world's Top 1% of tech talent, works remotely on roles that drive significant impact worldwide.

When you apply for this position, you're taking the first step in a process that goes beyond the ordinary. We aim to align your passions and skills with our vacancies, setting you on a path to exceptional career development and success.

Incident Response Engineer at Baires Dev

As an Incident Response Engineer, you will own the entire incident lifecycle to ensure high system availability and minimal downtime. You will serve as the primary point of contact during production outages, coordinating technical response efforts and driving issues to resolution through clear communication and expert troubleshooting.

What You’ll Do
  • Manage the end-to-end incident lifecycle, including detection, triage, mitigation, and resolution of production issues.
  • Utilize Pager Duty and monitoring tools to maintain high-velocity response times and meet established SLOs/SLIs.
  • Facilitate blameless post-mortems to identify root causes and prevent recurrence of critical system failures.
  • Develop and maintain comprehensive runbooks and automated response scripts to reduce mean time to recovery (MTTR).
  • Partner with SRE and engineering teams to improve system observability and design more resilient architectures.
  • Oversee incident communication channels, providing timely updates to technical stakeholders and leadership during active events.
What We Are Looking For
  • 4+ years of experience in Infrastructure, Site Reliability Engineering, or Incident Management.
  • Proven expertise in owning the incident lifecycle, including detection, response, and post-mortem analysis.
  • Advanced proficiency in using Pager Duty for incident orchestration and communication.
  • Hands-on experience building runbooks and defining SLOs/SLIs to improve system reliability.
  • Deep understanding of Linux systems and modern observability practices.
  • Advanced proficiency in English.
How we do make your work (and your life) easier:
  • 100% remote work (from anywhere).
  • Excellent compensation in USD or your local currency if preferred
  • Hardware and software setup for you to work from home.
  • Flexible hours: create your own schedule.
  • Paid parental leaves, vacations, and national holidays.
  • Innovative and multicultural work environment: collaborate and learn from the global Top 1% of talent.
  • Supportive environment with mentorship, promotions, skill development, and diverse growth opportunities.

Join a global team where your unique talents can truly thrive and make a significant impact!

#J-18808-Ljbffr
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary