×
Register Here to Apply for Jobs or Post Jobs. X

Sr. Software Engineer, Site Reliability

Job in Indianapolis, Hamilton County, Indiana, 46262, USA
Listing for: JobCubby
Full Time position
Listed on 2026-09-06
Job specializations:
  • IT/Tech
    SRE/Site Reliability, IT Support
Salary/Wage Range or Industry Benchmark: 115000 - 150000 USD Yearly USD 115000.00 150000.00 YEAR
Job Description & How to Apply Below
Location: Indianapolis

At Bloomerang, we believe change happens on purpose. We champion the power and potential of nonprofits, igniting next-level impact with the team and technology built for purpose. Our powerful giving platform and stellar support enable tens of thousands of nonprofits to raise more, recruit more, and retain more, fueling maximum impact and raising the bar on what’s possible for the nonprofit sector.

That's why, even as the nonprofit sector sees declines in giving, Bloomerang customers raise more year over year.

We're also in the business of creating thriving employees. Join a mission-driven culture built on our core values of Simplify, Care and Act. We know our people are the key to our success, and we're proud to be home to some of the most innovative and skilled individuals in the workforce today. Come feel invigorated and unstoppable with us!

The Role

We are evolving our Tier 3 Support Engineering team into a Site Reliability Engineering (SRE) organization focused on improving product reliability, observability, and operational efficiency. We're looking for an experienced SRE who brings strong software engineering fundamentals and is excited to help shape this transformation and mature our SRE practices.

This is a highly collaborative, hands-on role working across application code, telemetry, databases, APIs, and infrastructure to diagnose complex production issues and improve reliability. Production support and product defects remain part of today's work as you help reduce reactive effort through observability, SLOs, automation, permanent fixes, and proactive reliability engineering.

What Success Looks Like

Success isn't measured solely by issues resolved, but by issues that no longer require manual intervention. You'll help detect problems earlier, reduce recurring issues and toil, strengthen incident response, and create more capacity for proactive reliability engineering.

What You Will Do
  • Own complex production support escalations and ticket triage, providing hands-on troubleshooting and resolution alongside reliability work.
  • Partner with Software Engineering to investigate complex production issues, identify root causes and reliability risks, and drive permanent solutions to recurring problems and defects.
  • Bring proven SRE practices to the team and foster proactive reliability, continuous improvement, automation, and shared ownership.
  • Lead incident response from triage and mitigation through recovery, root cause analysis, and blameless post-incident reviews, turning lessons learned into reliability improvements.
  • Build observability across products, services, and critical customer workflows using meaningful metrics, logs, traces, dashboards, and actionable alerts.
  • Define and mature SLIs and SLOs that measure system reliability and customer experience.
  • Develop synthetic monitoring for critical customer journeys to detect failures before they impact customers.
  • Identify sources of recurring operational toil and drive automation, tooling, process improvements, or permanent fixes that reduce manual effort.
  • Use AI-assisted tools and source code repositories to accelerate triage, troubleshooting, code analysis, automation, and technical investigation.
  • Participate in a rotating on-call schedule, primarily during business hours, with limited after-hours and weekend support.
What You Need to SucceedSRE Experience & Transformation
  • Hands-on Site Reliability Engineering experience applying software engineering practices to production reliability and helping establish or mature SRE practices.
  • Strong knowledge of SLIs, SLOs, error budgets, observability, automation, and toil reduction.
Observability & Incident Management
  • Experience building monitoring, dashboards, alerts, and telemetry using tools such as Honeycomb, New Relic, Grafana, Cloud Watch, Kibana, or similar.
  • Experience with production incident management, root cause analysis, blameless post-incident reviews, and corrective-action follow-through.
Technical Depth
  • Strong programming and scripting skills to navigate and troubleshoot application code and build automation and operational tooling.
  • Strong SQL and relational database skills for production troubleshooting…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary