×
Register Here to Apply for Jobs or Post Jobs. X

Service Reliability Engineer

Job in Manchester, Greater Manchester, M9, England, UK
Listing for: Fitch Group
Full Time position
Listed on 2026-09-06
Job specializations:
  • IT/Tech
    Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, AWS, Cybersecurity
Job Description & How to Apply Below

Fitch Group is currently seeking a Service Reliability Engineer to embed with Fitch Ratings development squads. The role is based out of our Manchester office.

As a leading global financial information services provider Fitch Group delivers vital credit and risk insights robust data and dynamic tools to champion more efficient transparent financial markets. With over 100 years of experience and colleagues in over 30 countries Fitch Groups culture of credibility independence and transparency is embedded throughout its structure which includes Fitch Ratings one of the worlds top three credit ratings agencies and Fitch Solutions a leading provider of insights data and analytics.

With dual headquarters in London and New York Fitch Group is owned by Hearst.

Fitchs Technology & Data Team is a dynamic department where innovation meets impact. Our team includes the Chief Data Office Chief Software Office Chief Technology Office Emerging Technology Shared Technology Services Technology Risk and the Executive Program Management Office (EPMO).Driven by our investment in cutting-edge technologies like AI and cloud solutions were home to a diverse range of roles and backgrounds united by a shared passion for leveraging modern technology to drive projects that matter to our organization and clients.

We are also proud to be recognized by Built In as a Best Place to Work in Technology 3 years in a row. Whether youre an experienced professional or just starting your career we offer an exciting and supportive environment where you can grow innovate and make a difference.

Want to learn more about a career in technology and data at Fitch

Visit:

About the Team

Fitch Group SRE provides Service Reliability Engineering expertise to Fitchs development organizations. This squad joins Core Engineering and other SRE groups as part of Cloud Infrastructure & Platform Engineering (CI&PE) serving as subject matter experts in cloud technologies systems engineering infrastructure automation and Dev Ops tooling across Fitch Group. The role itself will be dedicated to ensuring excellence in Fitch Ratings services with focus on new AI development.

How Youll Make an Impact:

  • Youll lead the delivery of reliable scalable mission-critical services. Youll guide squads on Kubernetes and modern deployment patterns.
  • Mentoring associate engineers and setting best practices for areas of expertise will help the team grow in knowledge and capabilities.
  • Partner closely with Fitch Ratings Development Squads and Operations to design and advance service builds Dev Ops tooling and drive operational excellence.
  • Partner with Core Engineering to architect and govern Git Hub Actions CI/CD with quality gates canary/bluegreen strategies and AI assisted redeploy checks
  • Own observability in Datadogdefine SLIs/SLOs dashboards alerting and MS Teams integrations and reduce incidents via telemetry-driven automation and blameless postmortems.
  • Champion AIenabled operations using AWS Bedrock/Sage Maker and Model Context Protocol (MCP) for log analysis anomaly detection incident triage and workflow orchestration; establish adoption guardrails.
  • Define and enforce cloud guardrails and security controls (SCPs/IAM boundaries OPA policies tagging centralized logging with AWS Config/Cloud Trail/Security Hub) in partnership with Security and Risk.
  • Influence cross functional roadmaps lead complex release planning and drive strategic platform initiatives across CI&PE; serve as an escalation point and participate in the L3 oncall rotation.

You May be a Good Fit if:

  • You have deep hands-on experience in SRE Dev Ops or Platform Engineering across both AWS and Azure with a strong track record operating Docker and Kubernetes in production environments.
  • Youre highly proficient in administering both Linux and Windows and have practical enterprise-level experience supporting IIS/.NET applications as well as Java Spring Boot services.
  • You have built and maintained CI/CD pipelines (primarily Git Hub Actions; Bamboo experience a plus) with Dev Sec Ops  principles baked inintegrating security scans policy-as-code and compliance gatesand script confidently in Python Power Shell or Bash.
  • You have experience with cloud security best practices (IAM secrets management container/image scanning) and understand core infrastructure fundamentals (networking storage DNS) and APM/telemetry tooling.

What Would Make You Stand Out:

  • Practical experience with agentic AI for operations incident triage runbooks and change management with clear guardrails auditability and human-in-the-loop controls.
  • Supporting AI/ML workloads at scale:
    Sage Maker endpoints GPU node groups autoscaling and Kubernetes-based model serving.
  • Policy-as-code (OPA) and compliance implementation across CIS NIST ISO 27001 with automated remediation integrated via CSPM tools (e.g. Wiz).
  • Applying AI in CI/CD observability and incident response using AWS Bedrock/Sage Maker and Model Context Protocol (MCP).
  • Hands-on Agile delivery experience actively participating in stand-ups…
Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary