×
Register Here to Apply for Jobs or Post Jobs. X
More jobs:

Site Reliability Engineer

Job in Greater London, London, Greater London, W1B, England, UK
Listing for: Autonomai Recruitment
Full Time position
Listed on 2026-08-04
Job specializations:
  • Software Development
    Unix/Linux
Salary/Wage Range or Industry Benchmark: 90000 - 130000 GBP Yearly GBP 90000.00 130000.00 YEAR
Job Description & How to Apply Below
Location: Greater London

A high-performing trading technology firm is seeking an SRE Engineer to drive reliability, scalability, and operational excellence across critical production systems. This role is suited to a hands-on engineer who operates across infrastructure, software, and platform environments, with a strong focus on resilience, automation, and engineering quality.

You will work closely with software engineering, platform, security, and infrastructure teams to improve service availability, strengthen observability, and enhance operational practices in a high-performance, low-latency environment.

Responsibilities
  • Own the reliability, performance, and availability of business-critical production systems and infrastructure
  • Support incident response, service restoration, root cause analysis, and post-incident reviews
  • Implement SRE best practices including monitoring, alerting, and service health standards
  • Build and improve automation to reduce operational toil and enhance deployment consistency and recovery
  • Partner with engineering teams to improve system design, fault tolerance, and capacity planning
  • Contribute to production readiness for new services and infrastructure changes
  • Maintain and improve observability across systems (metrics, logging, tracing)
  • Continuously improve operational processes and platform reliability
Requirements
  • Strong hands-on experience operating and troubleshooting Linux-based production environments
  • Solid understanding of distributed systems and streaming architectures
  • Experience with messaging platforms
  • Proficiency in at least one systems-level language (C, C++, Rust, or similar)
  • Familiarity with large-scale data systems
  • Experience with CI/CD pipelines, build systems, and modern development workflows
  • Working knowledge of containerised environments
  • Proven ability to troubleshoot and resolve complex production issues in high-availability systems
Preferred Profile
  • Background in a high-scale or performance-sensitive engineering environment (e.g. trading, FAANG, or similar)
  • Experience supporting low-latency or highly distributed infrastructure
  • Track record of improving automation, observability, and system reliability
  • Strong production mindset with a focus on stability, performance, and continuous improvement

This is an opportunity to work on critical, real-time systems in a high-performance environment, with direct impact on production reliability and trading outcomes.

#J-18808-Ljbffr
Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary