×
Register Here to Apply for Jobs or Post Jobs. X

Site Reliability Engineer

Job in Belfast, County Antrim, BT1, Northern Ireland, UK
Listing for: Cantor Fitzgerald
Full Time position
Listed on 2026-08-29
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer
Job Description & How to Apply Below
At Lucera, we're seeking an experienced SRE/Dev Ops Engineer to join our engineering team. You'll play a crucial role in maintaining the reliability and performance of our global trading infrastructure and financial technology platforms.

This role offers an opportunity to work in a high-stakes, performance-driven environment, where your expertise in automation, infrastructure management, and operational excellence will be paramount. 5+ years of experience in Site Reliability Engineering, Dev Ops, Platform Engineering, or Infrastructure Engineering. Hands-on expertise with Infrastructure as Code technologies (Terraform, Ansible, etc.).

Experience with container orchestration platforms (Kubernetes, Nomad, Open Shift). Proven track record in designing and maintaining CI/CD pipelines. Strong Linux systems administration and scripting skills (Bash, Shell, Python).

Experience with observability platforms (Grafana, InfluxDB) and automation. Ability to troubleshoot complex distributed systems and applications. Familiarity with Git and modern source control workflows. Understanding of high availability, fault tolerance, and disaster recovery principles. Excellent communication and collaboration skills for effective team work. Design, build, and maintain highly available, scalable, and resilient production infrastructure. Implement and manage Infrastructure as Code (IaC) solutions for automated provisioning and configuration.

Enhance CI/CD pipelines to ensure efficient and reliable software delivery. Manage containerized workloads across modern orchestration platforms. Build and improve monitoring, alerting, and observability solutions for rapid issue resolution. Automate operational workflows, deployment processes, and administrative tasks. Troubleshoot complex production incidents across various systems. Collaborate with development teams to enhance system reliability and performance. Participate in incident management, root cause analysis, and continuous improvement initiatives.

Contribute to disaster recovery and platform resilience strategies.

Full time Posting Date:
Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary