×
Register Here to Apply for Jobs or Post Jobs. X

Site Reliability Engineering Manager

Job in Milan, Lombardy, Italy
Listing for: Altro
Full Time position
Listed on 2026-07-30
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, IT Project Manager
Salary/Wage Range or Industry Benchmark: 90000 - 130000 EUR Yearly EUR 90000.00 130000.00 YEAR
Job Description & How to Apply Below
About Core View
Core View is the global leader in Microsoft 365 (M365) tenant resilience, serving over 23 million users worldwide. We empower the world’s leading organizations to master the complexity of Microsoft M365. Through robust security and precise governance, we help ensure that our client’s environments stay cyber-resilient and productive, no matter how complex they are.

Our unified, cloud-native platform delivers powerful automation, rapid value, and end-to-end visibility across the entire M365 ecosystem. Backed by world-class support and a collaborative, innovative culture, Core View is a place where your ideas matter, and your work truly impacts global enterprises.

Job Summary
To support our growth, we are looking for an SRE Manager, in either UK or Italy. As our Site Reliability Engineering Manager, you will be responsible for building and leading our Site Reliability Engineering team from the ground up. In this role, you will define the team structure, establish SRE practices and processes, and actively contribute as a hands‑on engineer — especially in the early stages.

You will work closely with the Director of Cloud & IT Infrastructure and collaborate with software engineering, Dev Ops, and IT operations teams to ensure the reliability, scalability, and performance of our systems.

The ideal candidate has a strong background in IT operations or infrastructure engineering and has successfully led or coordinated technical teams. We are looking for someone who combines solid operational expertise with a natural inclination toward leadership, process building, and continuous improvement.

Job Responsibilities

Build the SRE team from scratch: define roles, participate in hiring, onboard and mentor engineers.

Define the SRE technical direction and operational roadmap, aligning reliability initiatives with business objectives and product priorities.

Establish SRE practices, processes, and culture within the organization, including on‑call rotations, incident management, and blameless post‑mortems, fostering a culture of psychological safety and continuous learning.

Define and track SLOs, SLIs, and error budgets in collaboration with engineering and product teams.

Act as a hands‑on contributor during the team ramp‑up phase, directly involved in designing and operating critical infrastructure on Azure.

Own the incident management process end‑to‑end: detection, response, escalation, resolution, and root cause analysis.

Drive automation initiatives — including AI‑assisted operations where applicable — to reduce toil and improve the operational efficiency of the team.

Design and oversee monitoring, alerting, and observability solutions to ensure full visibility across systems and services.

Own capacity planning and cloud cost governance, ensuring infrastructure scales efficiently while remaining cost‑effective.

Build and maintain a strong documentation culture: runbooks, operational procedures, incident playbooks, and architectural decisions.

Collaborate with software engineering teams to embed reliability and operational readiness into the development lifecycle.

Define and maintain disaster recovery plans, backup strategies, and business continuity procedures.

Ensure that security best practices and compliance requirements are applied consistently across infrastructure and operations.

Report on team performance, reliability metrics, and operational health to senior leadership.

Job Requirements

A minimum of 3+ years of experience in an SRE role.

A minimum of 1+ years of experience leading or coordinating a technical team, with demonstrated ability to hire, mentor, and develop engineers.

Proven experience in IT operations, infrastructure engineering.

Solid hands‑on experience with Azure cloud services and cloud‑based infrastructure management.

Strong understanding of IT operations best practices, including incident management, change management, and service continuity.

Experience with monitoring and observability tools (e.g., Prometheus, Grafana, Azure Monitor, ELK stack).

Familiarity with Infrastructure as Code (IaC) tools such as Terraform or Ansible.

Good knowledge of containerization and orchestration…
Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary