Senior Cloud Operations Engineer/Team Lead
Listed on 2026-07-15
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Senior Cloud Operations Engineer / Team Lead
At Loftware,It’s all right there— the scale, the expertise, and the opportunity to grow your career in a business-critical industry.
Job Title:Senior Cloud Operations Engineer/Team Lead
Possible locations:
- Portsmouth, NH, USA (Hybrid),Remote (U.S.
-based candidates working EST hours
Please note:Visa sponsorship is not available for this role.
About the role: Loftware is expanding its worldwide 24x7 Cloud Operations Team and we are looking for a highly skilled Senior Cloud Operations Engineer / Team Lead with strong cloud-based Linux and Windows expertise. This role is both hands‑on and leadership‑focused, responsible for building, maintaining, and troubleshooting customer environments for mission‑critical applications across AWS and Azure. As a senior member of the team, you will also lead the US‑based Cloud Operations group, running the US daily stand‑ups to ensure alignment, focus, and operational excellence.
The
Senior Cloud Operations Engineer / Team Lead
will work closely with the global Cloud Operations team, and alongside QA and Development, to continually improve automated infrastructure and application deployment. You will help build and maintain reliable cloud infrastructure and services, ensuring the highly available and scalable solutions that Loftware customers rely on.
This is an excellent opportunity to take on technical leadership responsibilities while remaining deeply involved in cloud engineering, helping shape the evolution of our cloud platforms and mentoring the next generation of Cloud Operations engineers.
the base salary range for this position is $115,000 to $160,000 annually. This range represents the good‑faith minimum and maximum salary that the organization believes it would pay for this role at the time of this posting (these ranges are dependent on location of hire).
Lead Cloud Operations Excellence
- Own and enhance monitoring, observability, and alerting frameworks across AWS, Azure, and other platforms.
- Define SLIs/SLOs and drive continuous reliability improvements.
- Define and promote best practices across cloud operations.
- Lead the US daily stand‑ups, ensuring clear priorities, effective communication, and rapid issue resolution.
- Manage and mentor the US‑based Cloud Operations Engineers.
- Conduct regular 1:1s, performance reviews, and development planning.
- Foster a culture of accountability, collaboration, and technical excellence.
Drive Automation & Infrastructure as Code
- Architect and maintain scalable, reusable IaC solutions using Terraform, Terragrunt, and Ansible.
- Champion automation‑first approaches to eliminate manual operational tasks.
Security & Compliance Leadership
- Design and enforce cloud security best practices, governance, and compliance standards.
- Lead vulnerability assessments and remediation strategies across environments.
Reliability & Resilience Engineering
- Design and implement disaster recovery (DR), backup strategies, and high‑availability architectures.
- Lead incident response, root cause analysis (RCA), and postmortem reviews with actionable improvements.
Cross‑Functional Collaboration
- Partner closely with Engineering, QA, and Product teams to improve system architecture and application performance.
- Influence design decisions to enhance scalability, resilience, and operational efficiency.
- Design and manage complex networking setups including VPNs, Direct Connect, Transit Gateways, and hybrid connectivity.
Performance Optimization
- Identify bottlenecks and lead initiatives to optimize system performance and efficiency.
On‑Call & Incident Leadership
- Participate in and lead on‑call rotations.
- Act as escalation point for critical incidents and ensure rapid resolution.
Required Qualifications:
- 8+ years of relevant experience in cloud operations, SRE, or infrastructure engineering
- Deep experience with Linux and/or Windows server environments
- Proven experience leading operational initiatives in production environments
- Strong communication skills in English (written and verbal)
Preferred Technical
Experience:- Scripting:
Python, Java, Bash, .NET/C#, Powershell - IAC and Automation:
Terraform, Terragrunt,Ansible, Rundeck, Jenkins - Cloud-native…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).