Senior DevOps Engineer
Listed on 2026-07-15
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, IT Infrastructure
Senior Dev Ops Engineer
Hybrid role with an expectation to work on average 2 days per week from an HPE office.
Who We AreHewlett Packard Enterprise is the global edge‑to‑cloud company advancing the way people live and work. We help companies connect, protect, analyze, and act on their data and applications wherever they live, from edge to cloud, so they can turn insights into outcomes at the speed required to thrive in today’s complex world. Our culture thrives on finding new and better ways to accelerate what’s next.
We value diverse backgrounds and succeed together. We provide flexibility to manage work and personal needs, and we make bold moves together as a force for good. If you are looking to stretch and grow your career, our culture will embrace you. Open up opportunities with HPE.
We are seeking a highly skilled Senior Dev Ops Engineer with deep expertise in Linux systems, Kubernetes platforms, virtualization technologies, automation, and modern Dev Ops practices. In this role, you will design, deploy, operate, troubleshoot, and automate enterprise‑scale infrastructure and platform services that support the development and delivery of Morpheus and VM Essentials products. You will partner closely with software engineering, architecture, QA, and operations teams to deliver scalable, secure, highly available, and automated platforms.
This position requires a hands‑on technical leader who thrives in a fast‑paced environment and is passionate about improving operational excellence through automation, observability, and cloud‑native technologies.
- Design, deploy, maintain, and troubleshoot Kubernetes platforms and cloud‑native infrastructure.
- Build, automate, and enhance infrastructure and platform services using technologies such as Ansible, Terraform, Go, shell scripting, and related automation frameworks.
- Develop and maintain CI/CD pipelines using Jenkins, Git Hub Actions, or similar tools.
- Administer and support Linux and Windows environments while ensuring reliability, scalability, security, and performance.
- Manage virtualization environments leveraging KVM, VMware, Hyper‑V, and associated technologies.
- Collaborate with software development, QA, architecture, Dev Ops, and operations teams to deliver robust platform solutions.
- Participate in architecture reviews and contribute to technology strategy, platform standards, and engineering best practices.
- Enhance observability through monitoring, logging, alerting, and operational readiness initiatives.
- Implement security best practices and ensure compliance with organizational policies and standards.
- Support software releases, platform upgrades, patching activities, and lifecycle management processes.
- Create and maintain technical documentation, operational runbooks, knowledge base articles, and architecture diagrams.
- Continuously evaluate emerging technologies and recommend improvements to platform architecture, automation, and operational efficiency.
- Troubleshoot and resolve complex infrastructure, platform, and deployment issues across distributed environments.
- Effectively manage multiple priorities while delivering high‑quality, reliable solutions within agreed timelines.
- Bachelor’s or Master’s degree in Computer Science, Engineering, Information Systems, or a related technical discipline.
- 7+ years of professional experience in Dev Ops, Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, or related fields.
- Strong Linux system administration, performance tuning, and troubleshooting skills.
- Solid understanding of servers, operating systems, storage, and infrastructure architecture.
- Experience operating highly available and distributed systems in enterprise environments.
- Hands‑on experience designing, deploying, and operating Kubernetes platforms.
- Strong expertise in container orchestration, deployment strategies, monitoring, and troubleshooting.
- Understanding of cloud‑native architecture patterns and platform operations.
- Strong scripting and…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).