More jobs:
Senior DevOps Engineer – Storage Platforms
Job in
San Ramon, Contra Costa County, California, 94583, USA
Listed on 2026-07-21
Listing for:
Jobtailor
Full Time
position Listed on 2026-07-21
Job specializations:
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, IT Infrastructure, Systems Engineer
Job Description & How to Apply Below
- Deploy, automate, and operate large-scale Software Defined Storage architectures across private and public cloud regions within ITIL methodology.
- Deploy and support enterprise storage platforms (Pure Storage, HPE, Net App) and SDS solutions (Ceph, Longhorn).
- Integrate self-service storage workflows for Kubernetes CSI and Open Stack consumers (VM and Baremetal).
- Implement and manage backup solutions (preferably Rubik).
- Build and maintain Infrastructure-as-Code for storage platforms using Ansible, Terraform, Helm and Git, with Python/Bash automation.
- Implement CI/CD pipelines for infrastructure updates, patching, upgrades, testing, and rollback.
- Implement and improve monitoring, alerting, and observability for storage systems (capacity, latency, IOPS, recovery health) using Git Ops and tools such as Prometheus, Loki, and Grafana.
- Perform deep troubleshooting across storage, Kubernetes, hypervisors, networking, and Linux systems.
- Develop and maintain technical documentation, architecture diagrams, operational procedures, and runbooks.
- Participate in on-call rotations, incident response, and root cause analysis.
- Collaborate globally on change management, documentation, and operational best practices.
- 6+ years of experience managing enterprise storage and Kubernetes platforms on Linux.
- Strong hands-on experience with SDS solutions (Ceph, Longhorn) and storage migrations from legacy systems.
- Expertise with block, file, and object storage, including Fibre Channel (Cisco MDS) and IP-based protocols (NVMe-oF or iSCSI).
- Expert knowledge of Kubernetes and Linux systems (Ubuntu, RHEL/CentOS).
- Proficiency with Infrastructure-as-Code (IaC) (Ansible, Terraform) for provisioning storage and backup schedules.
- Expertise in backup technologies (preferably Rubik).
- Strong scripting skills in Python and Bash (Golang a plus).
- Experience operating 24x7 mission-critical production environments.
- Hands-on experience with KVM hypervisors (Suse Harvester, Open Stack).
- Strong written and verbal communication skills.
- Proficiency with Git, CI/CD pipelines, and automated testing frameworks.
- Ability to write technical documentation and contribute to community wikis or knowledge bases.
- Bachelor’s degree in computer science or equivalent professional experience.
Demonstrates expertise in managing enterprise storage solutions and Kubernetes platforms, with a strong focus on Infrastructure-as-Code practices and automation. Proficient in implementing and optimizing storage architectures across cloud environments while ensuring operational excellence and effective documentation.
#J-18808-LjbffrPosition Requirements
10+ Years
work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×