QA Engineer – Kubernetes Platform Validation; Azure Local & GCP FedRAMP
Listed on 2026-07-29
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations
- Full-time
Engineering the AI-powered enterprise. With AI and cloud-native solutions, BETSOL accelerates cloud transformation for enterprises across 17+ countries. BETSOL holds several engineering patents, and is recognized with industry awards. BETSOL maintains a net promoter score that is 2x the industry average.
BETSOL’s open source backup and recovery product line, Zmanda (), delivers up to 50% savings in total cost of ownership (TCO) and delivers best-in-class performance.
BETSOL Global IT Services () builds and supports end-to-end enterprise solutions, reducing time-to-market for customers.
BETSOL offices are set against the vibrant backdrops of Broomfield, Colorado and Bangalore & Belagavi, India.
We take pride in being an employee-centric organization, offering comprehensive benefits and opportunities.
You will join the same technical pod that builds and operates our Kubernetes-based cloud platforms across Azure and GCP, working alongside our Dev Ops/Dev Sec Ops engineers as a peer rather than a downstream verifier. Your focus is end-to-end solution validation of two platforms in parallel — Azure Local (Microsoft's hybrid on-prem stack, formerly Azure Stack HCI) and our FedRAMP-authorized GCP deployment — from initial deployment through Day-2 operations, upgrades, and decommission.
You'll design and run tests across the full "ilities" spectrum, break things on purpose to prove resilience, and read Terraform/Ansible/Helm and observability data fluently enough to troubleshoot from platform down to code alongside engineering.
- Drive end-to-end validation for two Kubernetes-based platforms in parallel — Azure Local and the FedRAMP-authorized GCP deployment — covering initial deployment, Day-2 operations, upgrades, and decommission.
- Design, execute, and document solution-level test cases across reliability, availability, scalability, security, observability, performance, recoverability, and up gradability, written in Given/When/Then style that's readable by engineering, PM, SRE, and operations.
- Test and troubleshoot hands-on across kubectl, Helm, operators, CRDs, PVCs, Stateful Sets, ArgoCD or Flux, network policies, and node affinity — able to debug from a pod down to the underlying node without a hand-hold.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).