Sr. Site Reliability Engineer/SWE
Listed on 2026-09-13
-
IT/Tech
IT Support, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Systems Engineer
Location: Ham
About Us
Visa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid.
About UsVisa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid.
At Visa, you'll have the opportunity to create impact at scale — tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world.
Join Visa and do work that matters – to you, to your community, and to the world. Progress starts with you.
Job DescriptionSite Reliability Engineering (SRE) is essential to Visa’s Cloud platform strategy. In this role, you’ll ensure our development platform and tools let engineers focus on innovation instead of infrastructure. You’ll promote observability best practices and automate resolution of recurring issues, working closely with software engineering teams to support security, availability, and performance. Responsibilities include triaging issues, collaborating on infrastructure management, and setting up monitoring for full coverage.
Hands‑on expertise is required, especially with major Dev Tools like Git Hub, Jenkins, Jira, and Artifactory.
We seek a Software Engineer + SRE hybrid engineer. The ideal candidate deeply understands at least one major Dev Tool, quickly resolves tool‑related issues in collaboration with developers, and applies systems thinking to maintain reliable applications and infrastructure while improving developer productivity.
Key Responsibilities- Dev Tools Support You will be the primary point of contact for developers using tools like Git Hub, Jenkins, Jira, or Artifactory.
- Troubleshoot and resolve tool‑related issues promptly to minimize developer downtime.
- Maintain and optimize CI‑CD pipelines and integrations for reliability and scalability.
- Collaborate with development teams to improve workflows and automation.
- Site Reliability Engineering Design, implement, and maintain systems for high availability, scalability, and performance.
- Monitor and improve application reliability through proactive measures and incident response.
- Develop and maintain observability solutions (metrics, logging, tracing).
- Participate in on‑call rotations and drive root cause analysis for incidents.
- Collaboration & Continuous Improvement Partner with engineering teams to identify reliability risks and implement best practices.
- Document processes, troubleshooting guides, and reliability of playbooks.
- Advocate for automation and self‑service solutions to reduce operational overhead.
This is a hybrid position. Expectation of days in the office will be confirmed by your Hiring Manager.
QualificationsQualifications
- Bachelor's degree, OR 3+ years of relevant work experience
- Bachelor's degree in IT, CS or related field and-or 3+ Years Working Experience IT Operations and Delivery.
- Experience:
3-8 years in SRE and-or Dev Tools support roles. - Beginner level programming and-or scripting in 2 or more of the following:
Python, Java, Go, Power Shell, JavaScript, Terraform, Ansible, Helm, Chef, Cloud Formation. - Basic understanding of YAML, JSON, HTML, XML.
- Hands on experience in Linux and
-or Windows systems and good understanding of distributed computing environments. - Experience with CI‑CD tooling such as Jenkins, Github, Bitbucket, ArgoCD, Artifactory, Azure Dev Ops in a large-scale environment
- Experience with observability tooling such as Grafana, Prometheus,…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).