SRE Microservices
Job Description & How to Apply Below
Responsibilities
Design, build and maintain scalable and resilient microservices infrastructure focusing on automation and self‑healing capabilities. Implement and manage CI/CD pipelines for microservices, ensuring rapid and reliable deployments with robust rollback strategies. Proactively monitor microservices health, performance and security, establishing SLOs, SLIs and effective alerting mechanisms. Develop and maintain comprehensive disaster recovery and business continuity plans for critical microservices applications.
Qualifications- 5+ years of experience in SRE or Dev Ops roles, building and managing large-scale, high‑availability systems across banking, fintech, e‑commerce, or other data‑intensive digital ecosystems.
- Bachelor’s degree in Computer Science or equivalent technical experience.
- Strong experience with Linux environments and performance troubleshooting.
- Proven expertise in Terraform and Infrastructure as Code (IaC) methodologies.
- Proficiency with Kubernetes and container orchestration in microservices environments.
- Hands‑on experience with AWS (preferred); exposure to Azure or GCP is an advantage.
- Deep knowledge of Dynatrace (AIOps, Davis AI), Prometheus, Grafana and the ELK stack.
- Experience implementing AI/ML‑driven reliability or automation solutions (AIOps, anomaly detection, predictive alerting).
- Practical understanding of CI/CD pipelines (Git Hub Actions, Jenkins, Git Lab CI/CD or Azure Dev Ops).
- Experience with Kafka, Rabbit
MQ, Redis, Aurora and RDS databases. - Strong scripting or programming skills in Python, Bash or Go.
- Organized, structured and meticulous in approach.
- Experienced in cross‑functional collaboration and working with distributed teams.
- Strong analytical mindset with excellent troubleshooting skills for complex production systems.
- Calm and composed communicator under pressure, capable of leading during high‑impact incidents.
- Proactive problem‑solver who anticipates issues and drives preventive improvements.
- Passionate about AI‑driven automation, observability and reliability engineering.
- Continuously learning, keeping up‑to‑date with cloud‑native, microservices and SRE best practices.
- Collaborative and adaptable team player who thrives in a fast‑paced, regulated environment, and is passionate about building reliable, scalable systems that empower digital banking innovation.
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×