More jobs:
Site Reliability Engineer; Vietnamese speaker
Remote / Online - Candidates ideally in
UAE/Dubai
Listed on 2026-09-11
UAE/Dubai
Listing for:
AGAPI Technologies
Remote/Work from Home
position Listed on 2026-09-11
Job specializations:
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, IT Support
Job Description & How to Apply Below
Home / Job opening / Site Reliability Engineer (Vietnamese speaker)
Site Reliability Engineer (Vietnamese speaker)Dubai
Job Responsibilities:- Design, build, and maintain highly available, scalable, and resilient infrastructure.
- Define, implement, and monitor SLIs, SLOs, and SLAs to ensure service reliability.
- Develop and maintain monitoring, logging, tracing, and alerting systems.
- Respond to production incidents, perform root cause analysis (RCA), and drive postmortem improvements.
- Automate operational tasks, deployments, backups, and recovery processes using scripts and Infrastructure as Code (IaC).
- Optimize system performance, availability, scalability, and cost efficiency.
- Implement disaster recovery (DR), backup, and business continuity strategies.
- Perform capacity planning and infrastructure scaling based on traffic growth and business demands.
- Collaborate with development teams to improve application reliability, observability, and operational readiness.
- Define and enforce operational best practices, security standards, and reliability engineering principles.
- Continuously improve platform automation, self-healing capabilities, and operational efficiency.
- Evaluate and adopt new technologies, tools, and practices to enhance platform reliability and developer productivity.
- 4 years in an SRE, Dev Ops, or platform engineering role with production ownership at scale.
- Hands-on Kubernetes experience — deployment, scaling, networking, and troubleshooting in a production environment.
- Infrastructure-as-code fluency with Terraform (or Pulumi) across a major cloud provider (AWS preferred).
- Observability stack experience — Prometheus, Grafana, and/or equivalent tools with a track record of building meaningful dashboards and alerts.
- Proficiency in at least one scripting or programming language (Python, Go, or Bash) for automation and tooling.
- AWS certification (Solutions Architect, Dev Ops Engineer, or Sys Ops Administrator).
- Competitive Compensation: Enjoy a salary package tailored to your skills and experience, along with performance-based bonuses.
- Comprehensive Benefits: We support your well-being with accommodation, meal allowances, and assistance with work visa processing.
- Work-Life Balance: Unwind with generous holiday and New Year bonuses.
- Top-Tier Equipment: Stay productive with the latest tools, including a Mac Book and Iphone.
- Thriving Culture: Immerse yourself in a dynamic, inclusive work environment that fosters growth.
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×