More jobs:
Senior Specialist SysOps; DevOps
Job in
Abu Dhabi, UAE/Dubai
Listed on 2026-10-02
Listing for:
CPX Piceance
Full Time
position Listed on 2026-10-02
Job specializations:
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, IT Infrastructure, Systems Engineer
Job Description & How to Apply Below
Job Purpose:
The Senior Specialist - Sys Ops (Dev Ops) is responsible for managing, optimizing, and securing enterprise infrastructure, cloud platforms, Dev Ops tool chains, automation frameworks, and containerized environments. The role ensures reliability, scalability, performance, and operational excellence across internal and customer-hosted platforms while providing advanced technical support for incidents, service requests, and operational challenges.
Key Responsibilities:- Infrastructure and Operations
- Participate in the design, planning, and implementation of new projects and technologies, ensuring that solutions are highly available, secure, scalable, and performant.
- Provide day-to-day L2/L3 operational support, including incidents, service requests, problems, changes, and operational tasks.
- Monitor infrastructure performance, identify and resolve bottlenecks, troubleshoot outages, perform root cause analysis, and recommend improvements.
- Secure infrastructure by establishing and enforcing policies, defining and monitoring access, and supporting vulnerability and risk remediation.
- Maintain infrastructure health checks and proactively take action to minimize downtime and performance issues.
- Ensure services are backed up and recoverable in accordance with approved backup and recovery policies.
- Report operational infrastructure status, risks, progress, dependencies, and challenges to management.
- Actively participate in new projects, technology initiatives, and the onboarding of new customers and services.
- Design, deploy, administer, and optimize Microsoft Azure infrastructure and platform services.
- Support cloud migration, modernization, hybrid-cloud integration, governance, security, availability, and cost optimization initiatives.
- Implement resilient cloud-native architectures and standardized platform services that improve scalability and operational efficiency.
- Deploy, manage, upgrade, and troubleshoot containerized workloads using Kubernetes, Docker, and Helm.
- Implement Git Ops-based deployment and configuration management practices for consistent, auditable, and repeatable platform operations.
- Monitor and optimize cluster availability, performance, resource utilization, capacity, and security.
- Develop and maintain Infrastructure as Code using Terraform, Ansible, and Bicep.
- Create reusable modules, templates, playbooks, and automated provisioning workflows.
- Maintain version-controlled infrastructure definitions and configuration standards to ensure consistency across environments.
- Design, build, maintain, and optimize CI/CD pipelines using Jenkins, Git Hub, Git Lab, and Azure Dev Ops.
- Automate build, test, security validation, deployment, rollback, and release processes.
- Collaborate with development, security, and operations teams to embed security, compliance, and quality controls into delivery pipelines.
- Promote Dev Ops and Git Ops practices to improve deployment speed, consistency, traceability, and reliability.
- Develop and maintain automation scripts and configuration files using Bash, Power Shell, Python, and YAML.
- Automate repetitive infrastructure, deployment, monitoring, reporting, and service-management activities to reduce manual effort and operational risk.
- Integrate enterprise platforms, APIs, and workflows to streamline service delivery and operational processes.
- Implement and maintain monitoring, dashboards, alerting, and observability solutions using Prometheus, Grafana, Azure Monitor, and Log Analytics.
- Perform proactive monitoring, capacity planning, trend analysis, and performance optimization to maintain service stability and availability.
- Deploy and operationally support Azure AI services and AI-enabled solutions in accordance with security and governance requirements.
- Design and implement workflow automation that improves operational efficiency, response times, and service quality.
- Evaluate emerging AI, cloud-native, and automation technologies for practical adoption and service enhancement.
- Ensure compliance with approved ITSM policies, processes, procedures, and service-management guidelines.
- Write comprehensive technical reports, operational procedures, assessment findings, knowledge-base articles, known-error records, and…
Position Requirements
10+ Years
work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×