AI Ops Engineer UAE
Listed on 2026-09-29
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, AI Engineer (Applied/Software)
Position Summary
Dicetek is seeking an experienced, highly technical, and accomplished AI Ops Engineer to join our client banking project team on a full-time basis in Abu Dhabi, UAE. In this specialized artificial intelligence operations and production engineering role, you will spearhead operating cloud-native, AI-driven, and high-scale API platforms across enterprise banking systems. You will work closely with cross-functional software engineers, data science squads, and IT infrastructure teams to manage CI/CD automation, LLMOps pipelines, observability telemetry, and AI Fin Ops cost optimization.
Ideal candidates bring a strong academic background in computer science or engineering, deep practical mastery of cloud-native AI operations, and immediate availability for employment in the UAE with a maximum 45-day notice period.
Job Description
As an AI Ops Engineer at Dicetek in Abu Dhabi, UAE, you will take full ownership of ensuring the reliability, scalability, security, and performance of mission-critical enterprise AI and LLM platforms. Your day-to-day responsibilities encompass managing Git Hub Actions and CI/CD pipelines, enforcing release governance and LLMOps lifecycle management, executing progressive delivery (canary releases, rollbacks), and maintaining comprehensive observability telemetry and operational runbooks.
You will drive AI Fin Ops initiatives by monitoring token and model usage, managing platform capacity and quotas, and optimizing cloud infrastructure costs. Working in a fast-paced banking environment with urgent multiple openings requiring UAE residency, you will establish reusable release standards, operational playbooks, and self-service workflows to empower engineering teams.
- Operate, scale, and maintain cloud-native, AI-driven, and high-scale API platforms in enterprise banking environments.
- Manage and enhance robust CI/CD pipelines, automated workflows, and environment management via Git Hub Actions or equivalent tools.
- Administer LLMOps frameworks, AI lifecycle management, version control, and production release governance.
- Design and execute progressive delivery strategies, controlled rollouts, canary releases, and automated rollback readiness.
- Implement comprehensive observability telemetry, monitoring dashboards, alert systems, and production-readiness practices.
- Drive AI Fin Ops initiatives, monitoring token and model usage, platform capacity, quota allocation, and cost optimization.
- Establish and standardize reusable release policies, operational playbooks, and self-service workflows across engineering teams.
- Collaborate closely with software development squads, data scientists, and infrastructure engineers in Abu Dhabi.
- Troubleshoot complex production incidents, perform root-cause analysis (RCA), and implement preventive engineering measures.
- Maintain comprehensive technical documentation, incident runbooks, architecture topologies, and audit logs.
- Bachelor’s degree in Computer Science, Information Technology, Software Engineering, or a related technical field.
- Minimum 4+ years of professional engineering experience with cloud-native, artificial intelligence, or high-scale API platforms.
- Hands-on professional experience with CI/CD automation pipelines and Git Hub Actions or equivalent tools.
- Practical working knowledge of LLMOps, AI model lifecycle management, and enterprise software release governance.
- Proven experience implementing progressive delivery techniques, canary deployments, and zero-downtime rollouts.
- Strong expertise in observability, telemetry, monitoring tools, log aggregation, and production-readiness practices.
- Solid exposure to AI Fin Ops, token and model usage tracking, resource…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).