More jobs:
Technical Site Reliability Engineer
Job in
Abu Dhabi, UAE/Dubai
Listed on 2026-08-16
Listing for:
Linuxcareers
Full Time
position Listed on 2026-08-16
Job specializations:
-
IT/Tech
Systems Engineer, Unix/Linux, SRE/Site Reliability, Cloud Computing: Infrastructure & Operations
Job Description & How to Apply Below
Anduril Industries is a defense technology company building AI-powered military systems. As the founding Site Reliability Engineer for the Advanced Capabilities division, you will design, build, and operate infrastructure for next-generation wargaming simulation facilities that run massive-scale simulations of autonomous systems in contested environments.
What You'll Do- Maintain the simulation software stack including installation, configuration, updates, version management, and day-to-day functionality across simulation center tools and environments
- Own underlying infrastructure—compute, networking, storage, and environment configuration—keeping systems provisioned, patched, and performant
- Build and maintain post-release test suites with regression and smoke-test automation that runs after every software release to surface integration issues immediately
- Root-cause errors and failures in the system, drive them to permanent resolution, and implement guardrails, monitoring, or process changes to prevent recurrence
- Partner with development teams to review changes for reliability risk, surface concerns early, and implement mitigation strategies before releases
- Proficiency in Python for automation, tooling, and test development
- Working knowledge of C++ to read, debug, build, and trace issues in simulation codebases
- Solid general networking fundamentals including TCP/IP, UDP, multicast, DNS, routing, and firewall configuration with ability to diagnose latency, packet loss, and connectivity problems
- Experience maintaining production or production-adjacent systems and troubleshooting under time pressure
- Strong written and verbal communication skills for escalation, root cause explanation, and runbook documentation
- Eligibility to pass security and background check requirements for sensitive information systems
- Experience with modeling and simulation, wargaming, or distributed simulation standards such as DIS, HLA, TENA and platforms like AFSIM or VBS
- Test automation and CI/CD experience including building automated validation pipelines
- Infrastructure-as-code and configuration management tools such as Terraform, Ansible, Docker, and Kubernetes
- On-premises and cloud deployment experience
- Observability tooling such as Prometheus, Grafana, or ELK
- Linux systems administration depth with comfort in mixed Linux and Windows environments
- Prior work in defense, aerospace, or classified environments
- Active security clearance
Highly competitive equity grants are included in the majority of full‑time offers and are considered part of total compensation package. Top-tier benefits for full‑time employees including health and recovery support.
#J-18808-LjbffrTo View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×