System Development Engineer II, AWS ADC Monitoring & Observability
Listed on 2026-07-21
-
IT/Tech
Cloud Computing: Infrastructure & Operations, AWS, Unix/Linux
Do you enjoy helping U.S. Intelligence Community and Defense agencies implement innovative cloud computing solutions? Would you like to do this using the latest cloud computing technologies while being part of the world's most advanced cloud infrastructure?
AWS Cloud Watch is seeking System Development engineers who are passionate about troubleshooting technical issues, automating manual processes, and continuously improving service efficiency to enhance customer experience. In this role, you will support Cloud Watch Monitoring and Cloud Watch Logs services in specialized security regions. You'll work in a dynamic environment with our software development teams to deliver reliable services to our customers.
AWS Cloud Watch provides customers actionable visibility into the health of their applications and systems. Key features include web services for submission and retrieval of measurements, a console for presentation, and the ability to notify and automate based on metrics. Teams in Cloud Watch solve problems of massive telemetry data scale, distributed systems, data visualization, and operational workflow.
Security RequirementThis position requires the candidate selected to currently possess and maintain an active TS/SCI security clearance. After start, the selected candidate must maintain an active TS/SCI security clearance with polygraph or commensurate clearance for each government agency for which they perform AWS work.
Key job responsibilities- Identify when implementations lack high availability or contain defects; investigate systemic patterns.
- Resolve, mitigate, or mitigate operational concerns in a timely manner.
- Support the operational stability of Cloud Watch and Cloud Watch Logs services in ADC environments.
- Work with AWS engineers to identify process automation opportunities.
- Perform full‑lifecycle software development with constant customer interaction to scope, design, implement, test, deliver, and operate software solutions.
- Learn and grow using a Dev Ops, Agile Scrum, incremental delivery philosophy with highly supportive peers constantly sharing subject matter expertise and producing peer‑reviewed code.
- Investigate the root cause of a customer issue impacting Cloud Watch service visibility.
- Analyze metric trends to identify performance degradation or capacity concerns.
- Collaborate with the team on automation opportunities for common operational tasks.
- Participate in on‑call rotations, responding to service alerts and customer issues.
- Dive deep into logs and metrics to understand system behavior and dependencies.
- Discuss and implement improvements to team operational processes.
Strong candidates will have experience with Linux command line and bash or Python scripting in large‑scale environments, and solid understanding of networking and distributed systems concepts. You should demonstrate enthusiasm for technical deep‑dives, strong communication skills, and the ability to prioritize and organize complex work. You are motivated to grow your technical skills and advance your career, enjoy automating tasks in distributed systems that operate at scale, and thrive in performance engineering and service scaling.
You are committed to delivering high‑quality, reliable, always‑on services to our customers.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).