More jobs:
Site Reliability Engineer (SRE
Job in
Washington, District of Columbia, 20022, USA
Listed on 2026-08-06
Listing for:
Veriipro
Full Time
position Listed on 2026-08-06
Job specializations:
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, IT Support, Systems Engineer
Job Description & How to Apply Below
Roles and Responsibilities
- Implement and maintain observability solutions using Dynatrace, including dashboards, distributed tracing, telemetry, and anomaly detection.
- Design and manage CI/CD pipelines using Git Hub Actions, Jenkins, or AWS Code Pipeline.
- Automate cloud infrastructure provisioning using Terraform, Cloud Formation, or AWS CDK.
- Support production environments through incident response, troubleshooting, RCA, and ITIL-based processes.
- Define and monitor SRE metrics including SLIs, SLOs, and error budgets.
- Perform performance tuning, capacity planning, resiliency testing, and cloud cost optimization.
- Manage security configurations, access controls, service accounts, and certificates.
- Bachelor’s degree in Computer Science, Engineering, or a related technical field.
- 3+ years of experience in Site Reliability Engineering, Dev Ops, Cloud Engineering, or infrastructure-focused roles.
- Hands-on experience with cloud platforms such as AWS and Azure.
- Strong understanding of container technologies including Docker, Kubernetes, and Amazon ECS.
- Experience with Infrastructure-as-Code tools such as Terraform, Cloud Formation, or AWS CDK.
- Proficiency in Python or similar scripting languages for automation.
- Experience with configuration management tools such as Ansible.
- Strong knowledge of Linux systems administration, networking concepts, and troubleshooting.
- Familiarity with relational databases, cloud-native databases, and No
SQL technologies. - Experience with monitoring and observability tools, especially Dynatrace.
- Strong communication and collaboration skills with the ability to work independently.
- Ability to participate in on-call support rotations as needed.
- Experience implementing SRE best practices in enterprise environments.
- Knowledge of CI/CD, Git workflows, and cloud-native application architectures.
- Familiarity with ITIL processes and Service Now.
- Experience with security compliance frameworks and operational resiliency practices.
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×