Observability SRE/SME
Listed on 2026-08-17
-
IT/Tech
SRE/Site Reliability, Systems Engineer, Cloud Computing: Infrastructure & Operations, IT Support
Observability SRE/SME
Location:
Dallas, TX or Jersey City, NJ (Hybrid, 3 days onsite)
Duration:
Contract to hire
Client: DTCC
Requirements:
Has been done/been a part of a Grafana/Observability implementation SRE experience (understand how operation works), well versed with Splunk, Grafana and implement telemetry with Grafana Versed in scripting and automation Dashboarding with Grafana Good understanding of Open Telemetry
Join a leading organization as a Grafana Technical SME, where your expertise will be pivotal in designing, implementing, and optimizing observability solutions. Bringing deep knowledge in Grafana, log tracing, and application performance monitoring, you will help elevate the company's monitoring infrastructure, ensuring robust system performance and reliability in a complex environment. This role offers the chance to work on impactful projects in a collaborative, innovative setting, blending onsite engagement with flexible work arrangements.
RequiredSkills
- Extensive experience with Grafana dashboard design and implementation
- Proven track record in application performance monitoring (APM) and observability tooling
- Hands-on expertise in log tracing, root cause analysis, and troubleshooting distributed applications
- Strong knowledge of cloud platform integrations, especially with AWS environments
- Proficiency in using tools such as Prometheus, Open Telemetry, Dynatrace, App Dynamics, Datadog, and Splunk
- Experience with infrastructure as code (Terraform, scripting with Python and shell) for deploying observability solutions
- Ability to configure SLO-based alerting and optimize observability stacks
- Experience with performance testing and load injection
- Familiarity with automation pipelines and CI/CD integrations for monitoring tools
- Knowledge of complex, regulated environments such as financial or government sectors
Education and Experience
Bachelor’s degree in a technical field or related discipline Prior hands-on roles focused on observability, site reliability engineering, or performance engineering in similar environments.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).