More jobs:
Observability Operations Engineer
Job in
Phoenix, Maricopa County, Arizona, 85003, USA
Listed on 2026-09-12
Listing for:
Tata Consultancy Services
Full Time
position Listed on 2026-09-12
Job specializations:
-
IT/Tech
SRE/Site Reliability, Unix/Linux
Job Description & How to Apply Below
- Strong Observability Administration experience with Dynatrace, Splunk, and Open Search/Elasticsearch.
- Hands-on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning.
- Strong knowledge of Linux, Kubernetes, and cloud environments.
- Experience with Grafana, Prometheus, Open Telemetry, and related observability technologies.
- Automation experience using Python/Shell scripting and REST APIs.
- Experience supporting enterprise-scale production environments, troubleshooting, and RCA.
- Administer and optimize enterprise Dynatrace, Splunk, and Open Search/Elasticsearch platforms.
- Maintain platform availability, scalability, performance, security, and reliability.
- Build and manage monitoring, logging, tracing, dashboards, alerts, and operational metrics.
- Troubleshoot production issues and perform root cause analysis using observability tools.
- Support Linux, Kubernetes, container, and cloud-based environments.
- Automate operational activities and drive self-healing and AI-assisted operations.
- Manage upgrades, patching, capacity planning, backups, and operational governance.
- Collaborate with SRE, Dev Ops, Platform, Infrastructure, and Application teams.
Observability Operations Engineer
Must Have Technical/Functional Skills- Strong Observability Administration experience with Dynatrace, Splunk, and Open Search/Elasticsearch.
- Hands‑on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning.
- Strong knowledge of Linux, Kubernetes, and cloud environments.
- Experience with Grafana, Prometheus, Open Telemetry, and related observability technologies.
- Automation experience using Python/Shell scripting and REST APIs.
- Experience supporting enterprise‑scale production environments, troubleshooting, and RCA.
- Administer and optimize enterprise Dynatrace, Splunk, and Open Search/Elasticsearch platforms.
- Maintain platform availability, scalability, performance, security, and reliability.
- Build and manage monitoring, logging, tracing, dashboards, alerts, and operational metrics.
- Troubleshoot production issues and perform root cause analysis using observability tools.
- Support Linux, Kubernetes, container, and cloud-based environments.
- Automate operational activities and drive self-healing and AI-assisted operations.
- Manage upgrades, patching, capacity planning, backups, and operational governance.
- Collaborate with SRE, Dev Ops, Platform, Infrastructure, and Application teams.
Generic Managerial Skills, If any
Good Communication and assertiveness, Team Player
Salary Range- $100,000-$120,000 a year
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×