Monitoring and Observability Engineer-Charlotte NC; Hybrid
Job in
Charlotte, Mecklenburg County, North Carolina, 28202, USA
Listed on 2026-07-09
Listing for:
Georgia IT, Inc.
Full Time
position Listed on 2026-07-09
Job specializations:
-
IT/Tech
IT Support, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Job Description & How to Apply Below
Monitoring and Observability Engineer
Location:
Charlotte NC – Hybrid 3 days a week
Duration: 1 year
Rate: DOE
US Citizens and Green cards are Preferred.
Required:- Experience with monitoring and onboarding cloud applications into Dynatrace
- Implement monitoring solutions for critical business applications to facilitate observability using Dynatrace open telemetry by collecting, processing, and analyzing data including logs, performance metrics, system events, and distributed traces, logging frameworks
- Perform necessary maintenance for .Net applications availability using Dynatrace Alerting to improve system scalability, reliability, and performance, optimizing infrastructure, implementing redundancy and failover mechanisms, and conducting load testing to ensure systems can handle expected traffic volumes
- Expertise in working with AWS components like EC2, RDS, ELB, Route
53, Lambda, ECS, WAF, Kafka and onboard Logging and services for these components into Dynatrace - Create custom level dashboards that helps in troubleshooting issues related with Networking, Disk IO and services degradation
- Work with the application team to facilitate monitoring solutions based on their requirements and created custom alerts, dashboard, rehydration process and reporting
- Identify key metrics and create Monitors, Custom Alerts using Tags for various applications
- Troubleshooting issues related to .Net Core AWS in various prod environments and project environments
- Provide 24x7 support for inter-application groups (Development, QA, System Admin, Support D.B.A.'s)
- Experience with incident management processes responding to incidents, conduct post-incident reviews (PIRs), and contribute to improving system reliability and resilience
- Analyze and troubleshoot the application performance issues to pinpoint the root cause for the slowness to avoid downtime.
- Proficiency in scripting languages such as Python, Bash, or Power Shell for automating routine tasks, troubleshooting issues, and building tool to improve operational efficiency.
- Work on deployment applications using CI/CD tools like Jenkins, Puppet and Chef in clustered environments.
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×