Senior Performance Engineering and Observability Lead; Remote
Columbus, Franklin County, Ohio, 43224, USA
Listed on 2026-07-13
-
Software Development
DevOps, AI Reliability/ Performance Engineer
We are seeking a Performance Engineering & Observability Lead who elevates how teams think about quality across the enterprise. The focus is on shaping test strategy, guiding performance and automation, strengthening CI/CD quality signals, and enabling teams to build resilient, performant features from the start. Success is measured by improved reliability, faster feedback, and confident delivery at scale.
Your influence goes beyond writing tests—you’ll elevate how multiple teams across our ecosystem think about resilience, performance, and customer impact. You’ll help build smarter test strategies, strengthen service‑level automation, and establish the patterns, tooling, and guardrails that keep our distributed architecture dependable. You’ll collaborate with engineers, product leaders, and platform teams to ensure our teams ship with confidence and perform flawlessly under real‑world conditions.
You’ll begin by building a strong understanding of our service landscape, React front end, customer journeys, and operational rhythms. From there, you’ll lead performance initiatives across teams—developing k6-based performance tests (JavaScript), Playwright-driven synthetic journeys, meaningful performance budgets, and the observability practices that enable quick diagnosis and confident releases.
This isn’t a “performance team tests at the end” role. Delivery teams own quality and performance.
Your role is to enable, guide, and elevate
—making performance a natural part of how teams build software every day.
Your mission is simple: protect the customer experience by helping teams consistently deliver fast, stable, and peak-ready services and web experiences.
What You’ll Influence- Quality mindset embedded early in discovery and design
- Performance‑Driven Engineering Practices - stability, and readiness
- Automation strategy across UI, API, mobile, and integrations
- Scalable test patterns, tooling, and guardrails across squads
- Quality gates and feedback loops within CI/CD
- Drive early conversations around latency expectations, scalability assumptions, failure modes, and peak scenarios
- Help teams translate customer experience goals into clear performance targets, including API latency budgets, journey-level budgets, and Core Web Vitals thresholds
- Identify and communicate risks across service dependencies, caching behavior, data access patterns, and front‑end rendering paths
- Guide performance strategy execution across microservices, APIs, event flows, and React user journeys
- Help teams adopt consistent approaches to workload modeling, test execution, and result interpretation
- Design practical, efficient test strategies for high-confidence delivery
- Expand Performance & automation where it delivers leverage and reduces manual effort
- Promote reusable patterns that scale quality across teams
- Performance evaluation at every stage
- Lead readiness for major releases and seasonal traffic spikes
- Forecast risks and validate resilience under load
- Define and maintain NFRs (SLIs/SLOs, latency targets, throughput goals)
- Model workloads, concurrency levels, and peak-event projections
- Partner with architects to validate scalability and resilience patterns early
- Build and execute performance, load, stress, and endurance tests
- Design realistic cross-service performance scenarios
- Validate caching, queuing, retry, and rate‑limit behavior under load
- Ensure services and web journeys are observable using Dynatrace for distributed tracing and service metrics, Splunk for log analysis and correlation, Full Story for experience insights and session replay
- Help teams connect performance metrics to customer outcomes (latency, errors, drop-offs)
- Lead or support investigation of latency regressions and performance incidents
- Implement automated performance smoke checks and regression triggers
- Integrate performance baselines and fail‑fast rules into pipelines
- Bui…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).