Staff Engineer II
Listed on 2026-07-12
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Network Engineer
Staff Engineer II
Location:
Block 23
The Staff Engineer II – Monitoring & Performance Engineering is responsible for the implementation, configuration, administration, and support of enterprise monitoring, observability, and performance management platforms. This role supports critical applications, infrastructure, cloud services, and network technologies to ensure system reliability, availability, and operational excellence. The engineer partners with Infrastructure, Network, Security, Cloud, Architecture, Application Development, and Operations teams to implement monitoring solutions, resolve complex technical issues, and improve service reliability across the enterprise.
- Design, implement, configure, administer, and support enterprise monitoring and observability platforms, including Datadog and Tripwire.
- Configure and maintain APM, infrastructure monitoring, log management, distributed tracing, alerting, dashboards, and reporting capabilities.
- Onboard applications, infrastructure, cloud services, middleware platforms, and network technologies into enterprise monitoring solutions.
- Develop and maintain monitoring standards, alerting strategies, dashboards, and operational reporting.
- Troubleshoot complex application, infrastructure, cloud, and network issues and perform root cause analysis.
- Collaborate across Infrastructure, Network, Security, Cloud, Architecture, Application Development, and Operations teams to resolve cross-functional technical issues.
- Create and maintain technical documentation, operational procedures, and support runbooks.
- Evaluate and recommend improvements to monitoring, observability, and performance management capabilities.
- Participate in technology modernization and continuous improvement initiatives.
- Participate in a 24x7 on-call rotation supporting enterprise monitoring platforms and production services.
- Respond to production incidents, support major incident management activities, and participate in after-hours maintenance and operational support as required.
What you'll need:
- Bachelor's degree in computer science, Engineering, Information Technology, or equivalent experience.
- 7+ years of experience in Systems Engineering, Infrastructure Engineering, Monitoring, Observability, Performance Engineering, or related technical disciplines.
- Hands-on experience designing, implementing, configuring, administering, and supporting enterprise monitoring and observability platforms, including Datadog and Tripwire.
- Experience onboarding applications, infrastructure, cloud services, middleware platforms, and network technologies into monitoring solutions. Experience supporting:
Application Performance Monitoring (APM), Infrastructure Monitoring, Distributed Tracing, Log Management, Alerting and Notifications, Dashboards and Operational Reporting - Strong troubleshooting and root cause analysis skills across infrastructure, applications, middleware, cloud platforms, and networks.
- Strong understanding of TCP/IP networking fundamentals, including DNS, HTTP/HTTPS, TLS/SSL, routing, switching, network connectivity, and latency analysis.
- Experience supporting enterprise production environments through 24x7 on-call rotations, incident response, and major incident management.
- Strong verbal, written, presentation, and interpersonal communication skills.
- Proven ability to work effectively across multiple technology teams and organizational boundaries.
Benefits you'll love:
We offer all the important things you'd want — like competitive salaries, an ownership stake in the company, medical and dental insurance, time off, a great 401k matching program, tuition assistance program, an employee volunteer program, and a wellness program. In addition, you'll have the opportunity to bolster your business knowledge, learning the ins and outs of how successful companies operate and manage their finances, giving you invaluable hands-on experience to help grow your career!
About the company:
Western Alliance Bank, Member FDIC, is a wholly owned subsidiary of Western Alliance Bancorporation. Serving clients nationwide, Western Alliance Bank includes six legacy bank brands — Alliance Association Bank, Alliance Bank of Arizona,…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).