More jobs:
Vice President, DevOps Production Services
Job in
Pittsburgh, Allegheny County, Pennsylvania, 15201, USA
Listed on 2026-07-09
Listing for:
The Bank of New York Mellon
Full Time
position Listed on 2026-07-09
Job specializations:
-
IT/Tech
IT Support, SRE/Site Reliability, Systems Administrator
Job Description & How to Apply Below
Production Services Application Support
Job Schedule:
Full time
Posted Date: T12:59:57+00:00
Job Shift: Day
Base Salary Min: 69000
At BNY, our culture allows us to run our company better and enables employees' growth and success. As a leading global financial services company at the heart of the global financial system, we influence nearly 20% of the world's investible assets. Every day, our teams harness cutting-edge AI and breakthrough technologies to collaborate with clients, driving transformative solutions that redefine industries and uplift communities worldwide.
Recognized as a top destination for innovators, BNY is where bold ideas meet advanced technology and exceptional talent. Together, we power the future of finance - and this is what #LifeAtBNY is all about. Join us and be part of something extraordinary.
We're seeking a future team member for the role of Vice President, Dev Ops Production Services to join our Production Services Application Support team. This role is located in New York, NY, Pittsburgh, PA, Jersey City, NJ or Lake Mary, FL.
We are seeking a highly skilled professional with strong experience in Production Application Support to manage and support critical enterprise AI based applications in a fast-paced production environment. The role requires hands-on expertise in monitoring, incident management, troubleshooting, release support, and ensuring high availability and stability of business-critical platforms.
In this role, you'll make an impact in the following ways:
* Provide L2/L3 production support for enterprise applications and ensure platform stability, resiliency, and availability.
* Monitor application health, system performance, batch jobs, interfaces, and alerts using enterprise monitoring and observability tools.
* Investigate, troubleshoot, and resolve production incidents within defined SLAs.
* Perform root cause analysis (RCA) for recurring issues and drive permanent fixes.
* Analyze production logs, identify failure patterns, and create actionable dashboards to improve service monitoring and incident response.
* Coordinate with development, infrastructure, database, network, and business teams for issue resolution.
* Support application deployments, change requests, weekend releases, and post-release validations.
* Maintain incident, problem, and change records in service management tools.
* Drive continuous service improvement through automation, process optimization, and proactive monitoring.
* Participate in on-call support and major incident management as required.
* Prepare operational reports, service health summaries, and stakeholder communications.
* Write and analyze SQL queries for data validation, issue investigation, and production troubleshooting.
* Use Unix/Linux commands and scripting for application support, log reviews, file handling, and system-level troubleshooting.
* Leverage Splunk extensively for log analysis, issue diagnosis, trend identification, alerting insights, and dashboard creation.
To be successful in this role, we're seeking the following:
* Proven experience in production application support for business-critical applications.
* Strong understanding of incident management, problem management, and change management processes.
* Strong SQL skills for querying, troubleshooting, and data analysis in production environments.
* Extensive hands-on experience with Splunk for log analysis, search creation, troubleshooting, monitoring, and dashboard development.
* Strong Unix/Linux skills for navigating servers, reviewing logs, troubleshooting jobs/processes, and supporting application runtime environments.
* Experience with monitoring and alerting tools, log analysis, Grafana, and dashboard-based production support.
* Experience with ITSM tools such as Service Now, Jira, or similar platforms.
* Ability to analyze application, infrastructure, and integration issues across distributed systems.
* Experience supporting applications in cloud and/or on-prem environments.
* Familiarity with scripting and troubleshooting middleware/interfaces.
* Strong knowledge of release support, service recovery, and operational governance.
* Ability to work in a…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×