×
Register Here to Apply for Jobs or Post Jobs. X

Site Reliability Engineer​/L3 Support

Job in Mount Vernon, Westchester County, New York, 10550, USA
Listing for: SS&C Technologies, Inc.
Full Time position
Listed on 2026-07-31
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, IT Support
Salary/Wage Range or Industry Benchmark: 110000 - 120000 USD Yearly USD 110000.00 120000.00 YEAR
Job Description & How to Apply Below
## Site Reliability Engineer/L3 Support Apply locations:
Remote
- New York, US:
Remote
- Kansas, US:
Remote
- Pennsylvania US:
Remote
- New Hampshire, UStime type:
Full time posted on:
Posted Todayjob requisition :
R45217

As a leading financial services and healthcare technology company based on revenue, SS&C is headquartered in Windsor, Connecticut, and has 27,000+ employees in 35 countries. Some 20,000 financial services and healthcare organizations, from the world's largest companies to small and mid-market firms, rely on SS&C for expertise, scale, and technology.
** Job Description
*
* Job Title:

** Site Reliability Engineer (SRE) / L3 Support Engineer
** Getting to know us:

As a leading financial services and healthcare technology company based on revenue, SS&C is headquartered in Windsor, Connecticut, and has 27,000+ employees in 35 countries. Some 20,000 financial services and healthcare organizations, from the world's largest companies to small and mid-market firms, rely on SS&C for expertise, scale, and technology.

Kick off your software engineering career on our Quality & Automation team. You will learn modern test engineering practices while contributing real code, automated tests, and quality improvements. We welcome candidates new to finance
** Why You Will Love It Here!
*** Flexibility:
Hybrid Work Model & a Business Casual Dress Code, including jeans
* Your Future: 401k Matching Program, Professional Development Reimbursement
* Work/Life Balance:
Flexible Personal/Vacation Time Off, Sick Leave, Paid Holidays
* Your Wellbeing:
Medical, Dental, Vision, Employee Assistance Program, Parental Leave
* Wide Ranging Perspectives:
Committed to Celebrating the Variety of Backgrounds, Talents and Experiences of Our Employees
* Training:
Hands-On, Team-Customized, including SS&C University
* Extra Perks:
Discounts on fitness clubs, travel and more!
** What You Will Get To Do:
*** We are looking for a Site Reliability Engineer (SRE) to join our Platform Engineering team and take ownership of the operational health, reliability, and availability of our FedRAMP High cloud platform.
* This role combines modern Site Reliability Engineering practices with advanced production support responsibilities. You will act as the highest level of operational support (L3), proactively identifying and resolving issues before they impact customers, driving continuous improvement, and working closely with engineering teams to improve the reliability and operability of the platform.
* This is not a traditional operations role. You will use automation, observability, and engineering best practices to reduce operational toil while helping development teams build resilient, secure services.

Due to the nature of the environment, this position requires the successful candidate to be a
** U.S. Citizen
** and eligible to work on systems supporting
** FedRAMP High
** workloads.
** What you will get to do:
*** Monitor the health, availability, performance, and security of production services.
* Proactively identify emerging issues using telemetry, logs, metrics, and distributed tracing.
* Investigate, troubleshoot, and resolve complex production incidents across application and infrastructure layers.
* Act as the L3 escalation point for operational issues that cannot be resolved by L1 or L2 support.
* Participate in an on-call rotation for critical production incidents.
* Lead incident response activities, including coordination, communication, and post-incident reviews.
* Perform root cause analysis and ensure corrective actions are implemented to prevent recurrence.
* Develop and maintain operational runbooks, dashboards, alerts, and standard operating procedures.
* Improve platform observability by enhancing monitoring, alerting, dashboards, and service-level indicators.
* Work closely with software engineering teams to improve service reliability, scalability, and resilience.
* Identify opportunities to automate operational tasks and eliminate repetitive manual work.
* Support production deployments, infrastructure changes, and maintenance activities.
* Assist with disaster recovery exercises, resilience testing, and operational readiness reviews.
* Ensure operational activities comply with FedRAMP High security and compliance requirements.
* Contribute to continuous improvement initiatives across reliability, performance, and operational excellence.
** What you will Bring:
*** U.S. Citizenship (required).
* 3–6 years of experience in Site Reliability Engineering, Production Engineering, Dev Ops, Platform Engineering, or a senior production support role.
* Experience supporting mission-critical cloud-based production systems.
* Strong understanding of Linux operating systems and networking fundamentals.
* Experience troubleshooting distributed applications running in Kubernetes.
* Experience with public cloud platforms, preferably AWS.
* Experience with infrastructure as code and configuration management.
* Strong scripting or programming skills (e.g. Python, Bash, Power Shell, Go, or…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary