Senior Site Reliability Engineer; SRE
Listed on 2026-06-26
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Systems Engineer, IT Support, SRE/Site Reliability
Company Description
Tradeweb is a global leader in electronic trading across asset classes. As financial markets become increasingly interconnected, our technology enables efficient, multi-asset trading on a global scale. We serve more than 3,000 clients in more than 85 countries, including many of the world?s largest banks, asset managers, hedge funds, insurers, corporations, and wealth managers.
Creative collaboration and sharp client focus have helped fuel our organic growth. We facilitated average daily trading volume (ADV) of more than $2.8 trillion over the past four fiscal quarters, topping $3.3 trillion in ADV for the first quarter of 2026.
Since our IPO in 2019, Tradeweb has completed four acquisitions and doubled our revenues ? and 2025 was our 26th consecutive year of record revenues.
Tradeweb plays a central role in modernizing market structure by developing innovative trading protocols, embedding analytics into execution, and building technology infrastructure that supports the convergence of traditional and digitally native financial markets. Tradeweb is a great place to work, recognized in 2025 by Forbes as one of America?s Best Companies and by U.S. News & World Report as one of the Best Financial Services Companies to Work For.
Tradeweb Markets LLC ("Tradeweb") is proud to be an EEO Minorities/Females/Protected Veterans/Disabled/Affirmative Action Employer.
Group Details
As a Sr. Site Reliability Engineer (SRE) at ICD, you will play a critical role in ensuring the reliability and seamless operation of our global platform and AWS infrastructure to create scalable andhighly reliablesoftware systems.
Tradeweb Technology jobs are fully remote. The Tradeweb Technology hub is in our Jersey City office which can be used for team meetings and collaboration efforts. There may be days where travel to the Jersey City office is recommended for organizational off sites.
Job Responsibilities
- High performance Engineering Organization:contribute to an Agile-Agentic organization, planning, grooming, story ideation that will lead to iterative improvements of our platform.
- Security:
Prioritize security in all aspects of work, ensuring that it is the foundational consideration in every task performed. - IaCAutomation and Tooling:
Practice Git Sec Opsby contributing to the development and delivery ofa highly availableplatform through automation. Continually improve the reliability and efficiency of systems through iterative processes while reducing toil. - Reliability Engineering:
Work to ensure the reliability and availability of systems. Develop andmaintainmonitoring tools, analyze system performance, and implement solutions to improve overall system reliability. - Incident Triage and Resolution:
Triage issues, assess risk, and prioritize remediation with service teams.
Take full ownership and drive resolution of production, quality engineering and development-related infrastructure issues. - Communication and
Collaboration:
Effectively communicate issue statuses to both R&D and non-technical audiences. Ability to manage context switching when required. Collaborate closely with software development teams to influence architecture and design decisions thatimpactthe reliability and performance of systems. - Observability:
Develop observability tools to fulfill the needs of SLOs. Define and measure Service Level Objectives (SLOs) to ensure that the systems meet reliability standards. - On-call Responsibilities:
Fulfill regular on-call duties to enable high system availability.
Qualifications
- 6+ years of equivalent technology operations and engineering experience(ArgoCD,Kustomize,Pulumi, K8s, LGTM).
- 4+ years of scripting/coding experience in any modern language (Python Preferred).
- 4+ years as an SRE or similar individual-contributor role supporting public cloud (AWS) and cloud native technologies (Lambda, EKS, SNS, SMS, etc.).
- Bachelor?s Degree or higher in Computer Science or related field.
What does it take to be successful in this role?
- Cloud-based virtualizationexpertise, particularly with AWS native services.
- Strong multitasking skills in a dynamic environment.
- Proven ability to work independently with a proactive, task-ownership…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).