Senior Site Reliability Engineer
Listed on 2026-08-05
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Systems Engineer, SRE/Site Reliability
Senior Engineer
At Grubhub, we believe food is more than just a meal:
It's a source of discovery, connection, and pure enjoyment. There's a time and place for every type of dish, from hidden neighborhood gems to tried-and-true favorites, and we exist to connect people with the food they love in all the ways they like to dig in. We've been at it since 2004, but now, as part of Wonder, Grubhub is operating with a renewed sense of momentum and the high-velocity energy of a powerhouse startup.
As a leading U.S. ordering and delivery marketplace, we feature over 415,000 merchants in more than 4,000 cities, creating the ultimate food experience by elevating online ordering through innovative restaurant technology, easy-to-use platforms, and an improved delivery experience. We are constantly finding new ways to innovate—from integrated grocery delivery with groceries powered by Instacart to exclusive loyalty programs. Join our team, based out of New York City, Chicago and Denver, and help us give our diners the exceptional value they deserve.
TheImpact You Will Make
- Guide Technical Direction and Innovation: help define architectural strategies, introducing modern platform engineering practices, and driving complex technical decisions across infrastructure and CI/CD platforms.
- Coach and Mentor Engineers:
Foster a culture of technical excellence by mentoring and coaching mid-level and junior engineers, conducting rigorous code and architecture reviews, and promoting continuous knowledge sharing. - Scale Infrastructure as Code (IaC):
Architect, maintain, and modernize declarative infrastructure across cloud environments using Terraform and/or Pulumi to drive reliability and self-service capabilities. - Orchestrate Global CI/CD Infrastructure:
Own, scale, and optimize high-throughput multi-cloud deployment pipelines using Jenkins and Spinnaker, ensuring fast, secure, and friction-free software delivery workflows. - Secure and Standardize Payment Environments:
Architect and harden our card payment infrastructure, ensuring the highest standards of data security, compliance, and proactive vulnerability mitigation. - Optimize Infrastructure and Container Life cycles:
Own and scale the automated pipeline for AWS AMIs and Docker containers, using configuration management tools to ensure consistent, secure, and reliable image deployment workflows. - Optimize Systems:
Troubleshoot and optimize mission-critical infrastructure at scale, leveraging deep Linux internals knowledge and kernel-level tuning to maximize efficiency and stability. - Advance Enterprise Logging and Observability:
Scale and maintain our data ingestion workflows and centralized logging platform using Vector, Datadog, and Splunk, building robust dashboards and advanced alerting mechanisms to minimize downtime. - Foster Operational Excellence:
Participate in the team's on-call rotation as an incident responder, leading postmortem discussions and writing automation scripts to eliminate repetitive operational toil.
- Technical Leadership and Mentorship:
Proven experience or strong aptitude for helping guide technical direction, making sound architectural choices, and successfully coaching/mentoring other engineers. - Infrastructure as Code (IaC) Expertise:
Hands-on familiarity and engineering experience with modern IaC frameworks, specifically Terraform and/or Pulumi, to manage complex cloud resources. - CI/CD and Deployment Expertise:
Robust, hands-on familiarity with configuring, scaling, and troubleshooting enterprise CI/CD systems, specifically Jenkins pipelines and Spinnaker deployments. - Linux Internals and Systems Engineering:
Deep, expert-level proficiency in Linux operating systems with proven experience in low-level system performance troubleshooting. - Automation and Advanced Scripting:
High proficiency in scripting and software automation utilizing Python and Bash to build internal tools, optimize processes, and eliminate technical overhead. - Core Cloud Expertise:
Strong engineering experience with AWS services (specifically EC2, Load Balancers, and AMI building) along with robust operational familiarity with Google Cloud Platform (GCP). - Container…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).