Senior Systems Engineer - Delivery DevOps
Listed on 2026-08-22
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Systems Engineer, IT Infrastructure
Overview
At Intuit, we're building a culture that rewards curiosity, risk-taking, and imaginative thinking, and we look for engineers who fold AI and emerging tech into their daily work to solve real customer problems. Mailchimp is a leading marketing platform for small businesses, helping millions of customers around the world build their brands and grow their companies. Our Delivery team owns the systems that send our customers' email at massive scale, millions of messages a day, and keeps that infrastructure fast, reliable, and healthy.
OverviewAt Intuit, we're building a culture that rewards curiosity, risk-taking, and imaginative thinking, and we look for engineers who fold AI and emerging tech into their daily work to solve real customer problems. Mailchimp is a leading marketing platform for small businesses, helping millions of customers around the world build their brands and grow their companies. Our Delivery team owns the systems that send our customers' email at massive scale, millions of messages a day, and keeps that infrastructure fast, reliable, and healthy.
ResponsibilitiesSystems and Infrastructure Engineering
- Design, build, and operate the AWS cloud infrastructure the platform runs on, managing the sending fleet across AWS environments, including AMI and image build pipelines, instance and capacity management, autoscaling, and running the fleet cost-effectively.
- Develop and manage Infrastructure-as-Code (for example, Terraform, Ansible, or Puppet) to provision, configure, and maintain scalable, secure server and network environments.
- Design and manage system architecture across the sending path, including networking, load balancing, and security configuration, to keep production applications and services healthy.
- Design, implement, and maintain the CI/CD pipelines and release automation behind the platform (Jenkins, Artifactory and RPM packaging, containerized services, and Kubernetes/ArgoCD-based delivery), automating deployment workflows so releases to production are reliable, repeatable, and secure.
- Configure and maintain containerized environments (Docker, Kubernetes) for application deployment, scaling, and orchestration.
- Configure and tune MTA software and host-level parameters to manage throughput, queueing, and delivery behavior across the fleet.
- Take part in an on-call rotation, available to handle Delivery-related operational issues outside normal business hours, to keep the platform up and the business running.
- Monitor production systems and infrastructure for performance, availability, and security, and own the metrics, dashboards, and alerting that give the team and stakeholders visibility into fleet and sending health (for example, Big Query, Open Search, and Splunk).
- Respond to fleet, host, and delivery incidents, diagnosing issues that span the MTA software, the hosts, the AWS environment, and the network.
- Write scripts and automation tooling to streamline infrastructure provisioning, system configuration, and operational processes, and to support performance tuning across the fleet.
- Take full responsibility for delivering major platform features and projects, managing technical risk, watching them after launch, and keeping stakeholders informed.
- Review internal code and configuration changes for their impact on sending infrastructure and reliability.
- Build and maintain tooling that tracks sending health and IP and domain reputation, and that supports the team's response to receiver and deliverability trends.
- Apply working knowledge of email authentication and receiver behavior (SPF, DKIM, DMARC, feedback loops, and bounce handling) to keep the platform configured correctly and delivering well.
- Partner with Senior and Staff Engineers and Product Managers to analyze infrastructure and reliability requirements, weighing scalability, security, cost, and performance as you design solutions.
- Help improve deployment methods, monitoring tools, and infrastructure automation practices across the team.
- Look for opportunities to automate operations and improve the platform…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).