Senior Platform Engineer
San Diego, San Diego County, California, 92101, USA
Listed on 2026-09-02
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Senior Platform Engineer
This position is not eligible for fully remote work. Candidates must be willing to work from one of Synopsys' U.S. office locations. Current hiring locations include Sunnyvale, CA;
Canonsburg, PA;
Hillsboro, OR;
Austin, TX;
Bellevue/Seattle, WA;
San Diego, CA; and Morrisville, NC. Applications are welcomed from candidates who can be based at, or are willing to relocate to, one of these locations.
Synopsys is the leader in engineering solutions from silicon to systems, enabling customers to rapidly innovate AI-powered products. We deliver industry-leading silicon design, IP, simulation and analysis solutions, and design services. We partner closely with our customers across a wide range of industries to maximize their R&D capability and productivity, powering innovation today that ignites the ingenuity of tomorrow.
You AreYou are the kind of engineer who sees beyond the ticket in front of you and instinctively looks for the system behind it. When something breaks, you don't just rush to restore service and move on—you stay curious about why it happened, what assumptions failed, and what would prevent the same issue from resurfacing. You're energized by complexity, not intimidated by it, and you approach ambiguity as a space to bring structure, clarity, and calm.
You thrive in environments where reliability and speed both matter, and you naturally look for the path that improves both rather than treating them as tradeoffs. You prefer repeatable mechanisms over heroic effort, and you feel a quiet satisfaction when a once-manual process becomes a dependable workflow that others can trust. You value clean handoffs, clear ownership, and thoughtful defaults—because you know that the best platforms are the ones that make the right thing the easy thing.
You communicate in a way that builds confidence: transparent about risk, crisp about decisions, and generous with context so others can move independently. You collaborate with empathy, but you're also willing to challenge decisions that create hidden operational or security debt. You take pride in leaving things better than you found them—more observable, more resilient, and easier to operate—so teams can focus on building what matters.
WhatYou'll Be Doing
- Design and maintain robust web applications and internal tools using Angular or React, Node.js/NestJS, and Python to support engineering workflows at scale.
- Architect and implement hybrid infrastructure solutions spanning on-prem data centers and public cloud environments (Azure, AWS, and GCP) with reliability and security as default outcomes.
- Build and evolve CI/CD pipelines in Azure Dev Ops, Git Hub Actions, or similar platforms to standardize builds, testing, releases, and deployment automation.
- Develop and manage Infrastructure as Code using Terraform to enable repeatable provisioning, eliminate drift, and improve auditability across environments.
- Deploy, operate, and optimize Kubernetes platforms (AKS, EKS, GKE, and self-managed clusters), including upgrades, capacity planning, and incident response.
- Automate provisioning and operational workflows with Python, Bash, or Power Shell to reduce manual toil and improve platform consistency.
- Implement observability and reliability practices—metrics, logs, traces, dashboards, and alerting—so issues are detected early and resolved with clear operational insight.
- Enable development teams to ship changes faster and more confidently because delivery workflows are automated, consistent, and resilient.
- Reduce operational incidents and time-to-recovery by improving platform reliability, observability, and repeatable remediation patterns.
- Turn infrastructure provisioning into a self-service, auditable capability that scales across hybrid and multi-cloud environments.
- Keep Kubernetes platforms stable and efficient across providers through better standardization, capacity management, and operational hygiene.
- Improve security and compliance outcomes by embedding governance, identity controls, and secure defaults directly into platform and delivery practices.
- Lower total cost of ownership by eliminating manual toil,…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).