More jobs:
Senior/Infrastructure – Platform Engineer
Job in
Santa Clara, Santa Clara County, California, 95053, USA
Listed on 2026-09-09
Listing for:
Jobtailor
Full Time
position Listed on 2026-09-09
Job specializations:
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, IT Infrastructure
Job Description & How to Apply Below
- Design, build, and operate scalable, reliable, and secure infrastructure across cloud, Kubernetes, on-premises, and hybrid environments
- Identify complex infrastructure and engineering productivity challenges and drive solutions from problem definition through architecture, implementation, and production
- Build and improve internal developer platforms, tools, and services
- Design, build, and optimize CI/CD platforms and pipelines
- Automate provisioning, configuration, deployment, testing, monitoring, and operational workflows
- Improve developer experience across the software lifecycle with engineering teams
- Develop reusable Infrastructure as Code, automation, and platform capabilities
- Build systems with reliability, observability, resilience, security, and disaster recovery capabilities
- Troubleshoot complex infrastructure and platform issues; participate in incident response and root-cause analysis
- Evaluate technologies and architectural approaches and establish infrastructure standards and best practices
- Lead technical initiatives across multiple engineering teams and contribute to infrastructure and developer productivity strategy
- 6+ years of experience in infrastructure, platform engineering, software engineering, SRE, Dev Ops, or related fields
- Strong programming and software engineering fundamentals with experience developing production-quality tools and services
- Deep production Kubernetes experience, including troubleshooting complex environments
- Experience with AWS, Azure, or GCP and/or on-premises infrastructure; self-managed or bare-metal
- Experience with Infrastructure as Code such as Terraform, Ansible, or similar technologies
- Experience designing and improving CI/CD platforms and deployment automation
- Strong understanding of Linux, networking, distributed systems, and infrastructure architecture
- Strong experience with server provisioning, configuration management, patching, upgrades, and lifecycle management
- Hands‑on experience with Ansible, Chef, or similar configuration management and automation technologies
- Strong programming skills in at least one language such as Go, Python, Rust, or C++
- Experience with observability, monitoring, logging, and reliability engineering practices
- Ability to lead complex technical initiatives, evaluate tradeoffs, and drive scalable, maintainable solutions through production
- Strong communication and collaboration skills across technical and cross‑functional teams
- Candidates must be legally authorized to work in the United States at the time of hire
- Candidates must have a minimum of 24 months of current U.S. work authorization remaining without employer sponsorship
- Nice-to-have qualifications include internal developer platforms, on-premises/private-cloud/bare-metal Kubernetes, Kubernetes operators/controllers, productivity or reliability metrics, Kubernetes CNI/CSI, distributed technologies, Jenkins, VPN, SSO, infrastructure security, network infrastructure, hardware/server provisioning, data center operations, and enterprise customer deployment experience
Demonstrates expertise in designing and operating scalable, reliable, and secure infrastructure across cloud and hybrid environments, with a strong focus on automation, CI/CD, and Infrastructure as Code. Proven ability to lead technical initiatives and improve developer experience through effective collaboration and problem-solving.
Highest-signal resume keywords- Kubernetes Expertise
- Infrastructure as Code (Terraform, Ansible)
- CI/CD Platform Design and Automation
- Cloud Infrastructure (AWS, Azure, GCP)
- Strong Programming Skills (Go, Python, Rust, C++)
Hard Skills
- Infrastructure Engineering
- Platform Engineering
- Software Engineering
- Dev Ops Practices
- Configuration Management
- Server Provisioning
- Monitoring and Observability
- Distributed Systems
- Networking
- Disaster Recovery
- Strong Communication
- Collaboration Skills
- Problem-Solving
- Infrastructure Standards
- Best Practices
- Incident Response
- Root-Cause Analysis
- Developer Platforms
- Private Cloud
- Bare-Metal
- Productivity Metrics
- Reliability Engineering
- Data Center Operations
- Kubernetes
- Terraform
- Ansible
- Chef
- Jenkins
- AWS
- Azure
- GCP
- Linux
- VPN
Position Requirements
10+ Years
work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×