DevOps Engineer
Listed on 2026-09-12
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Systems Engineer
Enterprise Technologies is a central IT unit at Stanford University, responsible for delivering foundational computing and communication infrastructure that supports the University's academic, research, and administrative functions. We are seeking a Senior Dev Ops Engineer with proven expertise in API development, as well as administration and integration experience across Google Workspace, AWS, and Microsoft Azure. In this role, you will design and automate cloud infrastructure, optimize CI/CD pipelines, develop and maintain internal APIs, and manage cloud-based configurations, identity services, and security policies—leveraging scripting and API-driven automation across multi-cloud and enterprise collaboration platforms.
RESPONSIBILITIESThe Dev Ops Engineer will serve as a key member of the team, leading the design, development, planning, support, and security of Stanford’s infrastructure with a focus on Google Cloud technologies. This role is critical in optimizing software delivery pipelines, enabling collaboration across teams, and ensuring the reliability, performance, and scalability of systems running in Google Cloud. The ideal candidate will have deep experience with infrastructure automation, cloud-native tooling, and API integrations within the Google ecosystem—including GCP and Google Workspace.
Strong leadership, scripting, and communication skills are essential.
Key responsibilities include:
- Design, implement, and maintain scalable CI/CD pipelines (e.g., Git Hub Actions, Jenkins, Git Lab CI).
- Automate infrastructure provisioning using tools like Terraform, Cloud Formation, or Pulumi.
- Build, deploy, and manage containerized applications using Docker and orchestration platforms (Kubernetes, ECS, etc.).
- Develop and maintain RESTful and/or GraphQL APIs using Python, Go, or Node.js.
- Manage and integrate Google Workspace with internal systems, including API access, identity provisioning (via Directory API), and service automation.
- Monitor system health and performance using observability tools (e.g., Prometheus, Grafana, ELK, Datadog).
- Implement security best practices, including IAM, secrets management, and compliance enforcement.
- Collaborate with engineering, security, and IT teams to maintain a secure and efficient development environment.
- Support incident response, root cause analysis, and disaster recovery procedures.
- Document infrastructure, processes, and API endpoints thoroughly.
- Participating in 24x7 on-call support rotation.
Education & Experience
Bachelor's degree and eight years of relevant experience, or a combination of education and relevant experience.
Knowledge,Skills and Abilities
Knowledge, Skills and Abilities:
- Bachelor’s degree in Computer Science, Engineering, or equivalent work experience.
- 5+ years in Dev Ops, SRE, or Platform Engineering roles.
- Hands-on experience with public cloud platforms (AWS, Azure, or GCP—GCP preferred).
- Strong programming skills in one or more languages for API development (e.g., Python, Go, Node.js).
- Proficiency with CI/CD tools and IaC frameworks (e.g., Terraform, Ansible, Cloud Formation).
- Experience with Docker and container orchestration tools such as Kubernetes or ECS.
- Demonstrated experience managing and integrating Google Workspace in enterprise environments, including APIs, security settings, and automation.
- Solid knowledge of Linux administration, networking, and cloud security principles.
- Strong troubleshooting and debugging skills across the stack.
- Excellent communication, documentation, and collaboration abilities.
- Should demonstrate a willingness to learn and adapt to new technologies and industry trends.
- Experience building and maintaining internal…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).