More jobs:
DevOps Engineer
Job in
Sacramento, Sacramento County, California, 95828, USA
Listed on 2026-08-03
Listing for:
California Surveying & Drafting Supply, Inc. (CSDS)
Full Time
position Listed on 2026-08-03
Job specializations:
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, IT Infrastructure
Job Description & How to Apply Below
We are seeking a highly skilled Dev Ops Engineer to design, build, automate, secure, and operate our Digital Products Platform. This role will be responsible for managing cloud infrastructure, CI/CD pipelines, infrastructure as code, platform reliability, monitoring, security automation, and operational excellence across multiple digital products and services.
The successful candidate will collaborate closely with Software Developers, Solution Architects, Product Owners, and Technology Leadership to enable rapid, secure, and reliable software delivery while maintaining platform stability and scalability.
Key Responsibilities- Design, implement, and maintain cloud infrastructure supporting enterprise digital products and services.
- Manage production and non-production environments ensuring high availability, resiliency, and scalability.
- Support cloud resource optimization, capacity planning, and cost management initiatives.
- Develop and maintain Infrastructure as Code using automation frameworks and declarative configuration tools.
- Standardize infrastructure deployment patterns and reusable platform components.
- Establish environment provisioning and configuration management best practices.
- Design, build, and maintain automated CI/CD pipelines.
- Implement automated testing, deployment, and release management processes.
- Support blue/green, rolling, and zero-downtime deployment strategies.
- Improve deployment frequency while maintaining security and reliability standards.
- Monitor platform health, availability, and performance.
- Develop automated alerting, logging, and operational dashboards.
- Support incident management, root cause analysis, and post-incident reviews.
- Drive continuous improvement of platform reliability and operational maturity.
- Implement Dev Sec Ops practices throughout the software delivery lifecycle.
- Manage vulnerability scanning, secrets management, access controls, and infrastructure security.
- Ensure platform compliance with organizational security standards and regulatory requirements.
- Implement enterprise monitoring, logging, tracing, and performance management solutions.
- Create operational metrics, SLAs, SLOs, and dashboards.
- Proactively identify and resolve performance bottlenecks.
- Work closely with development teams to improve deployment automation and platform adoption.
- Provide technical guidance on cloud-native architecture and operational best practices.
- Create technical documentation, standards, and operational runbooks.
- Mentor development and operational teams on Dev Ops practices.
- 5+ years of experience in Dev Ops, Site Reliability Engineering, Cloud Engineering, or Infrastructure Automation.
- Experience managing enterprise cloud environments.
- Strong experience implementing CI/CD pipelines and deployment automation.
- Experience with containerization and container orchestration platforms.
- Hands-on experience with Infrastructure as Code tools.
- Strong scripting and automation skills.
- Experience with monitoring, logging, and observability platforms.
- Understanding of networking, identity management, security, and cloud governance.
- Experience supporting production environments and incident response activities.
- Experience supporting digital product platforms and distributed microservices architectures.
- Experience with Azure cloud services.
- Experience with AWS cloud services and Amazon AI/ML services.
- Experience with Kubernetes and container-based platforms.
- Experience implementing Dev Sec Ops practices.
- Experience integrating enterprise applications, APIs, and data platforms.
- Experience with SaaS platform integration and enterprise authentication technologies.
- Knowledge of ITIL, Site Reliability Engineering (SRE), and operational excellence frameworks.
- Relevant cloud certifications (Azure, AWS, Kubernetes, Terraform, etc.).
- Improved deployment frequency and release reliability.
- Reduced incident volume and faster recovery times.
- Increased platform availability and performance.
- Enhanced security posture and compliance…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×