Site Reliability Engineer - Top Secret; req
Listed on 2026-07-23
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Systems Engineer
We are looking for a dynamic Site Reliability Engineer (SRE) with a Top Secret clearance to join our team. The Site Reliability Engineer (SRE) will manage, monitor, and optimize clusters on Kubernetes and accelerate our clients’ digital transformation through data-driven, scalable AI solutions.
Responsibilities- Monitor and Manage Kubernetes Clusters: Ensure the stability, health, and scalability of Kubernetes Clusters, deploying applications and services on Kubernetes
- Kubernetes Management: Deploy, monitor, and scale applications on Kubernetes clusters. Maintain Helm charts, manage services, and ensure resource allocation for optimal cluster performance
- Containerization & Deployment: Design and maintain Docker-based microservices architecture, ensuring consistent and reproducible deployments across staging, QA, and production environments
- Cloud Infrastructure Management: Work with leading Cloud Platforms (AWS, Azure and/or GCP) to set up, configure, and manage infrastructure resources using Infrastructure as Code (Terraform, Cloud Formation, etc.)
- Monitoring & Incident Response: Set up monitoring solutions, define alerts, an manage the incident response process for any issues related to Jenkins or Kubernetes clusters
- Automate Infrastructure Processes: Build automation tools for scaling, monitoring, and maintaining infrastructure using modern tools like Terraform, Ansible, Linux, or equivalent
- Collaborate Across Teams: Work closely with development, services, and operations teams to ensure a seamless integration between application development, deployment, and infrastructure
- Security & Compliance: Ensure all systems follow best practices in terms of security and compliance with relevant regulations. This includes role-based access, encryption, and automated vulnerability scanning
- Active TOP SECRET clearance or higher is required
- Bachelor’s degree in Computer Science or related field
- A minimum of two (2) years of experience working with on-premise and off-premise cloud environments
- Experience with AWS and/or Azure
- Hands-on experience with a range of open-source technologies, such as Linux, Docker, Kubernetes, K8s, Terraform, Helm, PostgreSQL, or similar technologies
- Ability to program (structured and OOP) using one or more high-level languages, such as Python, Java, C/C++, Ruby, and Java Script
- Experience with distributed storage technologies such as NFS, HDFS, Ceph, and Amazon S3, as well as dynamic resource management frameworks (Apache Mesos, Kubernetes, Yarn)
- Proactive approach to identifying problems, performance bottlenecks, and areas for improvement
- Ability to lead and work independently in an Agile/Scrum environment
- Real passion for developing team-oriented solutions to complex engineering problems
- Thrive in an autonomous, empowering and exciting environment
- Great verbal and written communication skills to collaborate multi-functionally and improve scalability
- Interest in committing to a fun, friendly, expansive, and intellectually stimulating environment
- Hands-on experience deploying and operating applications using IaaS and PaaS on major cloud providers, such as Amazon AWS, Microsoft Azure, or Google Cloud Services
- Experience with deep learning, natural language processing, computer vision, or reinforcement learning
- Conveys highly technical concepts and information in written form to technical and non-technical audiences
- The ability to work on multiple concurrent projects is essential. Strong self-motivation and the ability to work with minimal supervision
- Must be a team-oriented individual, energetic, result & delivery oriented, with a keen interest on quality and the ability to meet deadlines
CATHEXIS offers competitive compensation packages to all eligible employees. Our goal is to provide a compensation package that reflects the value you bring to our team, is competitive with national average market rates, and promotes your financial security and personal well-being. The annual salary range for this role is $100,000 - $160,000. Please note that the salary information provided is a general guideline.
CATHEXIS considers various factors in its final offer, including…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).