Senior DevOps Engineer - Atlanta, GA
Listed on 2026-07-02
-
IT/Tech
Azure, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Systems Engineer
At Cortland, we operate with a forward‑thinking approach that challenges conventional norms and actively seeks insights beyond traditional industry boundaries. As a recognized leader in the multifamily sector, our focus on performance, innovation, and disciplined execution continues to drive strong growth and market leadership. We are committed to building a best‑in‑class organization by empowering top talent with the resources, autonomy, and support needed to deliver results and advance their careers in a high‑performance environment.
Role OverviewAs a Senior Dev Ops Engineer, you’ll design, build, and manage Azure‑based cloud platforms that support secure, reliable application delivery. You’ll lead CI/CD enablement, infrastructure automation, and observability efforts, while partnering across technology teams to improve performance, reliability, and security as Cortland’s technology footprint continues to grow. In this role, you’ll play a key part in shaping platform standards and operational practices that enable teams to build, test, and deploy at scale.
Responsibilities- Design, deploy, and operate secure, scalable cloud infrastructure on Microsoft Azure supporting application and data platforms.
- Automate infrastructure provisioning and configuration using Infrastructure as Code (Terraform and Azure Bicep) to enable consistent, repeatable deployments.
- Continuously improve platform reliability, performance, and scalability through automation and optimization.
- Enable development teams through standardized CI/CD pipelines, reusable templates, and tooling support using Git Hub and Azure Dev Ops.
- Partner with engineering teams to improve build, test, and deployment workflows while promoting Dev Ops best practices.
- Build and support a centralized monitoring and observability platform covering application availability and performance, infrastructure and network health, alerting, and Service Now integrations.
- Participate in an on‑call rotation and respond to incidents and outages, collaborating with application teams to diagnose and resolve complex production issues.
- Design and maintain high‑availability architectures, backup strategies, and disaster recovery solutions to support business continuity.
- Implement and enforce cloud security and governance best practices, including IAM/RBAC, Azure Policy, firewall rules, and secure network architectures.
- Collaborate closely with Software Engineering, Data Engineering, Data Analytics, and Technology Operations teams to deliver reliable, secure cloud platform services.
- Bachelor’s degree in Computer Science, Information Systems, or a related field preferred.
- 5+ years of experience in Dev Ops, Cloud Engineering, or Site Reliability Engineering, with deep expertise in Microsoft Azure.
- Proven experience designing, deploying, and operating production‑grade Azure infrastructure.
- Microsoft Azure certifications (AZ‑104, AZ‑204, AZ‑305, AZ‑400) are preferred.
- Familiarity with Microsoft’s Well‑Architected Framework and Cloud Adoption Framework is preferred.
- Experience with AI‑assisted development tools such as Git Hub Copilot, Codex, or Claude Code is preferred.
- Extensive hands‑on experience with Infrastructure as Code using Terraform and Azure Bicep.
- Proficiency in scripting and automation using Power Shell and Python.
- Experience with CI/CD platforms and tooling, including Git, Git Hub, and Azure Dev Ops.
- Strong experience with Azure governance and security tools, including Azure Policy, Azure Privileged Identity Management (PIM), and Entra (Azure Active Directory).
- Experience with application and infrastructure monitoring tools such as Datadog, New Relic, Azure Monitor, Prometheus, Grafana, or similar platforms.
- Experience deploying and supporting Windows Server‑based workloads, including Active Directory and Microsoft SQL Server.
- Excellent problem‑solving and troubleshooting skills with the ability to resolve complex infrastructure and inter‑related systems issues.
- Strong communication and collaboration skills, with experience working across teams and providing platform support to multiple…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).