DevOps Engineer
Listed on 2026-09-13
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Azure
Location:
Dallas / Plano, TX - Locals only
Work Arrangement:
Hybrid – 3 days onsite per week
Experience: 5+ years preferred; candidates with strong relevant hands‑on experience are encouraged to apply.
Job Overview
We are looking for a hands‑on Dev Ops Engineer with strong experience in Azure cloud infrastructure, Kubernetes, Infrastructure as Code, CI/CD, observability, security, and production application operations. The ideal candidate will have experience supporting scalable cloud‑native and AI‑enabled workloads.
Key Responsibilities & Requirements
- Strong hands‑on experience with Microsoft Azure
, including AKS, Azure Functions, Container Apps/App Service, API Management, Key Vault, Storage, and core networking. - Experience with Terraform and/or Bicep/ARM for Infrastructure as Code, including reusable modules, state management, and environment promotion.
- Strong experience with Kubernetes/AKS
, including container deployments, Helm charts, ingress, resource tuning, horizontal/cluster autoscaling, and workload security. - Design and maintain multi‑stage CI/CD pipelines using Azure Dev Ops or Git Hub Actions, including automated testing, security scanning, artifact management, and blue‑green/canary deployments.
- Proficiency in Python for automation, operational scripting, and Dev Ops tooling.
- Experience with Azure Monitor, Application Insights, Log Analytics/KQL
, Prometheus, Grafana, and/or Open Telemetry. - Experience deploying and operating GenAI/agentic workloads in production, including Azure OpenAI or Azure AI Foundry, model/prompt versioning, evaluation pipelines, vector databases, and monitoring of latency, token usage, and inference costs.
- Experience supporting Databricks and MLflow
, with knowledge of Delta Lake, PostgreSQL, MongoDB, Kafka, and/or Azure Event Hubs. - Strong understanding of Microsoft Entra , managed identities, RBAC, secrets/certificate management, Dev Sec Ops scanning, and network isolation
. - Experience supporting large‑scale production systems, including capacity planning, high‑concurrency workloads, SLOs, incident response, scalability, and cost optimization
. - Familiarity with Lang Chain/Lang Graph, Semantic Kernel, or MCP and operational patterns for agentic/tool‑calling systems.
- Experience with Git Ops tools such as ArgoCD or Flux
. - Exposure to Istio, Linkerd, or similar service‑mesh technologies
, API gateways, and rate‑limiting patterns. - Strong troubleshooting, problem‑solving, communication, and ownership skills.
Preferred Qualifications
- Experience with Fin Ops
, cloud cost attribution, GPU optimization, and token/inference cost management. - Familiarity with Node.js, React, and Next.js build and deployment workflows.
- Telecommunications industry experience.
- Experience working in Agile environments with rapidly changing requirements.
Additional Information
Candidates should have genuine hands‑on experience with the technologies listed above and be prepared to discuss their experience and projects during the interview.
Resume Requirement: Please submit a concise 2–3 page resume highlighting relevant Azure, Kubernetes, Dev Ops, IaC, CI/CD, and production experience.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).