More jobs:
AI Engineer
Job in
Grand Prairie, Dallas County, Texas, 75051, USA
Listed on 2026-08-22
Listing for:
AIT Global inc.
Full Time
position Listed on 2026-08-22
Job specializations:
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Systems Engineer, SRE/Site Reliability, Cybersecurity
Job Description & How to Apply Below
- Location:
Irving, TX - Employment Type:
Full-Time (W2 only) - Work Model: 5 days a week onsite
- Compensation:
Base salary range of $110,000-$140,000 per year, plus benefits - Experience Range:
Min 8+ Years and Max 25 Years - Work Authorization:
Who can work on a full-time W2 basis without sponsorship - This role is not open for C2C/C2H/1099 or any contract arrangements
- This opportunity is available for candidates who can work on a full-time W2 basis without sponsorship.
Must Have Technical/Functional Skills
We are seeking a highly skilled Senior Dev Ops Engineer with deep, hands‑on expertise in building and operating enterprise‑grade CI/CD platforms, cloud infrastructure, and developer tooling at global scale. In this role, you will own the design, automation, and reliability of the software delivery pipeline – enabling engineering squads to ship faster, safer, and with greater confidence. You will bring a strong platform engineering mindset, infrastructure‑as‑code discipline, and a proven track record of reducing toil, improving system resilience, and embedding security into every stage of the software lifecycle in a regulated financial services environment.
- Own the CI/CD Platform Design, build, and continuously improve enterprise CI/CD pipelines.
- Ensure pipelines are fast, reliable, secure, and scalable across dozens of engineering squads and technology stacks.
- Architect & Manage Cloud Infrastructure Provision, manage, and optimize cloud infrastructure across AWS, Azure, and GCP using infrastructure-as-code. Enforce cost controls, security guardrails, and operational best practices at scale.
- Drive Platform Reliability & SRE Practices Define and own SLOs, SLIs, and error budgets for platform services. Lead incident response, conduct blameless post‑mortems, and implement systemic fixes to eliminate recurring failures.
- Embed Security into the Delivery Pipeline Implement Dev Sec Ops practices – integrating SAST, DAST, SCA, container scanning, and secrets detection into every pipeline stage. Ensure compliance with enterprise security standards and regulatory requirements.
- Standardize Container & Kubernetes Operations Own the Kubernetes and Open Shift platform strategy – cluster lifecycle management, namespace governance, RBAC, networking policies, and workload autoscaling.
- Enable Developer Productivity Build internal developer platforms (IDP), self‑service tooling, golden path templates, and reusable infrastructure components that reduce friction and accelerate engineering teams.
- Lead Observability & Capacity Planning Design and maintain a unified observability stack covering metrics, logging, and distributed tracing. Drive proactive capacity planning and performance optimization across all environments.
- Champion Infrastructure as Code & Git Ops Establish IaC standards across the organization. Drive Git Ops adoption for all infrastructure and application configuration changes, ensuring auditability and repeatability.
- Mentor & Lead Technical Direction Define Dev Ops standards, conduct platform design reviews, mentor junior engineers, and collaborate with architects, security teams, and engineering leads to align platform strategy with business goals.
- CI/CD & Build Systems
- Jenkins
- Git Hub Actions
- ArgoCD
- Tekton
- CircleCI
- Maven / Gradle / npm build integration
- Release Automation & Versioning Strategies
- Docker
- Open Shift
- Operators & Custom Resource Definitions (CRDs)
- Pod Security Standards
- Cluster Lifecycle Management
- AWS (EC2, ECS, EKS, Lambda, RDS, S3, VPC, IAM, SQS, SNS, Route
53) - GCP (GKE, Cloud Run, IAM, VPC)
- Infrastructure as Code & Git Ops
- Terraform
- Ansible
- Cloud Formation
- Crossplane
- Packer
- Observability & Monitoring
- Grafana
- Open Telemetry
- ELK Stack
- Loki / Fluentd / Fluentbit
- Jaeger / Zipkin (Distributed Tracing)
- Pager Duty / Alert manager
- SLO / SLI / Error Budget Management
- Google Cloud Observability
- Storage
- S3
- NAS
- EFS
- EBS
- Dev Sec Ops & Security
- DAST Tools (OWASP ZAP, Burp Suite)
- Secrets Scanning
- Hashi Corp Vault
- mTLS & Certificate Management
- RBAC & Identity-Aware Proxy
- Networking & Service Delivery
- DNS & Load Balancing (Nginx, HAProxy, AWS ALB/NLB)
- CDN (Cloudflare, AWS Cloud Front)
- Ingress Controllers (Nginx Ingress,…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×