×
Register Here to Apply for Jobs or Post Jobs. X

Senior DevOps Engineer

Job in Miami, Miami-Dade County, Florida, 33222, USA
Listing for: Socket.dev
Full Time position
Listed on 2026-09-24
Job specializations:
  • IT/Tech
    Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Systems Engineer, AWS
Salary/Wage Range or Industry Benchmark: 140000 - 190000 USD Yearly USD 140000.00 190000.00 YEAR
Job Description & How to Apply Below

Senior Dev Ops Engineer — Miami (Hybrid)

(Must be local to Miami area or commutable distance. If willing to relocate, please specify on your application.)

About

The Role

As a Senior Dev Ops Engineer you will be part of the Infrastructure Team responsible for Boats Group's internal developer platform. The platform provides self-service infrastructure for 500+ repositories across 13 international marketplace brands, running on AWS with EKS as the primary compute target.

You’ll own and evolve the platform that engineering teams use every day:
Kubernetes clusters, Git Ops delivery pipelines, CI/CD automation, observability, Cloudflare edge infrastructure, and Terraform/Atlantis-based infrastructure provisioning. You will also drive active migrations from legacy infrastructure (Jenkins, ECS, Cloud Formation) to the modern platform.

A successful candidate is someone who thinks in systems — you understand how a Helm chart, an ArgoCD Application Set, a Terraform module, and a Git Hub Actions workflow connect to deliver a service from PR to production. You’re comfortable operating at scale (200+ deployed services, multi-cluster EKS) and you know when to automate and when to simplify. Boats Group is a heavy AI adopter — LLMs and AI coding tools are part of everyday engineering work, and we expect you to bring real experience using them.

Every day is different: one day you’re debugging a Karpenter node scaling issue, the next you’re writing Terraform modules with Atlantis PR workflows, the next you’re migrating a Jenkins pipeline to Git Hub Actions. You won’t be bored.

What You’ll Do
  • Operate and evolve our multi-cluster EKS platform, ArgoCD Git Ops (200+ Application Sets), and Helm charts
  • Write Terraform modules and manage infrastructure changes via Atlantis PR workflows
  • Maintain and extend Git Hub Actions reusable workflows and composite actions
  • Manage the Grafana LGTM observability stack (Mimir, Loki, Tempo, Open Telemetry)
  • Manage Cloudflare edge infrastructure — 70+ domains, WAF, DNS, Workers, Bot Management
  • Migrate legacy services and pipelines from ECS/Jenkins/Cloud Formation to the modern platform
  • Use AI coding tools (Claude Code, MCP integrations) daily to accelerate infrastructure work
  • Improve deployment safety with progressive delivery (Argo Rollouts, Kargo)
  • Monitor and optimize AWS spending, security policies (Kyverno, RBAC), and IAM trust chains
  • Participate in on-call coverage for production outages
What You Must Have
  • Kubernetes — EKS operations, pod/node troubleshooting, RBAC, networking
  • Git Ops — ArgoCD or Flux; reconciliation model, sync failure debugging
  • IaC — Terraform/Terragrunt at scale;
    Atlantis PR-based workflows
  • CI/CD — Git Hub Actions (or similar) reusable workflows, not just consuming pipelines
  • AWS — Multi-account, IAM/OIDC, EKS, ECR, RDS, S3, SQS, VPC, Cloud Watch
  • Helm — Writing and maintaining production charts, values layering, template debugging
  • Observability — Prometheus/Grafana stack;
    Open Telemetry, Mimir, or Loki a plus
  • SRE
    practices — SLOs/SLIs, error budgets, incident response, blameless postmortems
  • AI
    /
    LLM fluency — Daily use of AI coding tools; prompt engineering, code generation, critical evaluation of LLM output
What You Should Have
  • K8s Ecosystem:
    Karpenter, Cilium, KEDA, Kyverno, External Secrets, cert-manager, Argo Rollouts/Kargo
  • IaC & CI/CD:
    Terraform/Terragrunt module authoring, Atlantis, Git Hub Actions (composite actions + reusable workflows), ArgoCD Application Sets, ECR image lifecycle
  • Migration:
    Jenkins - GHA, ECS - EKS, Cloud Formation - Terraform, monolith decomposition
  • Languages:

    Type Script/Node.js (primary ecosystem), Python, Go (bonus), Bash
  • AWS: EKS, ECR, Lambda, RDS/Aurora, Document DB, Elasticsearch, S3, SQS, Event Bridge, IAM/OIDC, Cognito, SSM, VPC, Route
    53, Cloud Front,…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary