T2/T3 Kubernetes/ DevOps Engineer for Sovereign Cloud Onsite / ApeiroRA / EU AI Projects (m/f/d
Verfasst am 2026-08-20
-
IT/Informationstechnik
Site Reliability Ingenieur/in, Cloud Computing: IT-Infrastruktur & Betrieb, IT Infrastruktur, Systemingenieur
T2/T3 Kubernetes/ Dev Ops Engineer for Sovereign Cloud Onsite / ApeiroRA / EU AI Projects (m/f/d) Location requirement
This position requires the candidate to be physically present in the Berlin, Garching, Dresden or St. Leon Rot office. Please ensure you meet this location requirement before applying. SAP supports relocation for this position and will assist successful candidates with the moving process.
What you’ll doBuild enterprise cloud infrastructure that provides European data sovereignty and hyperscaler‑grade capabilities. You’ll work on SAP Cloud Infrastructure, solving complex distributed systems challenges at scale: multi‑region networking, container orchestration, storage systems, and the APIs that connect them.
You will contribute to building a sovereign, privacy‑first AI cloud for the European market that makes it easy to securely train, fine‑tune, and deploy foundation and domain models with strict data residency and isolation. You’ll develop solutions using Go, Open Stack, and Kubernetes, tackling problems such as auto‑scaling thousands of containers across regions, building resilient storage systems, and designing APIs that handle massive traffic spikes.
Your work will power SAP’s production systems and thousands of customer environments, enabling organizations to run mission‑critical applications with the performance and reliability they expect from leading cloud platforms.
In your role as a Dev Ops Engineer you’ll build and maintain the infrastructure automation that powers SAP Cloud Infrastructure, implementing CI/CD pipelines and deployment automation for enterprise cloud services. You’ll ensure reliable delivery of distributed systems components that handle massive scale across multiple regions, using Kubernetes, Terraform, and cloud‑native tooling to automate the deployment and operation of container orchestration platforms, networking systems, and storage solutions.
Your focus will be on creating robust automation that supports multi‑region deployments, handles traffic spikes gracefully, and maintains the high‑availability standards expected by enterprise customers.
Kubernetes Platform Mastery: at least one year of daily hands‑on experience working with Kubernetes in production environments. You have created and maintained your own Helm charts or Kustomize configurations, managed cluster life cycles (upgrades, scaling, troubleshooting), and deployed shared platform components. Experience with multi‑cluster operations, ingress, service meshes, and cert‑manager. CKA, CKAD, or CKS certification are also a plus.
- Advanced Cloud & Infrastructure Expertise:
Mastery of cloud platforms (AWS, Azure, GCP) and advanced Kubernetes management, including multi‑cluster operations, scaling, and monitoring - Automation Proficiency:
Experience with Infrastructure as Code tools (e.g., Terraform, Ansible) and designing end‑to‑end CI/CD pipelines - Infrastructure Mastery:
Deep experience with Terraform, Pulumi, and building self‑service developer platforms at hyperscaler grade - Advanced Scripting Capabilities:
Proficient in Go, Python or Bash for building complex automation workflows - Monitoring & Observability:
Advanced skills in Prometheus, Grafana, Open Telemetry, and designing SLI/SLO frameworks for distributed systems - Platform Engineering
Experience:
5–8 years in Dev Ops or infrastructure roles, with demonstrated ability to design, operate, and evolve Kubernetes‑based platforms erience with proven ability to lead teams and drive infrastructure initiatives - Innovative Problem‑Solving:
Proven ability to diagnose and resolve complex distributed system issues using metrics and tracing tools like Prometheus, Grafana and Loki; skilled in performance tuning, incident response, and automating operational resilience - Collaboration & Mentorship:
Strong track record leading cross‑functional teams, mentoring junior engineers, and driving organizational adoption of best practices - Strategic Vision:
Experience developing infrastructure roadmaps, technology strategy, and managing thousands of workloads in multi‑region environments - Security and Compliance:
Awareness of industry standards related to security…
Um nach Stellen zu suchen, sie anzusehen und sich zu bewerben, die Bewerbungen aus Ihrem Standort oder Land akzeptieren, klicken Sie hier, um eine Suche zu starten: