Senior Platform Engineer
Listed on 2026-07-03
-
Software Development
Joining Collibra’s Platform Infrastructure Engineering team
Become a key member of Collibra's Platform Infrastructure Engineering team, reporting to the Senior Platform Infrastructure Engineering Manager. This team builds and operates the essential cloud foundation for all Collibra services. In this role, your contribution is critical as we evolve our multi-cloud, Kubernetes, IaC (Terraform/Helm), Golang automation, and Git Ops infrastructure environment for maximum efficiency and resilience, directly impacting Collibra's operational excellence.
Guided by our value of "Embrace and Drive Change," we foster a collaborative culture focused on continuous improvement that never settles long for good enough, offering you a dynamic environment to advance your technical skills and make a tangible difference.
- Developer Enablement:
Develop controllers and automations, work with development teams on refinements to platform capabilities. - Platform Contribution:
Contribute to the overall architecture of the platform infrastructure, collaborating with other infrastructure engineers using Git Ops, IaC and Kubernetes. - Operational Excellence:
Participate in on-call rotations, troubleshoot complex service issues, implement security best practices, and maintain clear documentation (architecture, procedures). - Continuous Improvement:
Stay current with platform engineering trends and infrastructure automation, identifying and implementing improvements.
- 3+ years of experience in Platform Engineering, SRE, or infrastructure-focused roles with a Bachelor's degree in Computer Science or a related technical field, OR equivalent practical experience demonstrating the skills below.
- Proven experience designing, building, and managing production services using Kubernetes and gitops / IaC at a scale of between tens and hundreds of Kubernetes clusters.
- Experience managing production workloads and infrastructure on major cloud platforms (AWS, GCP, Azure).
- Hands‑on experience operating Kubernetes clusters and managing containerized services in production.
- Demonstrable experience writing and maintaining Infrastructure as Code (IaC), preferably with Terraform, and proficiency in Golang or Python for automation.
- Preferred skills: CKA / CKAD, Istio, ArgoCD, deep experience with networks, Linux and Kubernetes, experience with monitoring/logging tools and observability spans / traces (e.g., Datadog, Grafana, Honeycomb), and proficient in creating controllers and other automation patterns to manage Kubernetes resources.
- Demonstrated proficiency in leveraging AI tools (e.g., Claude, Gemini, ChatGPT, Copilot) to solve real‑world business challenges, drive measurable outcomes, or streamline workflows.
- Must be eligible to work in the USA without requiring sponsorship.
- Because this role supports the US government, it is required that this candidate be a US citizen who resides on US soil.
- A bachelor’s degree or equivalent related working experience is required.
- Experienced in applying systematic troubleshooting and critical thinking to diagnose root causes within distributed cloud infrastructure and propose effective solutions.
- Able to demonstrate initiative in learning and utilizing evolving technologies related to cloud platforms, container orchestration, Git Ops, IaC and automation.
- An effective communicator in articulating complex technical details, designs, and trade‑offs clearly to both technical peers and potentially other stakeholders within a distributed team setting.
- Possess a mindset geared towards efficiency, proactively seeking and evaluating ways to automate manual processes and improve system reliability.
- Independently manage and complete assigned work, ensuring deliverables consistently meet defined requirements and acceptance criteria.
- Within your first month, you’ll gain exposure to our technical environment and meet the teams you’ll be consistently partnering with.
- Within your third month, you’ll independently execute assigned tasks accurately, demonstrating understanding of core platform technologies (Cloud, K8s, IaC, Git Ops) and team workflows.
- With…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).