Sr Platform Infrastructure Engineer, AI Transformation
Listed on 2026-10-05
-
Software Development
DevOps, Cloud Engineer - Software
Vantor is forging the new frontier of spatial intelligence, helping decision makers and operators navigate what's happening now and shape what's coming next. Vantor is a place for problem solvers, changemakers, and go-getters-where people are working together to help our customers see the world differently, and in doing so, be seen differently. Come be part of a mission, not just a job, where you can:
Shape your own future, build the next big thing, and change the world.
To be eligible for this position, you must be a U.S. Person, defined as a U.S. citizen, permanent resident, Asylee, or Refugee.
Export Control/ITAR:
Certain roles may be subject to U.S. export control laws, requiring U.S. person status as defined by 8 U.S.C. 1324b(a)(3).
Please review the job details below.
Vantor is seeking a Senior Platform Engineer, AI Infrastructure to help build and operate the secure, scalable foundation for our AI-driven development organization. This role supports Site Sentry, Maritime Sentry, and Storyline-products delivered as part of the broader Tensor globe spatial intelligence platform.
You will work at the intersection of software engineering, cloud infrastructure, reliability, automation, and performance. You will help development teams move quickly and safely by improving the systems that build, deploy, monitor, and operate mission-critical applications across cloud, on-premises, and edge environments.
The ideal candidate is a strong software engineer who can design and develop production-quality solutions, automate repetitive operational work, instrument systems for observability, troubleshoot complex failures, and continuously improve platform performance. Our codebase is primarily Python, and success in this role requires a pragmatic, hands-on approach to engineering and operations.
RequirementsPut security first: protect company, customer, mission, and operational data through secure-by-design engineering, least-privilege access, secrets management, threat-aware automation, and strong security and compliance practices.
Design, implement, and maintain reliable CI/CD pipelines for software projects, including automated testing, quality gates, artifact management, release promotion, rollback, and deployment verification.
Automate deployment, scaling, configuration, and monitoring of cloud-based infrastructure using technologies such as Kubernetes, Docker, Terraform, and Google Cloud services including Cloud Run.
Troubleshoot infrastructure, deployment, availability, latency, capacity, and performance issues across distributed systems.
Collaborate with software development teams to optimize code and services for performance, scalability, reliability, operability, and cost efficiency.
Implement and enforce best practices for security, compliance, change management, incident response, and operational readiness.
Continuously evaluate and recommend improvements to Dev Ops, SRE, developer productivity, observability, and platform engineering processes and tools.
Demonstrated experience as a full-stack or backend software developer, with the ability to write maintainable production code-not only configuration and infrastructure definitions.
Professional experience developing in Python and working with APIs, services, databases, automated tests, and source-control workflows.
Strong communication and collaboration skills, with the ability to explain technical tradeoffs and partner effectively with engineering, product, security, and customer-facing teams.
Own and improve platform reliability through SRE practices, including service-level objectives, monitoring, alerting, incident response, root-cause analysis, and operational reviews.
Lead performance management and optimization for services, infrastructure, Cloud Run workloads, databases, and supporting platform components.
Build and maintain deployment automation that makes releases repeatable, observable, secure, and easy to recover.
Develop internal tools and platform capabilities in Python to reduce toil, improve developer experience, and make the right operational behavior the easiest behavior.
Design and operate observability across logs, metrics, traces, health checks, dashboards, and actionable alerts.
Partner with developers to improve application architecture, code quality, resource utilization, scalability, and production readiness.
Build, deploy, and maintain on-premises environments at customer sites, including installation, upgrades, configuration, diagnostics, and…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).