×
Register Here to Apply for Jobs or Post Jobs. X

Principal Platform Engineer

Job in Denver, Denver County, Colorado, 80285, USA
Listing for: Flexential Corp.
Full Time position
Listed on 2026-07-26
Job specializations:
  • IT/Tech
    Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, IT Infrastructure
Salary/Wage Range or Industry Benchmark: 180000 - 210000 USD Yearly USD 180000.00 210000.00 YEAR
Job Description & How to Apply Below

Flexential is hiring a Principal Platform Engineer in the IT organization to plan roadmaps, establish requirements, develop and operationally manage platform technologies including Observability, Dev Ops, ITSM and Integrations. Current platform initiatives include building a next-gen Open Telemetry observability platform for 40+ data center facilities and platforms using a LGTM stack (Loki, Grafana, Tempo, Mimir); and enabling secure high-velocity SDLC capability enabling paved pathways, engg excellence measurements and devsecops across multiple development teams.

This role sits at the intersection of engineering management and hands-on technical work. You will lead a small team of platform engineers, create/capture requirements, establish and own technical planning and implementation, and be accountable for platform reliability, security, and delivery timelines. This is a high-visibility, high-impact role - the platforms you build will be foundational to Flexential's IT services, as well as enable AIOps and AI infrastructure.

Key Responsibilities and Essential Job Functions
  • Lead the design, development, deployment and operational management of automated, resilient, high availability, self-healing, secure platforms with native-AI capabilities for IT needs, serving both internal as well as customer business capabilities.
  • Lead, Build and manage the Platform Engineering team and function — hiring, mentoring, performance management, and technical roadmap ownership.
  • Plan, build and operate an Open Telemetry Observability platform with technologies including Grafana, Mimir, Loki, Tempo, Alert manager on Kubernetes/RKE2 using Helm and ArgoCD.
  • Build an automated federated Observability Edge Stack — Prometheus + OTel collector nodes deployed per site and Zabbix auto-discovery configuration and Prometheus scrape profile library for 10+ device classes (Cisco, Juniper, Dell, Net App, etc.).
  • Design, develop and manage engineering lifecycle platforms for high-velocity secure SDLC using Gitlab and similar / related technologies.
  • Build and operate iaC and CI/CD platforms including Git Lab CI/CD, Terraform, Ansible AWX, Helm, and ArgoCD for automated provisioning and application deployment.
  • Own, enhance and operate critical IT platform technologies e.g Boomi for integrations, AWS for Cloud environments, including their hosted infrastructure.
  • Establish and enforce platform security posture: secrets management via Cyber Ark/Conjur, RBAC, mTLS, compliance boundary design, and zero inbound telemetry architecture.
  • Build and integrate ITSM capabilities for various platforms e.g automated incident creation, CI enrichment, and CMDB correlation.
  • Define and implement extensibility patterns including AIOps: e.g anomaly detection hooks, event correlation pipeline design, and integration with future ML/AI tooling.
  • Partner with other IT and business teams for App Dev, requirements capture, delivery validation and integration needs.
  • Represent platform engineering in cross-functional architecture reviews and executive-level program updates.
  • Perform other management and technical duties as required and assigned for team and operational resilience e.g team building, on-call rotation, etc.
  • Travel maybe required to team or project events.
Required Qualifications
  • 8+ years of relevant technical experience with 2+ years in a management (or Principal-level) role leading a engineering team.
  • Dev Ops / Platform Engineering - 8+ years, End-to-end ownership of developer/infrastructure platforms;
    Kubernetes, Helm, ArgoCD, service-mesh, containerized workloads.
  • Git Ops / CI-CD - 5+ years Git Lab CI/CD, pipeline authoring, infrastructure-as-code delivery.
  • 8+ years of expert level automation frameworks experience with Python, Terraform, Ansible, etc.
  • Infrastructure (Linux/VM) - 8+ years Linux systems administration, VM lifecycle (VMware vCenter/VCF), Netapp storage and compute provisioning.
  • Working knowledge of Networking - 3+ years, TCP/IP, BGP/OSPF, SNMP protocol.
  • AI tooling – Strong understanding (or 1+ years experience) with MCP, Agentic workflows, SRE workflows e.g AIOps for Anomaly detection, event correlation, alert noise reduction on Prometheus and…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary