Platform Engineer II
Listed on 2026-06-26
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Systems Engineer, SRE/Site Reliability
PLATFORM ENGINEER
(Plano Texas or Draper Utah, In-Office)
ABOUT UPBOUNDUpbound Group, Inc. (effective February 27, 2023: NASDAQ: UPBD) is an omni-channel platform company committed to elevating financial opportunity for all through innovative, inclusive, and technology-driven financial solutions that address the evolving needs and aspirations of consumers. The Company’s customer-facing operating units include industry-leading brands such as Acima,
Rent‑A‑Center,
and Brigit that facilitate consumer transactions across a wide range of store-based and digital retail channels, including over 2,400 company branded retail units across the United States, Mexico and Puerto Rico. Upbound Group, Inc. is headquartered in Plano, Texas.
The Platform Engineering II is a lead individual contributor responsible for driving the development, maintenance, and optimization of platform infrastructure and services. This individual will take ownership of significant platform initiatives, lead design efforts for new features, and actively mentor junior engineers. They will play a critical role in evolving our platform's reliability, efficiency, and security, and in defining and implementing best practices that significantly accelerate the Software Development Life Cycle (SDLC).
This role also serves as a key builder of AI-powered platform capabilities, including the design, integration, and ongoing maintenance of AI components embedded in the CI/CD pipeline.
- Lead the design, implementation, and management of complex, scalable, and secure platform components across cloud environments (e.g., AWS, GCP).
- Drive the adoption and continuous improvement of infrastructure as code (IaC) practices, developing reusable modules and patterns.
- Architect and implement advanced CI/CD pipelines, including Git Ops principles and progressive delivery strategies.
- Design, implement, and optimize comprehensive observability solutions, ensuring robust monitoring, logging, tracing, and alerting.
- Provide expert‑level troubleshooting and perform root cause analysis for critical platform issues.
- Lead architectural discussions and propose innovative solutions to complex platform challenges.
- Ensure high availability, disaster recovery, and performance of critical platform services through proactive engineering.
- Collaborate extensively with Dev Ex, development and security teams, acting as a technical liaison to integrate applications with platform services and provide strategic guidance.
- Mentor junior and mid‑level engineers, fostering their technical growth and promoting best practices within the team.
- Evaluate and recommend new technologies, tools, and processes to enhance platform capabilities.
- Design, build, and operate AI‑powered components integrated into the CI/CD pipeline, including LLM and MCP‑based automation capabilities.
- Build and maintain the Developer Portal, including service scaffolding, golden path templates, and adoption metrics to improve developer experience
- Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
- 6+ years of progressive experience in Platform Engineering, Dev Ops, or SRE, with demonstrated leadership in technical initiatives.
- 2+ years of working with AI/LLM‑powered components
- Deep expertise in cloud computing platforms (AWS and/or GCP), including advanced networking, compute, and security services.
- Mastery of infrastructure as code (Terraform, Cloud Formation), including developing and managing complex IaC repositories.
- Extensive experience with advanced CI/CD pipeline automation (Git Hub Actions, Git Lab CI, Argo CD, Jenkins).
- Expert‑level knowledge of containerization (Docker) and orchestration (Kubernetes), including cluster management and optimization.
- Highly proficient in scripting/programming languages (e.g., Python, Go, Bash) for automation, tool development, and API integration.
- Proven experience in designing and implementing comprehensive observability stacks (Grafana, Prometheus, New Relic, Open Telemetry, ELK, Splunk).
- Strong understanding of enterprise networking, security best practices, and…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).