Principal Platform Software Engineer - AI & Cloud Platform
Job in
Chicago, Cook County, Illinois, 60601, USA
Listed on 2026-08-24
Listing for:
Oracle
Full Time
position Listed on 2026-08-24
Job specializations:
-
Software Development
DevOps, Cloud Engineer - Software, Backend Developer, Software Engineer
Job Description & How to Apply Below
Product And Research Role
You will partner closely with product management, customer engineering teams, and OCI engineers to translate evolving requirements into secure, scalable, and reliable software. The successful candidate is comfortable working through ambiguity, iterating quickly based on customer feedback, and balancing immediate delivery needs with long-term platform quality.
Key ResponsibilitiesPlatform Software Development
- Design, develop, test, deploy, and operate cloud-native platform services supporting AI, data, IoT, and enterprise cloud workloads.
- Drive the evolution of reusable platform capabilities by transforming customer-specific functionality into configurable OCI services.
- Contribute to architectural decisions that improve scalability, resilience, interoperability, security, and operational excellence across distributed systems.
- Implement cloud-native automation for provisioning, deployment, configuration, upgrades, rollback, and lifecycle management using CI/CD and infrastructure-as-code.
- Develop observability capabilities including telemetry, metrics, dashboards, alerts, logging, tracing, and health monitoring to ensure reliable production operations.
Software Development and Coding – Design, Testing, and Optimization
- Designs software solutions and analyzes technical requirements to deliver scalable cloud platform capabilities aligned with customer and business needs.
- Contributes throughout all phases of the software development lifecycle, including design, implementation, testing, deployment, and operational support.
- Develops production-quality software for distributed systems, APIs, service integrations, automation workflows, and data-processing pipelines.
- Implements reusable, maintainable, and secure software components that support AI, cloud, and enterprise platform services.
- Conducts code reviews and contributes to engineering standards, documentation, and best practices.
- Diagnoses and resolves complex production issues through debugging, root-cause analysis, and corrective actions.
- Implements comprehensive testing strategies including unit, integration, performance, load, and resilience testing.
- Optimizes application performance, scalability, and operational efficiency through profiling, monitoring, and continuous improvements.
- Develops and maintains APIs, ensuring secure integration, compatibility, and lifecycle management.
Software Architecture – Software System Structural Design
- Contributes to the design and implementation of scalable, resilient, and highly available distributed systems.
- Participates in architecture reviews and recommends practical design improvements supporting elasticity, fault tolerance, and operational readiness.
- Collaborates with senior engineers, product managers, and technical stakeholders to ensure architectural alignment across services.
- Evaluates technical approaches that improve cloud infrastructure, AI platform capabilities, service interoperability, and long-term maintainability.
- Implements performance optimization, security, and scalability considerations throughout software design.
Customer-Focused Delivery
- Collaborates closely with product management, customer-facing teams, and engineering partners to understand technical requirements and business objectives.
- Translates customer workflows and feedback into reusable platform capabilities and production-ready features.
- Balances rapid customer delivery with long-term platform strategy and reusable service design.
- Provides technical recommendations that improve customer outcomes while maintaining engineering quality and platform consistency.
Reliability & Operations – Software Products Support
- Builds telemetry, monitoring, alerting, diagnostics, and operational dashboards for cloud services.
- Investigates production incidents, participates in incident response, and performs root-cause analysis to improve service reliability.
- Implements automation, operational runbooks, capacity planning, and service health improvements.
- Participates in operational support rotations and contributes to maintaining service availability, reliability, and performance objectives.
- Collaborates with engineering and operations teams to…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×