×
Register Here to Apply for Jobs or Post Jobs. X

Senior Core Infrastructure Engineer; OCI Object Storage

Job in Nashville, Davidson County, Tennessee, 37247, USA
Listing for: Oracle
Full Time position
Listed on 2026-07-25
Job specializations:
  • Software Development
Salary/Wage Range or Industry Benchmark: 120000 - 160000 USD Yearly USD 120000.00 160000.00 YEAR
Job Description & How to Apply Below
Position: Senior Core Infrastructure Engineer (OCI Object Storage)

Job Description

NOTE:

This role will sit onsite at our Nashville, TN location.

Are you interested in building large-scale distributed infrastructure for the cloud? Oracle’s Cloud Infrastructure team is building Infrastructure-as-a-Service technologies that operate at high scale in a broadly distributed multi-tenant cloud environment. Our customers run their businesses on our cloud, and our mission is to provide them with industry-leading compute, storage, networking, database, security, and an ever-expanding set of foundational cloud-based services.

As part of this effort, the Object Storage Service team is looking for hands‑on engineers with expertise and passion in solving difficult problems in distributed systems, large-scale storage, and highly available services. If this is you, you can be part of the team that drives the best-in-class Object Storage Service into the next phase of its development. These are exciting times for the service – we are growing fast, and delivering innovative, enterprise-class features to satisfy the most demanding big data and enterprise workloads for our customers.

An engineer at any level can have significant technical and business impact.

Responsibilities
  • Design, implement, and optimize components in distributed systems with an emphasis on scalability, resiliency, and operability.
  • Deliver features and load/performance tests; leverage data‑plane platforms and distributed state tools for high‑volume retrieval, storage, and processing; and review peers’ implementations for scalability compliance.
  • Build fault‑tolerant paths (redundancy, replication, automatic fail‑over), apply recovery‑oriented principles, and implement retries, circuit breakers, and timeouts.
  • Proactively detect and mitigate issues via tests, alarms, dashboards, and telemetry; author runbooks and participate in incident response and root‑cause analyses.
  • Implement standard replication and synchronization techniques, develop automation/IaC for troubleshooting and maintenance, and apply advanced security controls (encryption, access, remediation) while ensuring change, compliance, and documentation standards are met.
System Design & Architecture – System Scalability
  • Implement and contribute to the development of components of distributed systems that support horizontal and vertical scaling including leveraging distributed state management tools.
  • Optimize code and/or systems for large-scale data processing.
  • Implement scalability requirements for assigned components and review implementation of team members.
  • Leverage components of data‑plane platforms to handle large-scale data retrieval, storage, and processing.
  • Implement performance and load testing.
System Design & Architecture – System Reliability Design
  • Collaborate with the team to build fault‑tolerant components capable of withstanding in-service updates by implementing redundancy, replication, and automatic fail‑over mechanisms.
  • Apply recovery-oriented computing principles to design components that effectively handle service disruptions.
  • Implement retry mechanisms, circuit breakers, and timeouts to help handle network unreliability.
System Design & Architecture – System Reliability Performance
  • Implement tests and alarm configurations to proactively detect and address issues or failures.
  • Support efforts to recover from failures by drafting and executing runbooks and operational procedures.
  • Build and customize dashboards, telemetry systems, and alerting mechanisms to monitor component health.
System Design & Architecture – Correctness / Availability
  • Design and implement functional requirements and testing for assigned features within an existing system.
  • Implement test scenarios (e.g., fault-injection, brown-out) to evaluate system correctness.
  • Implement standard data replication and synchronization techniques to maintain data integrity and availability.
Operational Troubleshooting & Incident Management
  • Diagnose, debug, and resolve issues in system components to support ongoing operations.
  • Implement basic strategies to prevent interruptions, ensuring no maintenance windows are required for customers and users when resolving issues.
  • Design and implement automation scripts…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary