Infrastructure Engineer
Job in
Lansing, Ingham County, Michigan, 48915, USA
Listed on 2026-07-24
Listing for:
Oracle
Full Time
position Listed on 2026-07-24
Job specializations:
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations
Job Description & How to Apply Below
* *
* Job Description:
** In this role you will be at the forefront of designing and implementing next-generation accelerated computing and AI solutions on Oracle Cloud Infrastructure (OCI). You will engage directly with customers ranging from emerging AI startups to Fortune 500 enterprises, helping them architect and deploy complex HPC and GPU clusters, AI platforms, and intelligent agentic solutions across proof-of-concept and production environments.
This highly visible and influential role combines deep technical expertise with a consultative approach to pre-sales technical consulting, solution engineering, and AI transformation strategy. You will play a critical role in supporting OCI strategic customers by optimizing infrastructure for performance, reliability, scalability, and cost-effectiveness.
You will also help mature customer relationships by establishing strong partnerships with strategic customers and peers across OCI. In addition, you will develop and implement new processes that may have OCI-wide impact, contributing to Oracle's strategic vision for cloud and AI adoption.
** Responsibilities:*
* + Work directly with strategic customers and partners across industry verticals to onboard, integrate, and optimize solutions on Oracle Cloud Infrastructure.
+ Architect and deploy large-scale GPU, HPC, and AI infrastructure on OCI using tools such as Terraform, Ansible, Slurm, Kubernetes, and other automation frameworks.
+ Build automated solutions for cluster provisioning, software deployment, infrastructure as code, performance monitoring, and capacity planning.
+ Collaborate with Oracle's largest enterprise customers to define, tailor, and validate solutions that meet high-performance computing, AI, and accelerated computing requirements.
+ Support LLM-based applications, agentic AI systems, and robotic AI platforms from architecture and design through deployment and production readiness.
+ Act as a trusted technical advisor to customers, partners, Oracle sales teams, support teams, and engineering organizations.
+ Own and promote architecture best practices across customer environments and internal Oracle teams.
+ Identify, document, and communicate technical risks, design gaps, dependencies, and potential customer-impacting issues.
+ Author SME-level technical collateral, reference architectures, deployment guides, and implementation patterns to enable repeatable customer solutions.
+ Partner with OCI engineering teams to influence product direction, validate new features, and integrate new OCI services into customer-ready solutions.
+ Troubleshoot infrastructure performance, scalability, reliability, and availability issues; implement mitigations to reduce risk and minimize downtime.
+ Participate in service-impacting incidents, drive rapid remediation, own root cause analysis, and ensure corrective actions are completed.
+ Serve as a technical escalation path for support teams requiring direct customer engagement.
+ Define and improve processes, procedures, and change-management practices to reduce operational risk for strategic customers.
+ Act as the strategic customer technical advocate when working with engineering teams to improve OCI reliability, usability, and operational readiness.
+ Stay current on emerging cloud, AI, GPU, HPC, and infrastructure technologies; evaluate their relevance to OCI customer architectures and internal workflows.
+ Document infrastructure designs, configurations, operational procedures, and lessons learned to support knowledge sharing, maintainability, and repeatability.
LI-ES2
** Responsibilities*
* *
* Qualifications:
*
* + Experience in scripting and automation using tools like Python, Ansible, Terraform, and/or Kubernetes.
+ Solid understanding of networking concepts, security principles, and best practices.
+ Strong communication and collaboration skills, with the ability to work effectively in cross-functional teams and convey technical concepts to non-technical stakeholders.
+ Excellent problem-solving skills, with the ability to troubleshoot complex issues and drive resolution in a fast-paced environment.
+ Strong documentation skills with experience documenting infrastructure designs, configurations, procedures, and troubleshooting steps to facilitate knowledge sharing, ensure maintainability, and enhance team collaboration.
+ 5+ years of experience in Architecture, Operations, and/or System Engineering
+ Recent technical experience designing and/or deploying infrastructure and high-performance computing (including GPU technology) or at least one major public cloud (AWS, GCP, Azure, OCI). Muti-Cloud Management exposure is an added advantage.
+ Linux skills with hands-on experience in Oracle Linux/RHEL/CentOS, Ubuntu, and Debian distributions, including system administration, package management, shell scripting, and performance optimization.
** Preferred Qualifications*
* + Proficiency in at least one of the programming languages such as Python, Rust, Go, Java, or…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×