Infrastructure Operations Engineer Barstow, TX
Listed on 2026-07-17
-
IT/Tech
Network Engineer, SRE/Site Reliability, IT Support, Cloud Computing: Infrastructure & Operations
Infrastructure Operations Engineer Barstow, TX
Barstow, TX
About NscaleNscale is the GPU cloud engineered for AI. We provide cost‑effective, high‑performance infrastructure for AI start‑ups and large enterprise customers. Nscale enables AI‑focused companies to achieve superior results by reducing the complexity of AI development. Our GPU cloud bolsters technical capabilities and directly supports strategic business outcomes, including cost management, rapid innovation, and environmental responsibility.
Our CultureWe thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you’ll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you’ll be contributing to building the technology that powers the future.
Aboutthe Role
We’re looking for an Engineer that has good people, leadership & technical skills.
- A technical expert responsible for ensuring the efficiency, reliability, and scalability of data centre infrastructure.
- Comfortable problem‑solving & making decisions on complex topics with high levels of ambiguity in a results‑driven environment.
- Comfortable influencing without authority and exceptional at building relationships with senior stakeholders across the business to get things done.
- Understanding and skillset to grasp technical concepts and problems quickly.
- Strong analytical skills.
- Doer who is extremely organised and diligent.
- Self‑starter, curious, and quick to learn, knowing what questions to ask to get up to speed quickly.
- Join the Support duty rotation and handle day‑to‑day tickets and alerts, escalating early and appropriately. Collaborate with Engineering with guidance when incidents or changes require it.
- Accurately record, update, manage and resolve tickets using the ticketing system whilst keeping all parties informed of the tickets progression.
- Follow established runbooks to resolve common issues. Propose improvements and contribute incremental fixes with review.
- Keep tickets up to date with clear notes, next steps, and customer communications via the agreed channels.
- Learn the Platform fundamentals so you can help customers get value from our services, asking for support when deeper expertise is needed.
- Participate in monitoring, troubleshooting, and triage. Capture logs and facts to enable efficient handover.
- Deliver assigned tasks and project work to agreed quality and timelines. Flag blockers early and seek help when needed.
- Share knowledge by documenting steps you’ve validated and by contributing to training materials. Shadow seniors during complex work to build capability.
- Take part in incident reviews as a contributor and help track preventative follow‑ups in your scope.
- Identify areas for implementation for automation to optimise processes.
- Constantly endeavor to learn and upskill.
- Collaborate with cross‑functional teams for service improvements. Be the escalation point for onsite operations staff.
- Participate in on‑call or out‑of‑hours work when scheduled and after onboarding.
- Availability to travel to Nscale or Customer locations to assist with deployments, trouble shooting and operational tasks and attendance of supplier related training courses.
- Growth mindset. Curious, dependable, and collaborative. Seek feedback, ask questions and invest in learning to progress toward senior.
- Platform and DC fundamentals. Awareness of servers, networks, storage, and virtualization concepts, ideally from a support or operations background.
- Linux fundamentals. Comfortable with the CLI, services via systemd, file systems, permissions, and basic networking tools. Able to troubleshoot common issues and know when to escalate.
- Networking basics. Solid grasp of IP addressing, subnets, VLANs, routing at a high level, DNS, and firewalls. Advanced topics like BGP or VXLAN are a plus, not required.
- Kubernetes exposure. Understand core concepts like nodes, pods, services, and logs. Can perform basic troubleshooting and follow runbooks. Cluster‑level administration experience is a nice to have.
- GPU…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).