×
Register Here to Apply for Jobs or Post Jobs. X

Senior DevOps Engineer

Job in Santa Clara, Santa Clara County, California, 95053, USA
Listing for: NVIDIA Gruppe
Full Time position
Listed on 2026-09-21
Job specializations:
  • IT/Tech
    Unix/Linux, IT Infrastructure, SRE/Site Reliability, Systems Engineer
Salary/Wage Range or Industry Benchmark: 190000 - 350000 USD Yearly USD 190000.00 350000.00 YEAR
Job Description & How to Apply Below

As a Senior Dev Ops Engineer, you will help lead the evolution of infrastructure operations within our Networking Software group. Building on a strong Linux systems administration foundation, you will build, automate, and operate scalable platforms that support networking software development and testing. This role offers an outstanding opportunity to work with elite technology and collaborate with ambitious engineers across global sites.

If you are passionate about automation, reliability, technical leadership, and continuous improvement, this is the perfect opportunity for you!

What you'll be doing:
  • Build, provision, configure, and maintain scalable Linux infrastructure for networking feature creation and validation, including physical servers, network switches, virtualization platforms, containers, and remote-management interfaces.
  • Develop automation for infrastructure provisioning, configuration management, software deployment, upgrades, and day-to-day operations using infrastructure-as-code and configuration-management practices.
  • Build reusable tools and self-service capabilities that simplify infrastructure operations, improve engineering efficiency, and reduce repetitive manual work.
  • Diagnose and resolve complex issues spanning hardware, firmware, operating systems, virtualization, containers, storage, network communications, and application environments.
  • Implement monitoring, observability, capacity management, and reliability practices to improve infrastructure performance, availability, and operational readiness.
  • Partner with engineering, IT, facilities, security, and network teams to define technical standards, maintain documentation and runbooks, and establish scalable operational processes.
  • Provide technical leadership, guide infrastructure initiatives, and mentor team members in automation, troubleshooting, and operational guidelines.
What we need to see:
  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or equivalent experience.
  • 6+ years of experience in systems engineering, Dev Ops, site reliability engineering, or infrastructure operations, including significant hands‑on experience coordinating production or engineering Linux environments.
  • Experience working in the semiconductor industry or a hardware‑focused engineering environment, with deep hands‑on expertise in bare‑metal Linux systems and server components—including CPUs, GPUs, memory, PCIe devices, NICs, storage, BIOS/UEFI, BMC/IPMI/Redfish, power, and cooling.
  • Proficiency in generative AI tools and skill in applying them effectively to automation, troubleshooting, documentation, operational analysis, and engineering efficiency.
  • Skilled at diagnosing complex issues across hardware, firmware, and operating‑system layers.
  • Strong data‑center networking knowledge, including TCP/IP, DNS, DHCP, VLANs, routing, switching, firewalls, and network troubleshooting tools.
  • Strong analytical, problem‑solving, written communication, and cross‑departmental collaboration skills, with the ability to guide technical initiatives to completion.
Ways to stand out from the crowd:
  • Experience managing Linux KVM/QEMU virtualization, Kubernetes clusters, multi‑user engineering lab environments, NFS or distributed storage systems, and automated OS or cluster provisioning platforms.
  • Experience supporting fast‑growing engineering labs, large‑scale data‑center environments, or globally distributed infrastructure.
  • Familiarity with observability platforms, including metrics, logging, tracing, alerting, incident management, and service‑level objectives.
  • Experience crafting self‑service infrastructure platforms and reusable automation that improves developer efficiency and thea Ability to establish clear, reliable, and scalable…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary