×
Register Here to Apply for Jobs or Post Jobs. X

Principal Staff Engineer, AI Infrastructure

Job in Mountain View, Santa Clara County, California, 94039, USA
Listing for: PVH (Tommy Hilfiger/Calvin Klein)
Full Time position
Listed on 2026-08-02
Job specializations:
  • Software Development
    AI Engineer (Applied/Software), Software Architect, Cloud Engineer - Software, Machine Learning/ ML Engineer
Salary/Wage Range or Industry Benchmark: 260000 - 380000 USD Yearly USD 260000.00 380000.00 YEAR
Job Description & How to Apply Below

At Linked In, our approach to flexible work is centered on trust and optimized for culture, connection, clarity, and the evolving needs of our business. This role may be remote or hybrid. At Linked In, hybrid roles are performed both from home and from a Linked In office on select days, as determined by the business needs of the team. Remote roles are performed from the designated home work location upon time of hire, and any changes to this home work location requires a review of remote status and approval.

We're hiring a Principal Staff Software Engineer to lead Linked In's GPU-Based Retrieval Platform, a foundational AI infrastructure stack that powers candidate generation and retrieval across Feed, Ads, Search, Talent, and other critical product experiences. The platform sits on the hot path of billions of member interactions each day, making this one of the highest-leverage technical leadership roles within Linked In's AI Infrastructure organization.

In this role, you will own the platform's technical direction and architecture end to end, spanning large-scale indexing and retrieval, low-latency distributed serving, GPU scheduling, memory efficiency, batching, and kernel-level optimization. You will drive improvements in throughput, tail latency, retrieval quality, reliability, and cost, directly influencing member engagement, business outcomes, and AI engineer productivity across the company.

The GPU-Based Retrieval Platform team works at the intersection of GPU systems, distributed serving, information retrieval, machine learning, and product engineering. You will partner closely with teams across Feed, Ads, Search, Talent, Modeling, and Infrastructure, while setting the technical direction for a multi-team platform, influencing cross-company architecture, and mentoring senior engineers.

Responsibilities:
  • Set the long-term technical strategy and architecture for Linked In's GPU-Based Retrieval Platform.
  • Lead the design and evolution of large-scale indexing, candidate generation, vector search, and retrieval-serving systems.
  • Optimize GPU performance across CUDA, Triton, memory hierarchies, batching, scheduling, and multi-GPU communication.
  • Improve throughput, QPS per GPU, tail latency, recall quality, reliability, and infrastructure cost.
  • Build scalable, observable, and highly available multi-tenant serving systems for high-QPS production workloads.
  • Make critical architectural trade-offs across latency, quality, capacity, model complexity, and cost.
  • Partner with Feed, Ads, Search, Talent, Modeling, and Infrastructure teams to shape the platform roadmap.
  • Evaluate emerging GPU technologies, retrieval architectures, and serving frameworks for adoption at Linked In.
  • Lead complex cross-organizational initiatives from architecture and design through production rollout and adoption.
  • Mentor senior engineers, raise the technical bar, and influence AI infrastructure strategy across Linked In.
Basic Qualifications:
  • BS in Computer Science or equivalent.
  • 10+ years of industry experience in software design, development, and algorithm related solutions.
  • 5+ years in experience as an architect, or technical leadership position.
  • Experience in developing and scaling large scale databases or analytics systems
  • Experience with coding in Java, C++, or Rust; knowledge of query execution, indexing, and concurrency.
  • Hands on experience developing distributed systems, large-scale systems, databases and/or Backend APIs
Preferred Qualifications:
  • Master's or PhD in Computer Science or a related technical discipline, with experience operating at Principal Staff or equivalent scope.
  • 15+ years of software engineering experience, including 7+ years in senior technical leadership roles shaping architecture across multiple organizations.
  • 5+ years of hands on experience with CUDA, Triton, GPU kernel optimization, and hardware aware performance tuning.
  • 3+ years of experience with NCCL, distributed inference, multi-GPU communication, or optimizing workloads on modern accelerators such as NVIDIA H100 or H200 GPUs.
  • 3+ years of experience with inference optimization techniques such as quantization, mixed precision, batching, memory management, and throughput or latency tuning.
  • 5+ years of experience building large-scale retrieval systems, including ANN algorithms, hybrid retrieval, learned indexes, or billion-scale vector search.
  • 5+ years of experience building multi-tenant AI serving or retrieval platforms and balancing recall, latency, throughput, reliability, and cost across multiple products, models, or workloads.
  • Experience in one or more of the following domains: search, recommendations, feed, advertising, candidate generation, LLM serving, MLOps, or large-scale AI infrastructure.
  • Hands on experience with one or more of the following: CUDA, Triton, GPU scheduling, memory optimization, multi-GPU workloads, embeddings, vector search, ANN, or candidate generation.
Suggested

Skills:
  • AI / ML Infrastructure
  • Technical Strategy
  • Distributed Systems
  • Sta
#J-18808-Ljbffr
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary