×
Register Here to Apply for Jobs or Post Jobs. X

Senior Technical Project Manager – Applied AI

Job in Palo Alto, Santa Clara County, California, 94306, USA
Listing for: Nebius
Full Time position
Listed on 2026-08-28
Job specializations:
  • IT/Tech
    IT Project Manager, Cloud Computing: Infrastructure & Operations, Systems Engineer
Salary/Wage Range or Industry Benchmark: 147200 - 224000 USD Yearly USD 147200.00 224000.00 YEAR
Job Description & How to Apply Below

About Nebius

Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.

Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.

Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.

Role

As a Technical Project Manager in Token Factory, your primary focus will be coordinating complex cross-functional projects, including new model launches, bringing new GPU platforms into production serving, engineering quality and platform-wide improvement projects, and the end-to-end delivery pipeline for new AI models. You will bring together multiple engineering teams, understand the critical path, proactively manage dependencies and risks, and ensure that ambitious technical goals are delivered predictably, on time, and without unnecessary operational overhead.

Your

Responsibilities Will Include
  • Leading template and model delivery — from weights through validation to production endpoint.
  • Coordinating model onboarding and production delivery, from infrastructure readiness to successful customer availability.
  • Running the Applied AI / Customer Delivery sync — driving open issues to resolution and unblocking teams between meetings.
  • Coordinating the bring-up of new GPU platforms for serving — engine and kernel readiness, benchmarking against current hardware, and template migration once validated.
  • Partnering with Token Factory Product on the serverless launch path — ensuring validated templates hand off cleanly into catalogue, model card, and public availability.
  • Building and maintaining execution plans, identifying risks early, and ensuring blockers are resolved before they impact delivery.
  • Working closely with engineering managers, technical leads, product managers, SRE, infrastructure, networking, security, and other platform teams.
  • Facilitating technical decision-making and ensuring ownership and accountability across complex cross-team projects.
  • Continuously improving engineering delivery processes within Applied AI's delivery chain to make execution more predictable while maintaining a sustainable pace for engineering teams.
Requirements
  • Excellent project management and delivery skills with the ability to break down ambiguous initiatives into executable plans
  • Ability to identify critical paths, manage complex dependency graphs, and coordinate multiple parallel work streams
  • Strong risk management and prioritization skills, with the ability to make progress in fast-changing environments
  • Excellent written and verbal communication skills in English
  • Comfortable leading incident coordination, facilitating discussions, documenting decisions, and driving follow-up actions
  • Strong technical background that allows you to understand engineering discussions, infrastructure dependencies, and architectural trade-offs without being the primary implementer
It will be an added bonus if you have
  • Experience working with GPU infrastructure, AI/ML platforms, or large-scale inference systems
  • Experience delivering cloud infrastructure or data center deployment projects
  • Familiarity with capacity planning, production operations, and reliability engineering
  • Previous experience in high-growth infrastructure or platform engineering organizations where priorities change quickly and execution speed matters
Pay Transparency

We offer competitive compensation and benefits packages. Actual compensation will be determined based on job-related factors, including experience, skills, qualifications, the level at which the candidate is hired, and geographic location, consistent with applicable law.

Base Compensation Range

$147,200—$224,000 USD

Benefits & Perks
  • Competitive compensation
  • Career growth and learning…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary