×
Register Here to Apply for Jobs or Post Jobs. X

Senior Director, Inference Products and Optimizations

Job in Seattle, King County, Washington, 98127, USA
Listing for: PVH (Tommy Hilfiger/Calvin Klein)
Full Time position
Listed on 2026-07-20
Job specializations:
  • IT/Tech
    AI Engineer (Applied/Software), Machine Learning/ ML Engineer, Systems Engineer
Salary/Wage Range or Industry Benchmark: 274400 - 343000 USD Yearly USD 274400.00 343000.00 YEAR
Job Description & How to Apply Below

Senior Director of Engineering – Inference Engine organization at Digital Ocean

What You’ll Do
  • Recruit, mentor, and coach engineers to foster a culture of ownership, technical excellence, and continuous improvement.
  • Work with product teams to define and execute on the roadmap for Digital Ocean’s Inference Products, including Serverless, Dedicated, Inference Router, Batch, and Multimodal Inference.
  • Lead the design and evolution of the inference serving stack, driving technical strategy across vLLM, SGLang, and LLM‑D to optimize throughput, latency, and GPU utilization at scale.
  • Architect the model‑serving and optimization layer, spanning quantization, KV‑cache management, speculative decoding, and disaggregated serving, to deliver best‑in‑class performance‑per‑dollar.
  • Collaborate with product management, other engineering teams, and key stakeholders to align priorities, manage dependencies, and communicate progress and risks.
  • Ensure production health, stability, and on‑call rotation to maintain customer SLAs.
  • Institutionalize benchmarking frameworks, observability, and auto‑tuning capabilities to guide system and infrastructure tuning efforts.
  • Encourage contributions to open‑source inference engines to advance capabilities.
Qualifications
  • 10+ years of software engineering experience with 6+ years in technical leadership or management, ideally in inference or AI/ML systems.
  • Deep expertise in distributed systems design, modern AI/ML technologies, Kubernetes at scale, LLM inference, and AI workload orchestration.
  • Strategic knowledge of GPU architectures (NVIDIA/AMD), interconnects, and hardware topology for AI training and inference performance.
  • Familiarity with container runtime internals, system isolation, and security contexts for shared infrastructure risk management.
  • Expertise in defining, tracking, and operationalizing deep infrastructure and inference metrics (e.g., TTFT, TPOT) for performance improvements and SLO compliance.
  • Product mindset: ability to translate complex technical requirements into user‑focused product features while balancing innovation and reliability.
  • Excellent communication skills for explaining technical decisions to non‑technical stakeholders and aligning diverse teams.
  • Strong ownership mentality and proactive drive to identify and resolve issues preventing delivery of value.

Compensation: $274,400 – $343,000

Hybrid role

#J-18808-Ljbffr
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary