×
Register Here to Apply for Jobs or Post Jobs. X

Senior Machine Learning Engineer, Applied Science Data Frameworks

Job in San Jose, Santa Clara County, California, 95199, USA
Listing for: Adobe
Full Time position
Listed on 2026-09-14
Job specializations:
  • Software Development
    Machine Learning/ ML Engineer, Data Engineering
Salary/Wage Range or Industry Benchmark: 183300 - 265350 USD Yearly USD 183300.00 265350.00 YEAR
Job Description & How to Apply Below

About The Role

We are looking for a Senior Machine Learning Engineer to join our Applied Science Data Frameworks team responsible for building the foundational infrastructure that powers large-scale multimodal AI training and inference. This role is ideal for someone with strong distributed systems and data engineering fundamentals who is eager to work in an ML-adjacent environment, contributing to training data loaders, distributed inference frameworks, feature enrichment pipelines, and dataset management systems that enable ML teams to train foundation models at petabyte scale.

You'll work on high-impact projects involving distributed data loading for PyTorch training workloads, batch inference pipelines for feature enrichment, semantic search infrastructure for dataset discovery, and production-grade ML data pipelines that support generative AI model development. Your systems will process billions of images,

About

The Role

We are looking for a Senior Machine Learning Engineer to join our Applied Science Data Frameworks team responsible for building the foundational infrastructure that powers large-scale multimodal AI training and inference. This role is ideal for someone with strong distributed systems and data engineering fundamentals who is eager to work in an ML-adjacent environment, contributing to training data loaders, distributed inference frameworks, feature enrichment pipelines, and dataset management systems that enable ML teams to train foundation models at petabyte scale.

You'll work on high-impact projects involving distributed data loading for PyTorch training workloads, batch inference pipelines for feature enrichment, semantic search infrastructure for dataset discovery, and production-grade ML data pipelines that support generative AI model development. Your systems will process billions of images,
videos, and multimodal content across large-scale GPU clusters.
If you're excited about building distributed data frameworks, optimizing data pipelines at scale, and growing your expertise in ML infrastructure, we'd love to hear from you.

What You'll Do
  • Contribute to building and maintaining distributed training data loaders that handle multi-source data ingestion, temporal sampling, and real-time transformations for large-scale model training workflows.
  • Help implement and maintain feature enrichment pipelines and dataset registry systems that support multimodal model training across images, video, documents, and text.
  • Build and maintain batch inference pipelines for large-scale feature extraction,
  • processing assets through distributed GPU clusters with queue management and fault tolerance.
  • Develop data processing systems using frameworks like Apache Ray, Spark, DuckDB, or similar distributed computing tools for SQL-based data ingestion and Apache Arrow-based storage formats.
  • Support semantic search capabilities and vector database infrastructure (e.g., Open Search, LanceDB) for dataset discovery and embedding-based retrieval.
  • Contribute to CI/CD infrastructure for ML systems including self-hosted runner management, Docker image builds, automated testing pipelines, and deployment automation.
  • Collaborate with ML research teams to translate training requirements into reliable, scalable data loading and preprocessing infrastructure.
  • Write reusable framework components, SDKs, and documentation to help accelerate platform adoption across modeling teams.
  • Optimize data pipeline performance across dimensions like startup latency, throughput, memory footprint, and GPU utilization.
  • Contribute to observability and reliability standards for production data systems supporting 24/7 training workloads.
What You Need to Succeed
  • 5-6 years of professional experience building and operating distributed…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary