×
Register Here to Apply for Jobs or Post Jobs. X

Machine Learning Engineer Intern (AML-Engine-Orchestration) - 2027 Start Undergraduate​/Master I

Job in Northern, Floyd County, Kentucky, USA
Listing for: Pangle
Apprenticeship/Internship position
Listed on 2026-08-14
Job specializations:
  • Software Development
    DevOps, Software Engineer, Cloud Engineer - Software
Salary/Wage Range or Industry Benchmark: 20000 - 33000 USD Yearly USD 20000.00 33000.00 YEAR
Job Description & How to Apply Below
Position: Machine Learning Engineer Intern (AML-Engine-Orchestration) - 2027 Start Undergraduate/Master I[...]
Location: Northern

Join us as we work together to inspire creativity and enrich life around the globe.

Location:

San Jose

Team:

Technology

Employment Type:

Intern

Job Code:

A221977

Share this listing:

Responsibilities

Responsibilities

The Data-AML-Engine Orchestration team builds large-scale machine learning infrastructure that powers online model serving across Byte Dance products, including Tik Tok. We develop the orchestration, scheduling, and resource management systems that connect heterogeneous compute infrastructure with production ML workloads.

You will work on systems that directly affect GPU utilization, serving latency and availability, infrastructure reliability, and MLE productivity. Depending on your background and interests, you may focus on one or more of the following areas.

We are looking for talented individuals to join us for an internship. Our internship program offers students hands-on experience, industry exposure, and opportunities to apply their knowledge to real-world challenges while building a strong foundation for personal and professional growth.

Interns will gain practical experience, explore potential career paths, and participate in social events, learning programs, and development workshops alongside industry professionals.

Candidates may apply to a maximum of two positions across Our Company and its affiliates globally. Applications will be considered in the order they are submitted.

Applications are reviewed on a rolling basis, so we encourage you to apply early. Please clearly state your availability in your resume, including your start and end dates.

  • Design and build foundational orchestration capabilities for machine learning platforms, including Kubernetes Operators, container runtimes, and lifecycle management for jobs, services, and stateful workloads.
  • Build multi-tenant resource and quota systems that support priorities, preemption, fair sharing, elasticity, and cross-cluster scheduling. Improve GPU utilization and cost efficiency through resource pooling and Fin Ops.
  • Build lifecycle orchestration for online model serving, including model and image distribution, deployment, upgrades, rollback, autoscaling, multi-cluster operation, and disaster recovery.
  • Build serving orchestration and traffic management capabilities for disaggregated serving clusters, including topology-aware scheduling, KV Cache affinity, intelligent request routing, and QoS/SLA management.

Qualifications

Minimum Qualifications
  • Currently pursuing a Bachelor's or Master's degree in Computer Science, Software Engineering, Artificial Intelligence, or a related technical field.
  • Proficiency in at least one of Go, C++, or Python, with a solid foundation in data structures, algorithms, and software engineering principles.
  • Familiarity with Linux and a foundational understanding of operating systems, computer networks, concurrent programming, and distributed systems.
  • Strong hands-on and exploratory abilities, with a willingness to investigate systems through source code, metrics, logs, profiling, and experiments.
  • A systematic and quantitative approach to problem solving, with the ability to define measurements, test hypotheses, and validate system improvements.
  • Demonstrated ownership and collaboration through coursework, research, internships, open-source contributions, or other engineering projects.
Preferred Qualifications
  • Experience with Kubernetes, container runtimes, resource scheduling, quota management, multi-tenant systems, or Fin Ops.
  • Contributions to open-source infrastructure projects such as Kubernetes, Volcano, Koordinator, or Open Kruise.
  • Experience with model serving systems such as vLLM, SGLang, Triton, KServe, or Ray Serve, or an understanding of KV Cache, Continuous Batching, Prefill/Decode disaggregation, or model parallelism.
  • Experience with online services, gateways, traffic management, autoscaling, performance optimization, or highly available distributed systems.
  • Experience with GPU/NPU programming, heterogeneous resource scheduling, model distribution, or inference performance analysis.

Job Information

Job Information

About Us

Founded in 2012, Byte Dance's mission is to inspire creativity and enrich life. With a suite…

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary