Senior Research Scientist, ByteBrain Infrastructure Operation Technology - Infrastructure San Jose Regular
Listed on 2026-09-25
-
IT/Tech
AI Business & Operations, Machine Learning/ ML Engineer, AI Engineer (Applied/Software), Data Scientist -
Research/Development
AI Business & Operations, Data Scientist
Senior Research Scientist, Byte Brain Infrastructure Operation
Location:
San Jose
Team:
Infrastructure
Employment Type:
Regular
Job Code:
A06931
Share this listing:
ResponsibilitiesTeam Introduction Byte Brain is Byte Dance’s AI for Infrastructure (AI4
Infra) platform, dedicated to improving the efficiency, reliability, and intelligence of large-scale infrastructure systems through AI and machine learning. Byte Brain supports a wide range of infrastructure domains, including AI data center supply chains, databases, storage systems, networking, containers and virtualization, and big data platforms, powering infrastructure optimization at massive scale.
This role sits at the intersection of:
Operations Research × AIOps × AI for Infra You will have the opportunity to solve some of the most challenging optimization problems behind large-scale AI datacenters while pioneering the next generation of AI-powered decision-making systems, where LLMs, and optimization algorithms work together to improve efficiency, resource utilization, and operational intelligence across Byte Dance's global infrastructure.
- Design and develop AI, machine learning, and optimization algorithms to improve the efficiency, reliability, and performance of large-scale infrastructure systems. Areas may include AIOps, operations research, software engineering, and system optimization.
- Drive the deployment, scaling, and continuous improvement of algorithms in production environments, supporting large-scale services.
- Identify optimization opportunities and emerging challenges from real-world infrastructure scenarios, translating them into impactful research and engineering solutions.
- Conduct cutting-edge research and publish high-quality papers in top-tier conferences and journals.
Minimum Qualifications
- Proven research track record with multiple publications in top-tier conferences or journals related to AI, machine learning, operations research, systems, or related fields.
- Deep expertise in AI, machine learning, and/or operations research, with hands‑on experience in large-scale data analysis and algorithm development.
- Strong coding, implementation, and problem‑solving skills, with the ability to bridge research and production systems.
- Excellent communication and cross‑functional collaboration skills.
Industry experience applying AI and optimization techniques to real-world infrastructure challenges, such:
- AI data center supply chain optimization
- AI Ops and intelligent operations
- Software engineering productivity optimization
- Operations research and resource scheduling
- System tuning and performance
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).