1,433 open roles
Senior Research Scientist, ByteBrain Infrastructure Operation
Job description
Team Introduction ByteBrain is ByteDance’s AI for Infrastructure (AI4Infra) platform, dedicated to improving the efficiency, reliability, and intelligence of large-scale infrastructure systems through AI and machine learning. ByteBrain supports a wide range of infrastructure domains, including AI data center supply chains, databases, storage systems, networking, containers and virtualization, and big data platforms, powering infrastructure optimization at massive scale.
This role sits at the intersection of: Operations Research × AIOps × AI for Infra
You will have the opportunity to solve some of the most challenging optimization problems behind large-scale AI datacenters while pioneering the next generation of AI-powered decision-making systems, where LLMs, and optimization algorithms work together to improve efficiency, resource utilization, and operational intelligence across ByteDance's global infrastructure.
Responsibilities
- Design and develop AI, machine learning, and optimization algorithms to improve the efficiency, reliability, and performance of large-scale infrastructure systems. Areas may include AIOps, operations research, software engineering, and system optimization.
- Drive the deployment, scaling, and continuous improvement of algorithms in production environments, supporting large-scale services.
- Identify optimization opportunities and emerging challenges from real-world infrastructure scenarios, translating them into impactful research and engineering solutions.
- Conduct cutting-edge research and publish high-quality papers in top-tier conferences and journals.
Qualifications
Minimum Qualifications
- Proven research track record with multiple publications in top-tier conferences or journals related to AI, machine learning, operations research, systems, or related fields.
- Deep expertise in AI, machine learning, and/or operations research, with hands-on experience in large-scale data analysis and algorithm development.
- Strong coding, implementation, and problem-solving skills, with the ability to bridge research and production systems.
- Excellent communication and cross-functional collaboration skills.
Preferred Qualifications
Industry experience applying AI and optimization techniques to real-world infrastructure challenges, such as:
- AI data center supply chain optimization
- AI Ops and intelligent operations
- Software engineering productivity optimization
- Operations research and resource scheduling
- System tuning and performance optimization
- Large-scale infrastructure management and automation
Description copied from ByteDance's careers page. Read the full posting before you apply.
More jobs at ByteDance
Talent Acquisition Partner (Global AI To B) - BytePlus
ByteDance· San Jose, California, United StatesTalent Acquisition Partner (Global AI To B) - BytePlus
ByteDance· London, England, United KingdomCorporate Finance Associate - Group FP&A
ByteDance· Hong Kong (China), Hong Kong Island, Hong Kong, ChinaBackend Software Engineer (SRE) - Cloud Infrastructure
ByteDance· SingaporeCustomer Success Manager - Lark Japan
ByteDance· Tokyo, Japan
More jobs in San Jose
Advanced AI for Industry & Society Staff TA - College of Engineering - Integrated Innovation Institute
Carnegie Mellon University· Silicon Valley, CAWarehouse Part Time Overnight
Lowe's· San Jose, CA (S San Jose) 1756· $40k – $41kRetail Manager - CLUB Membership
Bass Pro Shops· San Jose, CA· $70k – $83kSoftware Architect - PNS AI Governance
TikTok· San Jose, California, United StatesLead Organizer
Californians for Justice· San Jose, CA 95133· $82k