Back to jobs

Software Engineer Intern (AML-Engine-Orchestration) - 2027 Start

ByteDance · San Jose, CA

Source: Intern List

Spotted 4d agoInternship

Job details

Employment
Internship
Level
Internship
Education
Bachelor's degree
Posted
Oct 5, 2026
Last confirmed open
Oct 7, 2026
Job description

About this role

Public source summary from intern-list.com / Jobright.

ByteDance is a technology company developing products and platforms that connect people and support content creation. The Software Engineer Intern will help build large-scale machine learning infrastructure for online model serving, focusing on orchestration, scheduling, resource management, serving lifecycle operations, and traffic management.

Responsibilities

Design and build foundational orchestration capabilities for machine learning platforms, including Kubernetes Operators, container runtimes, and lifecycle management for jobs, services, and stateful workloads Build multi-tenant resource and quota systems that support priorities, preemption, fair sharing, elasticity, and cross-cluster scheduling.

Improve GPU utilization and cost efficiency through resource pooling and FinOps Build lifecycle orchestration for online model serving, including model and image distribution, deployment, upgrades, rollback, autoscaling, multi-cluster operation, and disaster recovery Build serving orchestration and traffic management capabilities for disaggregated serving clusters, including topology-aware scheduling, KV Cache affinity, intelligent request routing, and QoS/SLA management

Qualifications: Currently pursuing a Bachelor's or Master's degree in Computer Science, Software Engineering, Artificial Intelligence, or a related technical field Proficiency in at least one of Go, C++, or Python, with a solid foundation in data structures, algorithms, and software engineering principles Familiarity with Linux and a foundational understanding of operating systems, computer networks, concurrent programming, and distributed systems Strong hands-on and exploratory abilities, with a willingness to investigate systems through source code, metrics, logs, profiling, and experiments A systematic and quantitative approach to problem solving, with the ability to define measurements, test hypotheses, and validate system improvements Demonstrated ownership and collaboration through coursework, research, internships, open-source contributions, or other engineering projects Experience with Kubernetes, container runtimes, resource scheduling, quota management, multi-tenant systems, or FinOps Contributions to open-source infrastructure projects such as Kubernetes, Volcano, Koordinator, or OpenKruise Experience with model serving systems such as vLLM, SGLang, Triton, KServe, or Ray Serve, or an understanding of KV Cache, Continuous Batching, Prefill/Decode disaggregation, or model parallelism Experience with online services, gateways, traffic management, autoscaling, performance optimization, or highly available distributed systems Experience with GPU/NPU programming, heterogeneous resource scheduling, model distribution, or inference performance analysis

Benefits: Interns have day one access to health insurance, life insurance, wellbeing benefits and more. Interns also receive 10 paid holidays per year and paid sick time (56 hours if hired in first half of year, 40 if hired in second half of year). Interns who are not working 100% remote may also be eligible for housing allowance.

Hands-on experience, industry exposure, and opportunities to apply their knowledge to real-world challenges. Social events, learning programs, and development workshops alongside industry professionals.

Interested in this role?Continue on ByteDance's careers page.
Apply on ByteDance