ByteDance
Posted 2w ago

AI Infrastructure Engineer Intern (Compute Efficiency & Scheduling) - 2027 Summer

ByteDance
Seattle, Washington, United States
OnsiteInternship
Responsibilities
  • optimizing compute
  • building tools
  • improving scheduler
Requirements
  • Currently pursuing a Bachelor's or Master's in Computer Science or related discipline
  • Available for a 12-week Summer 2027 internship
  • Coding experience in Go, C++, Python, Java, or C# required
Technical tools mentioned
GodelGoC++PythonJavaC#LinuxKubernetes

Job description

We decide how ByteDance's massive fleet of computers gets used — turning hundreds of clusters and millions of daily workloads (microservices, big data, and large AI/LLM jobs) into a system that's highly efficient, reliable, and increasingly self-managing, powered by our own intelligent scheduling platform Godel (open-sourced to the community). Our mission: get the most value out of every unit of computing power, at a scale few places in the world can offer, and let AI increasingly help run the infrastructure.

We are looking for talented individuals to join us for an internship. Our internship program offers students hands-on experience, industry exposure, and opportunities to apply their knowledge to real-world challenges while building a strong foundation for personal and professional growth.
Interns will gain practical experience, explore potential career paths, and participate in social events, learning programs, and development workshops alongside industry professionals.
Candidates may apply to a maximum of two positions across Our Company and its affiliates globally. Applications will be considered in the order they are submitted.
Applications are reviewed on a rolling basis, so we encourage you to apply early. Please clearly state your availability in your resume, including your start and end dates.

Candidates who pass resume screening will be invited to participate in Our Company's technical online assessment.

What You Might Work On
One focused, real project within our scope — for example:
- Help make our intelligent, AI-driven scheduler smarter at placing and balancing workloads.
- Explore new ways to get more useful work out of the same hardware — smarter use of compute, memory, or power.
- Build tools that help a global platform run reliably and increasingly manage itself, including AI-agent-assisted automation.

Why This Role Matters
The AI era runs on compute — and using it efficiently is one of the defining challenges for our company and the whole industry. Our team owns that challenge at ByteDance's global scale, making the infrastructure behind TikTok and our AI/LLM products faster, smarter, and far more efficient. As an intern, you'll get a real project at the center of it.

Minimum Qualifications
- Currently pursuing a Undergraduate/ Master's in Computer Science or a related discipline.
- Able to commit to a 12-week internship during Summer 2027.
- Experience coding in Go, C++, Python, Java, or C#.

Preferred Qualifications
- Intent to return to your degree program after the internship.
- Strong fundamentals in systems, distributed computing, or software engineering — shown through internships, work, research, competitions, or projects.
- Curiosity about large-scale or AI/ML infrastructure; exposure to Linux, containers, or Kubernetes is a plus but not required.
- High creativity and quick problem-solving.

About ByteDance

Developing AI-driven content platforms and mobile applications.

Similar jobs

AI Infrastructure Engineer roles near Seattle, Washington
4d
Save
Mark Applied
Hide
AI Infrastructure Engineer
Boston or San Francisco or Scottsdale or Seattle
$134k-$247k/yr OnsiteFull Time
Axon
AxonNASDAQ: AXON: Develops weapons and software for law enforcement and safety.
4+ YOERequires 4+ years in platform, DevOps, infrastructure, internal tools, automation, or software engineering; cloud infrastructure, CI/CD, infrastructure-as-code, production support, and AI application knowledge.
Vercel, Azure, AWS, GitHub Actions, Terraform, Bicep, Pulumi, CloudFormation, Python, TypeScript, JavaScript, Node.js, GCP, SSO, OAuth, OIDC, Entra ID, Azure AD, RBAC, Slack, Jira, Confluence, Quip, Microsoft 365, Salesforce, Snowflake, ServiceNow, React
4w
Save
Mark Applied
Hide
AI Infrastructure Engineer, Sandbox Platform
San Francisco or Seattle or New York City
$180k-$225k/yr OnsiteFull Time
Scale AI
Scale AI: Provides data and infrastructure for training artificial intelligence models.
4+ YOE4+ years building high-performance systems software; deep Linux internals, containerization/virtualization, systems programming (Go/Rust/C/C++); strong debugging and API/SDK design skills.
Docker, Firecracker, gVisor, QEMU, Kata Containers, Go, Rust, C/C++, Kubernetes, OpenHands, Agent2Agent, MCP, CRIU
1mo
Save
Mark Applied
Hide
AI Infrastructure Engineer
Seattle, Washington, United States
$152k-$332k/yr OnsiteFull Time
Zoom
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
5+ YOEBachelor's in CS/Engineering/AI, 5+ years software engineering experience in infrastructure and distributed systems, GPU/CUDA expertise, container and cloud experience, Python/C++ programming, PyTorch and Transformers experience.
Docker, Kubernetes, CUDA, Python, C++, PyTorch, Transformers, BrightHire
4d
Save
Mark Applied
Hide
AI & HPC Infrastructure Engineer
Albany or Arlington or Atlanta or Austin or Beaverton or Bentonville or Boston or Carmel or Charlotte or Chicago or Cincinnati or Cleveland or Columbus or Culver City or Denver or Des Moines or Detroit or Hartford or Houston or Irvine or Irving or Kirkland or Miami or Milwaukee or Minneapolis or Morristown or Mountain View or Nashville or New York City or Oklahoma City or Overland Park or Philadelphia or Pittsburgh or Raleigh or Redmond or Sacramento or San Diego or San Francisco or Scottsdale or Seattle or St. Louis or St. Petersburg or Walnut Creek or United States
$80k-$266k/yr HybridFull Time
Accenture
AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
5+ YOERequires 5+ years designing AI infrastructure, accelerated computing, clusters, orchestration, and automation; bachelor's degree or equivalent experience. Strong Kubernetes, Python, Terraform, and cloud expertise required.
Slurm, Run:ai, Kubernetes, NVIDIA Base Command Manager (BCM), NVIDIA NGC, NCCL, NVLink, CUDA, TensorRT-LLM, vLLM, SGLang, Triton Inference Server, NVIDIA Dynamo, llm-d, MLPerf, NCCL tests, fio, iperf, InfiniBand, Ethernet, SONiC, NVMe, NVMe-oF, VAST, Weka, DDN, AWS, Azure, GCP, VMware, Nutanix, Python, Terraform, Ansible, REST, OpenAPI, JSON, YAML, TensorFlow, PyTorch, JAX, Jupyter notebooks, Google Colab, MCP
2mo
Save
Mark Applied
Hide
Cloud Infrastructure and AI Efficiency Engineer
Seattle, Washington, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Not specified in the provided description.
4mo
Save
Mark Applied
Hide
Senior DGX Cloud AI Infrastructure Software Engineer
Santa Clara or Austin or Redmond or Washington or Oregon
$184k-$357k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOE8+ years building software infrastructure for large-scale AI systems; BS in Computer Science or related (or equivalent experience); strong debugging/root-cause skills; experience with observability (ELK, Prometheus, Loki), distributed systems, AI training/inference infra; Python and C/C++.
ELK, Prometheus, Loki, Python, C/C++, NCCL, IB verbs, ucx, libfabrics, PyTorch, TensorFlow, JAX, Ray, RDMA