ByteDance
Posted 1mo ago

Research Scientist - Driven Agent Self-Evolution - Global Frontier Tech Recruitment Program - 2027 Start (PhD)

ByteDance
San Jose, California, United States
OnsiteFull Time
Responsibilities
  • researching agents
  • building pipelines
  • collaborating teams
Requirements
  • PhD in CS/AI/ML or related
  • Strong ML/Deep RL foundations
  • Research publications
  • Programming in Python
  • Experience with PyTorch/TensorFlow/JAX, and interest in LLM agents and continual learning
Technical tools mentioned
PythonPyTorchTensorFlowJAX

Job description

We are looking for talented individuals to join our team in 2027. As a graduate, you will get opportunities to pursue bold ideas, tackle complex challenges, and unlock limitless growth. Launch your career where inspiration is infinite at our Company.

Successful candidates must be able to commit to an onboarding date by end of year 2027. Please state your availability and graduation date clearly in your resume.

Team Introduction:
The Applied Machine Learning Ark team combines system engineering and machine learning to develop and operate Large Language Model (LLM) service platforms that offer businesses Model-as-a-Service (MaaS) solutions, serving both large model providers and downstream users. The US team drives the design, development, and operation of MaaS solutions across the US and international markets outside mainland China. We are building full-stack, end-to-end solutions spanning text and multimodal LLM algorithms, LLM training/fine-tuning/inference frameworks, prompt engineering, model alignment, and intelligent agent systems. Beyond model serving, we operate large-scale log analytics pipelines that process massive volumes of invocation logs from text models, multimodal models, and agent systems — extracting usage patterns, quality signals, and actionable insights to inform model improvement, system optimization, and product decisions through continuous, data-driven feedback loops. We are actively seeking talented engineers and researchers specializing in Large Language Models and AI Agent systems to join our dynamic team.

Topic Content:
As model capabilities improve and computation becomes cheaper, the key challenge in real-world deployment is no longer building a capable one-off assistant, but building agent systems that improve through use. This research studies a self-evolving agent framework in which execution traces, environmental responses, and human feedback are converted into signals for continual improvement. The goal is to establish a closed loop from execution to feedback, attribution, accumulation, and reuse, so that system capability grows with real-world interaction. We focus on three tightly coupled directions: adaptive runtime, which enables online adjustment of planning, tool use, and control policies; experience compilation, which abstracts reusable skills, rules, and failure patterns from trajectories; and evaluation-governance loops, which ensure that each system update is measurable, comparable, and reversible. Together, these components support a synergistic co-evolution of the model layer and the harness layer, improving task quality, reducing manual intervention, and accumulating durable capability over time. More broadly, this work reframes agent deployment as a continual learning systems problem: not how to build a stronger static agent, but how to build an operational system that learns reliably from experience.

Responsibilities:
- Research and develop agent frameworks that continuously learn and improve from execution traces, user feedback, and environmental signals.
- Build large-scale log analytics pipelines to extract quality signals, usage patterns, and actionable insights from model and agent invocation logs, driving data-informed system and model improvements.
- Explore and apply frontier techniques in LLM post-training, reasoning, and planning to enhance agent capabilities.
- Collaborate across algorithm research, platform engineering, and product teams to turn research ideas into production-grade systems at scale.

Minimum Qualifications:
- Individuals who are completing or have recently completed a Ph.D. in Computer Science, Artificial Intelligence, Machine Learning, or a closely related discipline.
- Strong theoretical and practical foundation in machine learning, deep learning, reinforcement learning, or optimization.
- Research experience in at least one of the following areas: LLM-based agents, planning and reasoning, multi-agent systems, continual/lifelong learning, or LLM post-training (e.g., RLHF, DPO, GRPO, self-play).
- Strong programming skills in Python and proficiency with ML frameworks (e.g., PyTorch, TensorFlow, JAX).
- Publication record at top-tier venues (e.g., NeurIPS, ICML, ICLR, ACL, EMNLP, NAACL, AAAI, AAMAS, COLM).
- Strong problem-solving skills and ability to thrive in a fast-paced, collaborative environment.

Preferred Qualifications:
- Publications in areas directly related to agent learning and adaptation, such as tool use, self-improvement, skill discovery, trajectory optimization, reward modeling, or agent evaluation.
- Research experience in LLM reasoning and planning, including chain-of-thought, tree/graph search, Monte Carlo methods, or inference-time compute scaling.
- Experience training or fine-tuning large language models, including supervised fine-tuning, preference optimization, or curriculum learning.
- Hands-on experience building or evaluating LLM-based agent systems (e.g., ReAct, function calling, code generation agents, or multi-agent orchestration).
- Familiarity with meta-learning, few-shot generalization, or transfer learning in the context of LLM-based systems.
- Experience with feedback-driven optimization loops, such as online learning, bandit methods, or evolutionary strategies applied to agent improvement.
- Strong interest in bridging frontier AI research with production-grade engineering — turning papers into systems that work at scale.
- Internship experience at technology companies or research organizations.

About ByteDance

Developing AI-driven content platforms and mobile applications.

Similar jobs

Research Scientist roles near San Jose, California
17h
Save
Mark Applied
Hide
Research Scientist, All-Optical Working Memory
Emeryville, California, United States
$80k-$150k/yr OnsiteFull Time
Astera Institute
Astera Institute: Conducts fundamental research in artificial intelligence and longevity science.
PhD or equivalent research experience, hands-on mouse survival surgery, and expertise in neuroscience research. Preferred: two-photon imaging, holographic optogenetics, neural-data analysis, rodent behavior, and scientific computing.
Python, MATLAB, Suite2p, CaImAn, Ti:Sapphire, ChRmine, GCaMP, GLM, IACUC, AAALAC, SLM
2d
Save
Mark Applied
Hide
Research Scientist, Life Sciences (Chemistry)
San Francisco, California, United States
$300k-$320k/yr OnsiteFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
5+ YOEPhD in chemistry and sustained medicinal chemistry experience with hands-on small-molecule synthesis, SAR-driven design, optimization, and cross-functional research collaboration.
NMR, LC-MS, HPLC, Python, RDKit, pandas
2d
Save
Mark Applied
Hide
Research Scientist
Presidio, California, United States
OnsiteFull Time
Arcade
Arcade: Design and buy custom physical products using generative AI.
Ph.D. in a related AI or computer science field; top-tier publications; expert Vision-Language Model knowledge, Python, PyTorch, multimodal architectures, and transformer-based computer vision.
Vision-Language Model (VLM), Qwen-VLM, PaliGemma2, Python, PyTorch, CVPR, ICCV, ECCV, NeurIPS, ICLR, ACL
2d
Save
Mark Applied
Hide
Staff Research Scientist, Siri Innovation Studio
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Strong background in machine learning modeling with passion for incubating and deploying innovative AI technologies, prototyping user experiences, and collaborating across software, hardware, and design teams.
AI, ML, iOS, iPadOS, macOS, watchOS, visionOS
3d
Save
Mark Applied
Hide
Senior Research Scientist, Foundational AI in Health
Mountain View, California, United States
$174k-$252k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
2+ YOEPhD in computer science, biomedical engineering, or related quantitative field; 2+ years with machine learning frameworks and distributed accelerator training; scientific publications required; Python and health data research preferred.
JAX, PyTorch, TensorFlow, Python, Electronic Health Records (EHR)
4d
Save
Mark Applied
Hide
Research Scientist , WW Sustainability
Seattle or San Francisco
$136k-$184k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
4+ YOEPhD or master's degree with 4+ years of quantitative field research experience; scientific problem-solving, experimental and observational data analysis required.
R, MATLAB, Python
5d
Save
Mark Applied
Hide
AI Research Scientist - Optexity
Palo Alto, California, United States
OnsiteFull Time
TrueMeter
TrueMeter: AI-powered platform for commercial energy savings and bill management.
Research experience, model training experience, Python fluency, and a master's degree or higher in machine learning or a related field. Strong ownership and comfort with ambiguity required.
Python, LLM
5d
Save
Mark Applied
Hide
Synthetic Biologist - Research Scientist
Livermore, California, United States
$146k-$186k/yr OnsiteFull Time, Temporary
Lawrence Livermore National Laboratory
Lawrence Livermore National Laboratory: Develops science and technology for United States national security.
Bachelor’s degree or equivalent experience in a biological science field, molecular biology techniques, bacterial culture, recombinant protein expression and purification, analytical problem-solving, and collaborative research skills.
PCR, Gibson Assembly, Golden Gate Assembly, Bash, ZSH, R, Python