Scale AI
Posted 9mo ago

Machine Learning Research Engineer, Agent Data Foundation - Enterprise GenAI

Scale AI
San Francisco or New York
$181k-$315k/yrHybridFull Time
Responsibilities
  • build pipelines
  • train models
  • develop agents
Requirements
  • 3+ years in building with LLMs in production
  • Experience creating high-quality data for LLM/Agent
  • Publications in top conferences
  • Advanced degree in CS
Technical tools mentioned
PythonPyTorchTensorFlowReinforcement Learning

Job description

AI is becoming vitally important in every function of our society. At Scale, our mission is to accelerate the development of AI applications. For 9 years, Scale has been the leading AI data foundry, helping fuel the most exciting advancements in AI, including generative AI, defense applications, and autonomous vehicles. With our recent investment from Meta, we are doubling down on building out state of the art post-training algorithms to reach the performance necessary for complex agents in enterprises around the world. 

The Enterprise ML Research Lab works on the front lines of this AI revolution. We are working on an arsenal of proprietary research, tools, and resources that serve all of our enterprise clients. As MLRE on the Data Foundation team, you’ll work on cutting edge research to define the data flywheel that makes the whole machine move. This includes research around synthetic environments from task definitions, building agents for trace analysis, and contributing to a cutting edge framework that automatically hill-climbs agent-building from an eval set. This will involve creating best-in-class Agents that achieve state of the art results through a combination of post-training + agent-building algorithms.

If you are excited about shaping the future of the modern GenAI movement, we would love to hear from you!

You will: 

  • Build synthetic data pipelines to generate enterprise environments to use for RL post-training
  • Create agents to convert traces from production into actionable insights to use to improve agents
  • Contribute to our agent building product which can construct other agents using coding agents + proprietary algorithms
  • Train state of the art models, developed both internally and from the community, to deploy to our enterprise customers. 

Ideally you’d have:

  • 3+ years of building with LLMs in a production environment
  • Clear experiences with constructing high quality data to use to improve an LLM/Agent
  • Publications in top conferences such as NEURIPS, ICLR, or ICML within the last two years
  • PhD or Masters in Computer Science or a related field

Compensation packages at Scale for eligible roles include base salary, equity, and benefits. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position, determined by work location and additional factors, including job-related skills, experience, interview performance, and relevant education or training. Scale employees in eligible roles are also granted equity based compensation, subject to Board of Director approval. Your recruiter can share more about the specific salary range for your preferred location during the hiring process, and confirm whether the hired role will be eligible for equity grant. You’ll also receive benefits including, but not limited to: Comprehensive health, dental and vision coverage, retirement benefits, a learning and development stipend, and generous PTO. Additionally, this role may be eligible for additional benefits such as a commuter stipend.

Please reference the job posting's subtitle for where this position will be located. For pay transparency purposes, the base salary range for this full-time position in the locations of San Francisco, New York, Seattle is:
$180,600$315,000 USD

PLEASE NOTE: Our policy requires a 90-day waiting period before reconsidering candidates for the same role. This allows us to ensure a fair and thorough evaluation of all applicants.

About Us:

At Scale, our mission is to develop reliable AI systems for the world's most important decisions. Our products provide the high-quality data and full-stack technologies that power the world's leading models, and help enterprises and governments build, deploy, and oversee AI applications that deliver real impact. We work closely with industry leaders like Meta, Cisco, DLA Piper, Mayo Clinic, Time Inc., the Government of Qatar, and U.S. government agencies including the Army and Air Force. We are expanding our team to accelerate the development of AI applications.

We believe that everyone should be able to bring their whole selves to work, which is why we are proud to be an inclusive and equal opportunity workplace. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability status, gender identity or Veteran status. 

We are committed to working with and providing reasonable accommodations to applicants with physical and mental disabilities. If you need assistance and/or a reasonable accommodation in the application or recruiting process due to a disability, please contact us at [email protected]. Please see the United States Department of Labor's Know Your Rights poster for additional information.

We comply with the United States Department of Labor's Pay Transparency provision

PLEASE NOTE: We collect, retain and use personal data for our professional business purposes, including notifying you of job opportunities that may be of interest and sharing with our affiliates. We limit the personal data we collect to that which we believe is appropriate and necessary to manage applicants’ needs, provide our services, and comply with applicable laws. Any information we collect in connection with your application will be treated in accordance with our internal policies and programs designed to protect personal data. Please see our privacy policy for additional information.

About Scale AI

Provides data and infrastructure for training artificial intelligence models.

Year founded
2016
Employees
1200
Organization type
Private
Latest investment
Raised $14.30B Series G (2025) — led by Meta
Headquarters
US

Similar jobs

Machine Learning Research Engineer roles near San Francisco, California
4h
Save
Mark Applied
Hide
Sr. Machine Learning Research Engineer, Siri Speech
Cupertino or North America
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
PhD or equivalent industry experience; expertise in efficient deep learning; Python and JAX, PyTorch, or TensorFlow proficiency; software design, coding, parallel computing, and large-scale ML experience preferred.
Python, JAX, PyTorch, TensorFlow
2w
Save
Mark Applied
Hide
Machine Learning/Research Engineer Graduate (Monetization Technology-Ads Core Global) - 2027 Start (PhD)
San Jose, California, United States
$162k-$317k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
PhD in CS or related field; strong C/C++ and Python skills; ML/DL/algorithm theory; experience with TensorFlow/PyTorch/MXNet; familiarity with Linux and large-scale distributed systems.
C, C++, Python, Linux, TensorFlow, Pytorch, MXNet, Spark
1mo
Save
Mark Applied
Hide
Machine Learning Research Engineer
Redwood City or Palo Alto
$140k-$240k/yr HybridFull Time
WindBorne Systems
WindBorne Systems: Operates smart weather balloons to provide global atmospheric data.
Strong ML research and engineering skills with Python and PyTorch, experience with large messy datasets, research taste, and ability to convert experiments into reusable systems.
Python, PyTorch
3mo
Save
Mark Applied
Hide
Senior / Staff Machine Learning Research Engineer
South San Francisco, California, United States
$209k-$233k/yr HybridFull Time
Calico
Calico: Researching aging biology to develop longevity-improving therapies.
5+ YOE5+ years ML software engineering experience; strong Python and JAX or PyTorch skills; experience building model training/serving/eval platforms; willing to work onsite at least four days/week.
Python, JAX, PyTorch, Ray, Spark, BigQuery, AlphaFold
3mo
Save
Mark Applied
Hide
Machine Learning Research Engineer
Emeryville, California, United States
$200k-$330k/yr HybridFull Time
Profluent
Profluent: Designs functional proteins using deep generative AI models.
3+ YOEBS/MS in CS/ML or related,3+ years building/training ML models in PyTorch,strong Python and engineering skills,experience with transformer architectures,cloud/container familiarity,and GPU optimization/benchmarking.
PyTorch, Python, GCP, AWS, Azure, Kubernetes, Docker, CUDA, Triton, DDP, FSDP
3mo
Save
Mark Applied
Hide
Research Engineer, Machine Learning (RL Velocity)
San Francisco or New York City
$500k-$850k/yr HybridFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
Strong software engineering fundamentals; experience with ML infrastructure or research tooling; able to operate across stack; bias toward shipping and quick iteration.
JAX, PyTorch
4mo
Save
Mark Applied
Hide
Research Scientist
San Francisco, California, United States
$225k-$300k/yr OnsiteFull Time
Latent Health
Latent Health: AI platform for automating healthcare administrative workflows and documentation.
Strong ML/DL background with PyTorch; track record from idea to validated results; able to work independently in ambiguity; experience with longitudinal or clinical data is a plus.
PyTorch
4mo
Save
Mark Applied
Hide
Machine Learning Research Engineer/Scientist
Redwood City, California, United States
OnsiteFull Time
Sunday
Sunday: Developing autonomous robots to perform household chores.
3+ YOE3+ years of ML work for robotics, Python and deep learning (PyTorch preferred), hands-on robot learning, data collection, and cross-functional collaboration.
Python, PyTorch, Machine Learning