Eightfold.ai
Posted 3mo ago

Staff Machine Learning Engineer - Agentic Models, LLM, RAG, GenAI

Eightfold.ai
Santa Clara, California, United States
$232k-$310k/yrHybridFull Time
Responsibilities
  • leading research
  • architecting systems
  • developing models
Requirements
  • Expert in ML/GenAI/LLMs, NLP
  • Distributed systems
  • Python
  • TensorFlow/PyTorch
  • AWS
  • Docker/Kubernetes
  • SQS/Kafka
  • ML infra
  • 6-10+ years
  • Strong problem-solving and collaboration
Technical tools mentioned
PythonTensorFlowPyTorchAWSDockerKubernetesAWS SQSKafkaLangGraphCrewAIAutoGenPineconepgvectorQLORADPOvLLMTensorRT-LLM

Job description

About Eightfold.ai
Eightfold is a global leader in AI-native enterprise talent intelligence, trusted by the world's largest and most respected Fortune 500 organizations. Our platform is built from the ground up, operating at scale across Azure and AWS, deployed in multiple regions globally, including IL4-compliant environments for the US Government, supporting users in 100+ countries and 30+ languages.

Today, Eightfold is at the forefront of agentic AI, delivering intelligent agents that actively drive outcomes across hiring and talent workflows, while much of the industry is still experimenting with prototypes. We are defining the next era of agentic talent systems.

What sets Eightfold apart is not just the technology and our mission, but the team behind it. We are a deeply technical, execution-driven organization that values ownership, collaboration, and high standards. Our engineers, product leaders, and go-to-market teams work closely together—in person and across functions—to build systems that scale in the real world. If you're excited to work on hard problems, move with urgency, and raise the bar every day.
Eightfold is the place to build agentic systems that transform how the world works.

About AII/ML Team
Our AI/ML team is building the core technology that makes our platform autonomous and agentic. This team is the driving force behind Eightfold's cutting-edge solutions. We're a group of passionate experts who thrive on pushing the boundaries of applied machine learning. We work with massive datasets, tackle complex challenges, and focus on real-world reliability, shipping systems that manage multi-step reasoning and distributed state management at enterprise scale.
You will work with enormous data sets to build intelligent agents that proactively assist millions of users across 100+ countries and 30+ different languages.

Responsibilities:
  • Lead the team in: research, design, development, and deployment of advanced AI agents and agentic systems.
  • Architect and implement complex multi-agent systems, including planning, decision-making, and execution capabilities.
  • Develop and integrate large language models (LLMs) and other state-of-the-art AI techniques to enhance agent autonomy and intelligence.
  • Build robust, scalable, and reliable infrastructure to support the deployment and operation of AI agents at scale.
  • Diagnose and troubleshoot issues in complex distributed environments and optimize system performance.
  • Contribute to the team's technical growth and knowledge sharing.
  • Stay up-to-date with the latest advancements in AI research and agentic AI and apply them to our products.
  • Leverage enterprise data, market data, and user interactions to build intelligent and personalized agent experiences.
Qualifications:
  • Knowledge and passion in machine learning algorithms, Gen AI, LLMs, and natural language processing (NLP).
  • Understanding of agent-based modeling, reinforcement learning, and autonomous systems.
  • Experience with large language models (LLMs) and their applications in Agentic AI.
  • Proficiency in programming languages such as Python, and experience with machine learning frameworks like TensorFlow or PyTorch.
  • Experience with cloud platforms (AWS) and containerization technologies (Docker, Kubernetes).
  • Understanding of distributed system design patterns and microservices architecture.
  • Experience with message queuing systems (AWS SQS, Kafka).
  • Hands-on experience with system integration patterns and API design.
  • Excellent problem-solving and data analysis skills.
  • Strong communication and collaboration skills.
  • Master’s or Ph.D. in Computer Science, Artificial Intelligence, or a related field, or equivalent years of experience.
  • Min 6-10-+ years of relevant work experience in AI, Machine Learning, and applying data science to real-world use cases.
  • Strong track record of taking systems from prototype to production with a focus on scalability and reliability.
  • Experience with RAG architectures, including hybrid retrieval, vector databases (Pinecone, pgvector), and rerankers.
  • Proficiency in building multi-agent workflows using frameworks like LangGraph, CrewAI, or AutoGen.
  • Knowledge of fine-tuning strategies (QLORA, DPO) and inference optimization (vLLM, TensorRT-LLM).
Desired Skills & Experience:
  • Research experience in agentic AI or related fields.
  • Experience building and deploying AI agents in real-world applications.

Pay Transparency
Please note this role is categorized as onsite or hybrid in Zone A: Santa Clara, CA The base salary ranges below are provided for pay transparency. Base pay is only one piece of our total compensation package as this role is also eligible for annual bonus and equity awards. Compensation varies depending on a number of factors including qualifications, skills, competencies, experience and zones determined by location. 

Zone A: Base annual salary range: $232,000 to $310,000 + annual performance bonus up to 20% + preIPO equity (stock options).

Hybrid Work @ Eightfold: We embrace a hybrid work model that aims to boost collaboration, enhance our culture, and drive innovation through a blend of remote and in-person work. We are committed to creating a dynamic and flexible work environment that nurtures the collaborative spirit of our team. Starting May 1, 2025, employees residing near Santa Clara, CA office location will return to the office three days a week.
Eightfold.ai provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, veteran or disability status. 

Experience our comprehensive benefits with family medical, vision and dental coverage, a competitive base salary, and eligibility for equity awards and discretionary bonuses or commissions.


#LI-Hybrid

About Eightfold.ai

AI-native platform for talent management and workforce optimization.

Year founded
2016
Employees
900
Organization type
Private
Latest investment
Raised $220.00M Series E (2021) — led by SoftBank Vision Fund, General Catalyst, Lightspeed Venture Partners, Foundation Capital, IVP
Headquarters
US

Similar jobs

Staff Machine Learning Engineer roles near Santa Clara, California
2mo
Save
Mark Applied
Hide
Senior Staff Machine Learning Engineer
San Francisco or Sunnyvale
$243k-$357k/yr OnsiteFull Time
DoorDash
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
5+ YOE5+ years building and scaling ML/AI models; expertise in deep learning, ranking, relevance, LLMs; strong programming in Python/Java/C++; experience with PyTorch, TensorFlow, XGBoost; full ML lifecycle and experimentation.
Python, Java, C++, PyTorch, TensorFlow, XGBoost, Claude Code, Codex, Cursor
2mo
Save
Mark Applied
Hide
Staff Machine Learning Engineer
Santa Barbara or San Diego or San Francisco or Denver
$200k-$250k/yr RemoteFull Time
AppFolio
AppFolioNASDAQ: APPF: Provides cloud-based property and investment management software.
Production ML infra on AWS; ML tooling; cost discipline; collaboration with ML researchers and engineers.
Python, Docker, CI/CD, AWS, ECS, SageMaker, GPUs, LangChain, LangGraph, TensorRT, Triton
2mo
Save
Mark Applied
Hide
Staff ML Engineer, Fine Tuning - Slack
Seattle or Atlanta or San Francisco
$197k-$314k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
5+ YOE5+ years NLP model training and fine-tuning; experience with PyTorch, TensorFlow, JAX; production ML systems; programming languages; strong communication.
PyTorch, TensorFlow, JAX, Python, Go, Scala, Java, PHP, Ruby, C
2mo
Save
Mark Applied
Hide
Staff ML Engineer, Fine Tuning - Slack
Seattle or Atlanta or San Francisco
$197k-$314k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: Provides cloud-based customer relationship management and enterprise software.
5+ YOE5+ years in NLP model training and fine-tuning; strong DL framework experience; production-grade ML pipelines; proficient in multiple programming languages.
PyTorch, TensorFlow, JAX, TorchScript, TensorRT, ONNX, Python, PHP, Ruby, Go, C, Scala, Java
2mo
Save
Mark Applied
Hide
Staff MLE, GAI Search Platform - Moveworks
Mountain View, California, United States
HybridFull Time
ServiceNow
ServiceNowNYSE: NOW: Provides a cloud platform for automating enterprise digital workflows.
5+ YOE5+ years engineering; 2+ years as senior engineer; experience with search platforms; ML/NLP; Python; familiarity with LLMs.
Python, Golang
2mo
Save
Mark Applied
Hide
Staff Machine Learning Engineer
San Francisco or Minneapolis
HybridFull Time
Shipt
Shipt: Provides same-day delivery services from local retailers via app.
5+ YOE5+ years of machine learning and backend software engineering; backend in Go/Java and Python; embeddings, similarity search, ranking models; ML pipelines; distributed systems; SQL/NoSQL; API serving; A/B testing.
Go, Java, Python, MLflow, Kubeflow, Airflow, REST, gRPC, Model servers, SQL, NoSQL
2mo
Save
Mark Applied
Hide
Staff Machine Learning Engineer (Research Scientist) - DFAI
San Francisco or New York City or Seattle
$249k-$368k/yr HybridFull Time
Plaid
Plaid: Provides financial data connectivity and payment infrastructure via APIs.
7+ YOESenior ML engineer with 7–12+ years (MS) or 5–9+ years (PhD); strong technical leadership; expertise in transformers/LLMs; end-to-end production ownership; Python and software engineering fundamentals; fintech domain experience a plus.
Python, PyTorch, TensorFlow, Distributed training, Pretraining infrastructure, ML platforms
2mo
Save
Mark Applied
Hide
Staff Machine Learning Engineer, Search Ranking
Palo Alto or Seattle or San Francisco or Santa Monica or Bellevue
$229k-$343k/yr HybridFull Time
Snap
SnapNYSE: SNAP: Provides visual messaging software and augmented reality wearable devices.
8+ YOE8+ years ML experience; degree in a technical field; experience with ranking models and large-scale ML systems; strong collaboration and leadership.
Spark, Flink, Beam, TensorFlow, PyTorch, JAX, Python, C++, Java, Scala, ML Infrastructure, Airflow