779 llm engineer jobs at 402 companies in Cotati, CA
3mo
Save
Mark Applied
Hide
3mo
Distributed LLM Inference Engineer
San Francisco or Palo Alto
$170k-$247k/yrHybridFull Time
Anyscale: Cloud platform for scaling distributed machine learning applications.
Familiarity with running ML inference at large scale with high throughput and low latency; experience with PyTorch; solid understanding of distributed systems.
Clera: AI talent agent matching professionals with high-growth startup roles
8+ YOE8+ years engineering experience with production LLM systems, building evals and observability, experience with LLM agents and data-residency/SOC2/GDPR constraints, strong communication and product-engineer instincts.
Machine Learning Engineer, Model Evaluations (Speech LLM) - San Francisco
San Francisco, California, United States
$180k-$270k/yrHybridFull Time
Plaud: Develops AI-powered voice recorders and automated transcription software.
Python software engineering; building distributed systems, data pipelines, and evaluation harnesses at scale; partner with ML researchers to define benchmarks; build dashboards and monitor model health; debug mid-training anomalies; communicate results clearly.
Mira Mace: AI-powered healthcare advocacy and navigation for Medicare beneficiaries.
Staff-level backend/full-stack or ML engineering experience, production LLM/ML systems experience, architecture and build-vs-buy judgment, hands-on coding and code review, mentoring and technical leadership.
Baseten: Scalable infrastructure platform for deploying and serving AI models.
4+ YOE1+ MgmtLead a team of Forward Deployed Engineers; strong Python, ML inference, LLM experience; 4+ years software engineering; leadership experience; excellent communication.
Python, vLLM, TensorRT, Triton, Hugging Face, Ray Serve
San Francisco or New York City or Vancouver or Amsterdam or London or Paris or Singapore or Tokyo
$190k-$316k/yrHybridFull Time
AmplitudeNASDAQ: AMPL: Digital analytics platform for understanding and optimizing customer behavior.
3+ YOE3+ years software engineering experience with 2+ years building production LLM or applied AI systems; experience designing and scaling AI products, strong product sense, and collaboration skills.
GatherUp: Software platform for customer review and reputation management.
3+ YOE3+ years in GTM engineering/rev ops/growth engineering; coding in Python, JavaScript, or SQL; experience with AI voice tools, Clay, ZoomInfo, Apollo, n8n, Zapier, Salesforce/HubSpot; building agentic AI/LLM workflows; strong analytical and cross-functional communication skills.
Austin or Chicago or New York City or Salt Lake City or San Francisco or United States
$115k-$175k/yrRemoteFull Time
Gong: AI platform analyzing customer interactions to improve sales performance.
Background in software/data engineering or applied AI, experience shipping AI/LLM tools to production, systems thinking, API and workflow integration experience, strong independent delivery and stakeholder partnership skills.
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years at the intersection of design and engineering; hands-on experience building and shipping LLM-powered apps, internal tools, and high-fidelity prototypes; strong craft, collaboration, and product judgment.
Mondrio: AI-native platform for agentic pricing and revenue management.
8+ YOE8+ years engineering experience with production LLMs, built evals/observability, shipped LLM features, strong communication, product-engineer instincts, US work authorization required.
San Francisco or Toronto or United States or Canada
$190k-$249k/yrRemoteFull Time
EvenUp: AI-powered document generation and analysis for personal injury law.
5+ YOE5+ years engineering experience, building production-quality software; ownership of projects end-to-end; experience with frontend systems for deploying LLM-based algorithms; strong communication, mentoring, and cross-team collaboration skills.
San Francisco or Seattle or Los Angeles or New York City or United Kingdom or Ireland or Poland or Germany or Australia or North America or Europe
$200k-$260k/yrRemoteFull Time
Whatnot: Social marketplace for buying and selling via live streams
4+ YOE4+ years industry experience, 1+ years applied AI/LLM product experience; strong fluency with prompt engineering, RAG, MCP, agents, evals; systems-integration experience with APIs/webhooks; strong communication and mentorship skills.
Variance: Autonomous AI platform for managing digital fraud and adversarial risk.
Experience shipping production systems, hands‑on LLM work (prompting, RAG, evals), strong software engineering, customer communication, and CS degree or equivalent.
Staff Engineer, Agentic Intelligence - San Francisco
San Francisco or Seattle or Argentina
HybridFull Time
HomeVision: AI-powered automation for mortgage document and appraisal reviews.
Experience building and operating LLM- or agentic-based systems, designing configuration-driven/rules-engine systems, strong coding skills, experience across Go/Python backends and TypeScript frontends, and rigorous evaluation practices.
Harper: AI-native brokerage providing commercial insurance to businesses
Full-stack engineer with applied-AI fluency; experience shipping production systems, proficiency in Python or TypeScript, experience with LLM/agent features and frontend/backend work.
Phylo: An applied research lab building AI agents for biomedical discovery.
Experienced software engineer with ML/LLM experience, familiarity with LLM APIs and agent runtimes, strong quantitative judgment, production systems and distributed systems experience, and ability to design evaluations.
Twelve Labs: AI foundation models for video search and human-like understanding.
Proven experience shipping production agentic/LLM systems, owning developer-facing APIs, building auth (OAuth/OIDC/RBAC), and designing enterprise-ready platform components for reliability and scale.
Vals AI: Building enterprise benchmarks for evaluating LLM performance.
Strong engineering fundamentals, professional Python expertise, familiarity with LLMs, Git workflow experience, ability to analyze model error modes and work with cross-functional teams. In-person in San Francisco; relocation/transportation support provided.
Senior Staff Machine Learning Engineer, LLM/VLM Model Architecture & Optimization
Mountain View or San Francisco
$298k-$368k/yrOnsiteFull Time
Waymo: Autonomous driving technology for ride-hailing and logistics.
7+ YOE7+ years ML experience with large-scale model development (LLM/VLM), expertise in on-device inference and hardware acceleration, deep learning frameworks (PyTorch, JAX), large-scale training, and a master's degree in CS/EE or equivalent experience.