965 llm engineer jobs at 468 companies in American Canyon, CA

3mo
Save
Mark Applied
Hide
Distributed LLM Inference Engineer
San Francisco or Palo Alto
$170k-$247k/yr HybridFull Time
Anyscale
Anyscale: Cloud platform for scaling distributed machine learning applications.
Familiarity with running ML inference at large scale with high throughput and low latency; experience with PyTorch; solid understanding of distributed systems.
PyTorch, Ray, vLLM, TensorRT-LLM
5d
Save
Mark Applied
Hide
Senior Research Engineer, LLM Training & Post-Training
New York City or San Francisco or Seattle or London
$165k-$310k/yr HybridFull Time
Lightning AI
Lightning AI: Unified platform to build, train, and deploy AI models.
Requires significant PyTorch LLM training experience, distributed multi-GPU systems expertise, Python software engineering, experiment design, and a master's degree, PhD, or equivalent experience in a related field.
PyTorch, Python, DeepSpeed, FSDP, Megatron-LM, Hugging Face Transformers, TRL, PEFT, Lightning Fabric, CUDA, Triton, vLLM, SGLang, TensorRT, DPO, PPO, GRPO, SFT, RLHF
2w
Save
Mark Applied
Hide
Founding AI Engineer
San Francisco, California, United States
$225k-$255k/yr OnsiteFull Time
Clera
Clera: AI talent agent matching professionals with high-growth startup roles
8+ YOE8+ years engineering experience with production LLM systems, building evals and observability, experience with LLM agents and data-residency/SOC2/GDPR constraints, strong communication and product-engineer instincts.
LLM, MCP, Anthropic, Google, LangChain, LlamaIndex, Braintrust, OpenRouter
3mo
Save
Mark Applied
Hide
Machine Learning Engineer, Model Evaluations (Speech LLM) - San Francisco
San Francisco, California, United States
$180k-$270k/yr HybridFull Time
Plaud
Plaud: Develops AI-powered voice recorders and automated transcription software.
Python software engineering; building distributed systems, data pipelines, and evaluation harnesses at scale; partner with ML researchers to define benchmarks; build dashboards and monitor model health; debug mid-training anomalies; communicate results clearly.
Python, Distributed systems, Data pipelines, Evaluation harnesses, Dashboards, Weighs & Biases, MLflow
3w
Save
Mark Applied
Hide
Staff Engineer
San Francisco, California, United States
HybridFull Time
Mira Mace
Mira Mace: AI-powered healthcare advocacy and navigation for Medicare beneficiaries.
Staff-level backend/full-stack or ML engineering experience, production LLM/ML systems experience, architecture and build-vs-buy judgment, hands-on coding and code review, mentoring and technical leadership.
Python, TypeScript, FastAPI, Next.js, Postgres, LLM
3mo
Save
Mark Applied
Hide
Engineering Manager - Forward Deployed Engineering (LLM)
San Francisco or New York
$260k-$380k/yr HybridFull Time
Baseten
Baseten: Scalable infrastructure platform for deploying and serving AI models.
4+ YOE1+ MgmtLead a team of Forward Deployed Engineers; strong Python, ML inference, LLM experience; 4+ years software engineering; leadership experience; excellent communication.
Python, vLLM, TensorRT, Triton, Hugging Face, Ray Serve
1w
Save
Mark Applied
Hide
Distinguished Engineer
McLean or Richmond or New York City or San Jose or Cambridge or Plano or San Francisco
$245k-$335k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Provides credit card, banking, and auto loan services.
7+ YOEBachelor's degree; 7+ years in software engineering and solution architecture; 5+ years shipping cloud platforms; 3+ years with LLM systems, RAG, embeddings, prompt tooling, and model governance.
LLM, RAG, Java, Python, Go, JavaScript, TypeScript, Swift, BGP, Wi-Fi, SD-WAN
6d
Save
Mark Applied
Hide
Distinguished Engineer
McLean or Richmond or New York City or San Jose or Cambridge or Plano or San Francisco
$245k-$335k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: A diversified financial services providing banking and credit products.
7+ YOEBachelor's degree, 7+ years in software engineering and solution architecture, 5+ years shipping large-scale cloud platforms, and 3+ years with LLM systems including RAG, embeddings, prompt tooling, and model governance.
LLM, RAG, Java, Python, Go, JavaScript, TypeScript, Swift, BGP, Wi-Fi, SD-WAN
2w
Save
Mark Applied
Hide
Software Engineer, Artificial Intelligence/LLM (Multiple Seniority Levels)
San Carlos, California, United States
$135k-$260k/yr HybridFull Time
Beacon AI
Beacon AI: Developing an AI-powered-pilot for safer flight operations.
Experience building LLM-powered features, RAG and tool-calling, production services in Python or TypeScript, vector search and embeddings, evals/metrics, and safety/compliance for a regulated domain.
LangChain, Python, TypeScript, AWS Bedrock, OpenAI, Anthropic, OpenSearch, pgvector, Pinecone, Weaviate, S3, Aurora, DynamoDB, Triton, TensorRT-LLM
1mo
Save
Mark Applied
Hide
Senior AI Engineer
San Francisco or New York City or Vancouver or Amsterdam or London or Paris or Singapore or Tokyo
$190k-$316k/yr HybridFull Time
Amplitude
AmplitudeNASDAQ: AMPL: Digital analytics platform for understanding and optimizing customer behavior.
3+ YOE3+ years software engineering experience with 2+ years building production LLM or applied AI systems; experience designing and scaling AI products, strong product sense, and collaboration skills.
LLM, agent frameworks, evals
2w
Save
Mark Applied
Hide
GTM AI Engineer
Austin or Chicago or New York City or Salt Lake City or San Francisco or United States
$115k-$175k/yr RemoteFull Time
Gong
Gong: AI platform analyzing customer interactions to improve sales performance.
Background in software/data engineering or applied AI, experience shipping AI/LLM tools to production, systems thinking, API and workflow integration experience, strong independent delivery and stakeholder partnership skills.
Gong, Salesforce, APIs, agent frameworks, workflow orchestration, LLM
3w
Save
Mark Applied
Hide
Founding AI Engineer
San Francisco, California, United States
$225k-$255k/yr OnsiteFull Time
Mondrio
Mondrio: AI-native platform for agentic pricing and revenue management.
8+ YOE8+ years engineering experience with production LLMs, built evals/observability, shipped LLM features, strong communication, product-engineer instincts, US work authorization required.
TypeScript, React, Vercel Chat SDK, Python, FastAPI, Mongo, Atlas, GCP, Pulumi, Cloudflare Pages, FastMCP, Langfuse, Claude Code, Cursor, Vercel Eve, LangChain, LlamaIndex, Braintrust, OpenRouter
1mo
Save
Mark Applied
Hide
Senior Frontend Engineer, Platform
San Francisco or Toronto or United States or Canada
$190k-$249k/yr RemoteFull Time
EvenUp
EvenUp: AI-powered document generation and analysis for personal injury law.
5+ YOE5+ years engineering experience, building production-quality software; ownership of projects end-to-end; experience with frontend systems for deploying LLM-based algorithms; strong communication, mentoring, and cross-team collaboration skills.
LLM, Kubernetes
2w
Save
Mark Applied
Hide
Senior Forward Deployed Engineer
San Francisco, California, United States
OnsiteFull Time
Giga
Giga: Real-time voice AI agents for automated enterprise customer support
5+ YOE5+ years production software engineering; expertise in Python, APIs, CI/CD, cloud integrations, LLM agent systems; strong debugging, customer-facing communication, and mentoring skills.
Python, APIs, CI/CD, LLM
2mo
Save
Mark Applied
Hide
AI Tooling Engineer
San Francisco or Seattle or Los Angeles or New York City or United Kingdom or Ireland or Poland or Germany or Australia or North America or Europe
$200k-$260k/yr RemoteFull Time
Whatnot
Whatnot: Social marketplace for buying and selling via live streams
4+ YOE4+ years industry experience, 1+ years applied AI/LLM product experience; strong fluency with prompt engineering, RAG, MCP, agents, evals; systems-integration experience with APIs/webhooks; strong communication and mentorship skills.
Anthropic, OpenAI, Google, APIs, webhooks, LLM, RAG, MCP, evals
2mo
Save
Mark Applied
Hide
Infrastructure Engineer
Redwood City, California, United States
HybridFull Time
Vantaca
Vantaca: AI software for community association and HOA management.
8+ YOE8+ years in infrastructure/DevOps/SRE; strong cloud expertise; experience with CI/CD, PostgreSQL, Redis, APM, model serving, vector databases, GPU optimization, and LLM deployment.
PostgreSQL, Redis, APM, CI/CD, vector databases, model serving frameworks, LLM
3w
Save
Mark Applied
Hide
Senior Staff Software Engineer, Internal Tools
Redwood City, California, United States
$241k-$331k/yr HybridFull Time
Chan Zuckerberg Initiative
Chan Zuckerberg Initiative: Funds and builds technology to solve major societal challenges.
12+ YOE12+ years building production software, tech-leading small teams, full-stack engineering, cloud infrastructure/IaC, API integrations, LLMs/agents and retrieval/knowledge systems.
MCP, LLM, Infrastructure as Code (IaC), knowledge graphs, vector stores, RAG
3w
Save
Mark Applied
Hide
Agent Engineer
San Francisco, California, United States
OnsiteFull Time
Variance
Variance: Autonomous AI platform for managing digital fraud and adversarial risk.
Experience shipping production systems, hands‑on LLM work (prompting, RAG, evals), strong software engineering, customer communication, and CS degree or equivalent.
LLMs, RAG
2mo
Save
Mark Applied
Hide
Implementation Engineer (Forward Deployed), Customer Success & Integrations
San Mateo, California, United States
HybridFull Time
Parspec
Parspec: AI platform for construction product selection and procurement.
5+ YOE5+ years in forward-deployed/integration implementation engineering for B2B SaaS; degree in CS/engineering or equivalent; hands-on API/REST/JSON experience; Workato/iPaaS and ERP/CRM integrations (Eclipse, SAP, Salesforce); AI/LLM tooling; strong customer-facing and project management skills.
Workato, iPaaS, API, REST, JSON, Webhooks, Eclipse, SAP, Salesforce, AI, LLM
3w
Save
Mark Applied
Hide
Product Engineer
San Francisco, California, United States
$140k-$280k/yr OnsiteFull Time
Harper
Harper: AI-native brokerage providing commercial insurance to businesses
Full-stack engineer with applied-AI fluency; experience shipping production systems, proficiency in Python or TypeScript, experience with LLM/agent features and frontend/backend work.
Cursor, Claude Code, Windsurf, Python, TypeScript