952 llm engineer jobs at 488 companies in Vallejo, CA

4w
Save
Mark Applied
Hide
LLM Inference Engineer
San Francisco or United States
RemoteFull Time
NEAR AI
NEAR AI: Building private, verifiable infrastructure for autonomous AI agents.
Expert in LLM inference and serving systems, optimizing throughput/latency/cost for open-source LLMs, deep GPU architecture knowledge, and experience with PyTorch, Triton, CUDA and inference engines like vLLM/SGLang/TensorRT.
SGLang, vLLM, TensorRT, PyTorch, Triton, CuTe, CUDA
3mo
Save
Mark Applied
Hide
Distributed LLM Inference Engineer
San Francisco or Palo Alto
$170k-$247k/yr HybridFull Time
Anyscale
Anyscale: Cloud platform for scaling distributed machine learning applications.
Familiarity with running ML inference at large scale with high throughput and low latency; experience with PyTorch; solid understanding of distributed systems.
PyTorch, Ray, vLLM, TensorRT-LLM
1w
Save
Mark Applied
Hide
Senior Machine Learning Engineer, LLM Inference Optimization
Palo Alto or California
$195k-$262k/yr OnsiteFull Time
Nebius
NebiusNasdaq: NBIS: Builds cloud infrastructure and software for artificial intelligence development.
Expert Python and PyTorch skills, hands-on LLM/VLM inference deployment and optimization, knowledge of modern inference stacks, quantitative reasoning about latency/throughput/cost, and strong communication.
Python, PyTorch, vLLM, SGLang, TensorRT-LLM, Triton Inference Server, NVIDIA Dynamo, Ray Serve, KServe, CUDA, FlashInfer, LMCache, Ray
2d
Save
Mark Applied
Hide
Founding AI Engineer
San Francisco, California, United States
$225k-$255k/yr OnsiteFull Time
Clera
Clera: AI talent agent matching professionals with high-growth startup roles
8+ YOE8+ years engineering experience with production LLM systems, building evals and observability, experience with LLM agents and data-residency/SOC2/GDPR constraints, strong communication and product-engineer instincts.
LLM, MCP, Anthropic, Google, LangChain, LlamaIndex, Braintrust, OpenRouter
2mo
Save
Mark Applied
Hide
Software Engineering Manager, LLM Training
Mountain View, California, United States
$170k-$277k/yr HybridFull Time
LinkedInNASDAQ: MSFT: Professional social network for career development and job recruitment.
1+ Mgmt5+ years software engineering; 1+ year management; experience with LLMs and distributed systems; strong leadership and strategic planning
PyTorch, CUDA, Megatron, Hugging Face, vLLM, Liger, Ray, SGLang, VERL, NCCL, TorchDistributed
2mo
Save
Mark Applied
Hide
Machine Learning Engineer, Model Evaluations (Speech LLM) - San Francisco
San Francisco, California, United States
$180k-$270k/yr HybridFull Time
Plaud
Plaud: Develops AI-powered voice recorders and automated transcription software.
Python software engineering; building distributed systems, data pipelines, and evaluation harnesses at scale; partner with ML researchers to define benchmarks; build dashboards and monitor model health; debug mid-training anomalies; communicate results clearly.
Python, Distributed systems, Data pipelines, Evaluation harnesses, Dashboards, Weighs & Biases, MLflow
6d
Save
Mark Applied
Hide
Staff Engineer
San Francisco, California, United States
HybridFull Time
Mira Mace
Mira Mace: AI-powered healthcare advocacy and navigation for Medicare beneficiaries.
Staff-level backend/full-stack or ML engineering experience, production LLM/ML systems experience, architecture and build-vs-buy judgment, hands-on coding and code review, mentoring and technical leadership.
Python, TypeScript, FastAPI, Next.js, Postgres, LLM
2mo
Save
Mark Applied
Hide
Engineering Manager - Forward Deployed Engineering (LLM)
San Francisco or New York
$260k-$380k/yr HybridFull Time
Baseten
Baseten: Scalable infrastructure platform for deploying and serving AI models.
4+ YOE1+ MgmtLead a team of Forward Deployed Engineers; strong Python, ML inference, LLM experience; 4+ years software engineering; leadership experience; excellent communication.
Python, vLLM, TensorRT, Triton, Hugging Face, Ray Serve
3d
Save
Mark Applied
Hide
Software Engineer, Artificial Intelligence/LLM (Multiple Seniority Levels)
San Carlos, California, United States
$135k-$260k/yr HybridFull Time
Beacon AI
Beacon AI: Developing an AI-powered-pilot for safer flight operations.
Experience building LLM-powered features, RAG and tool-calling, production services in Python or TypeScript, vector search and embeddings, evals/metrics, and safety/compliance for a regulated domain.
LangChain, Python, TypeScript, AWS Bedrock, OpenAI, Anthropic, OpenSearch, pgvector, Pinecone, Weaviate, S3, Aurora, DynamoDB, Triton, TensorRT-LLM
1mo
Save
Mark Applied
Hide
Infrastructure Engineer
Menlo Park, California, United States
OnsiteFull Time
Shakudo
Shakudo: Develops an operating system for enterprise AI applications.
8+ YOE8+ years engineering experience, 5+ years Kubernetes operation, proficiency in Rust, experience with production infrastructure (physical servers, GPU/DGX clusters), CI/CD, security hardening, observability, and LLM/AI infrastructure.
Kubernetes, Rust, CI/CD, DGX, GPU, LLM, ETL
1mo
Save
Mark Applied
Hide
Growth Engineer (GTM / Go-To-Market Engineer)
Walnut Creek or California or United States
$80k-$150k/yr HybridFull Time
GatherUp
GatherUp: Software platform for customer review and reputation management.
3+ YOE3+ years in GTM engineering/rev ops/growth engineering; coding in Python, JavaScript, or SQL; experience with AI voice tools, Clay, ZoomInfo, Apollo, n8n, Zapier, Salesforce/HubSpot; building agentic AI/LLM workflows; strong analytical and cross-functional communication skills.
Python, JavaScript, SQL, ElevenLabs, Clay, ZoomInfo, Apollo, n8n, Zapier, Salesforce, HubSpot, LLM
3d
Save
Mark Applied
Hide
GTM AI Engineer
Austin or Chicago or New York City or Salt Lake City or San Francisco or United States
$115k-$175k/yr RemoteFull Time
Gong
Gong: AI platform analyzing customer interactions to improve sales performance.
Background in software/data engineering or applied AI, experience shipping AI/LLM tools to production, systems thinking, API and workflow integration experience, strong independent delivery and stakeholder partnership skills.
Gong, Salesforce, APIs, agent frameworks, workflow orchestration, LLM
1mo
Save
Mark Applied
Hide
Staff Design Engineer, Insights
Minnesota or San Francisco or San Jose
$159k-$302k/yr RemoteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years at the intersection of design and engineering; hands-on experience building and shipping LLM-powered apps, internal tools, and high-fidelity prototypes; strong craft, collaboration, and product judgment.
LLM, Vega, D3, Adobe Acrobat Studio, Adobe Express, Adobe Firefly, Creative Cloud, Adobe Experience Platform, Adobe Experience Manager, GenStudio
1mo
Save
Mark Applied
Hide
Senior Frontend Engineer, Platform
San Francisco or Toronto or United States or Canada
$190k-$249k/yr RemoteFull Time
EvenUp
EvenUp: AI-powered document generation and analysis for personal injury law.
5+ YOE5+ years engineering experience, building production-quality software; ownership of projects end-to-end; experience with frontend systems for deploying LLM-based algorithms; strong communication, mentoring, and cross-team collaboration skills.
LLM, Kubernetes
6d
Save
Mark Applied
Hide
Senior Forward Deployed Engineer
San Francisco, California, United States
OnsiteFull Time
Giga
Giga: Real-time voice AI agents for automated enterprise customer support
5+ YOE5+ years production software engineering; expertise in Python, APIs, CI/CD, cloud integrations, LLM agent systems; strong debugging, customer-facing communication, and mentoring skills.
Python, APIs, CI/CD, LLM
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Backend/Infra
Palo Alto, California, United States
$185k-$300k/yr HybridFull Time
Pika
Pika: AI-powered platform for generating and editing professional videos
5+ YOE5+ years software engineering experience, 2+ years with LLM/agentic systems preferred; deep backend proficiency (Node.js, Python, Go), distributed systems, AWS/GCP, Kubernetes, CI/CD, real-time architectures, and LLM integration skills.
Node.js, Python, Go, Express, FastAPI, TypeScript, AWS, GCP, Kubernetes, CI/CD, WebSocket, Claude, GPT, Gemini, LLM, LangChain, CrewAI, AutoGPT, SQL, NoSQL
1mo
Save
Mark Applied
Hide
AI Tooling Engineer
San Francisco or Seattle or Los Angeles or New York City or United Kingdom or Ireland or Poland or Germany or Australia or North America or Europe
$200k-$260k/yr RemoteFull Time
Whatnot
Whatnot: Social marketplace for buying and selling via live streams
4+ YOE4+ years industry experience, 1+ years applied AI/LLM product experience; strong fluency with prompt engineering, RAG, MCP, agents, evals; systems-integration experience with APIs/webhooks; strong communication and mentorship skills.
Anthropic, OpenAI, Google, APIs, webhooks, LLM, RAG, MCP, evals
1mo
Save
Mark Applied
Hide
Infrastructure Engineer
Redwood City, California, United States
HybridFull Time
Vantaca
Vantaca: AI software for community association and HOA management.
8+ YOE8+ years in infrastructure/DevOps/SRE; strong cloud expertise; experience with CI/CD, PostgreSQL, Redis, APM, model serving, vector databases, GPU optimization, and LLM deployment.
PostgreSQL, Redis, APM, CI/CD, vector databases, model serving frameworks, LLM
1w
Save
Mark Applied
Hide
Senior Staff Software Engineer, Internal Tools
Redwood City, California, United States
$241k-$331k/yr HybridFull Time
Chan Zuckerberg Initiative
Chan Zuckerberg Initiative: Funds and builds technology to solve major societal challenges.
12+ YOE12+ years building production software, tech-leading small teams, full-stack engineering, cloud infrastructure/IaC, API integrations, LLMs/agents and retrieval/knowledge systems.
MCP, LLM, Infrastructure as Code (IaC), knowledge graphs, vector stores, RAG
1mo
Save
Mark Applied
Hide
Implementation Engineer (Forward Deployed), Customer Success & Integrations
San Mateo, California, United States
HybridFull Time
Parspec
Parspec: AI platform for construction product selection and procurement.
5+ YOE5+ years in forward-deployed/integration implementation engineering for B2B SaaS; degree in CS/engineering or equivalent; hands-on API/REST/JSON experience; Workato/iPaaS and ERP/CRM integrations (Eclipse, SAP, Salesforce); AI/LLM tooling; strong customer-facing and project management skills.
Workato, iPaaS, API, REST, JSON, Webhooks, Eclipse, SAP, Salesforce, AI, LLM