432 llm engineer jobs at 98 companies in Patterson, CA

PromotedHiringCafe
Founding Machine Learning / AI Search Engineer
Cupertino, CA, US
$160k-$310k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Build the ML and AI search behind HiringCafe — ranking, recommenders, retrieval, and LLM agents that surface jobs people would never find on their own.
Python, PyTorch, Elasticsearch, LLMs
2w
Save
Mark Applied
Hide
Principal LLM Inference Engineer
Santa Clara, California, United States
$195k-$285k/yr HybridFull Time
d-Matrix: Develops high-performance semiconductor chips for generative AI inference.
10+ YOEBachelor's in CS/EE (or equivalent) with 10+ years experience (Master/PhD with 6+ years preferred); strong Python and C/C++; experience optimizing LLM inference, quantization, batching, GPU kernel programming and contributor-level work on inference frameworks.
Python, C, C++, vLLM, SGLang, TensorRT-LLM, ONNX Runtime, CUDA, Triton, JAX
1w
Save
Mark Applied
Hide
Senior/Staff LLM Application Engineer - Data Application
San Jose, California, United States
$213k-$450k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Experience with data products and LLM application development, strong coding skills in Python, knowledge of prompt engineering, retrieval and benchmarking, and ability to analyze user feedback.
Python
2mo
Save
Mark Applied
Hide
Principal High-Performance LLM Training Engineer
Santa Clara, California, United States
$272k-$431k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
12+ YOEMS/PhD or equivalent with 12+ years experience; principal-level impact in large-scale AI training, GPU and distributed systems performance; experience with transformer/LLM workloads, distributed training techniques, profiling and benchmarking; strong communication and leadership.
PyTorch, JAX, NeMo, NeMo RL, CUDA
2mo
Save
Mark Applied
Hide
Principal High-Performance LLM Training Engineer
Santa Clara, California, United States
$272k-$431k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
12+ YOEMS or PhD in CS/EE/CE, with 12+ years of relevant work; strong impact in large-scale AI training, GPU performance, distributed systems, or similar.
PyTorch, JAX, NeMo, NeMo RL, CUDA, Profiling tools
1w
Save
Mark Applied
Hide
LLM AIOps Development Engineer - Data Center Networking
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Deep data-center networking and Linux networking knowledge; proficiency in Golang or Python; experience with telemetry, observability, big-data pipelines, and LLM/AIOps approaches; familiarity with protocols like EVPN/VXLAN and BGP/OSPF.
gNMI, Netconf, IPFIX, NetFlow, SNMP, Golang, Python, Docker, Kubernetes, CI/CD, Kafka, Flink, ClickHouse, TSDB, Prometheus, OpenTelemetry, Neo4j, SONiC, P4, eBPF, DPDK, RDMA, RoCE
2mo
Save
Mark Applied
Hide
Staff Machine Learning Engineer - Agentic Models, LLM, RAG, GenAI
Santa Clara, California, United States
$232k-$310k/yr HybridFull Time
Eightfold
Eightfold: Global AI-native talent intelligence platform provider.
6+ YOELead AI/ML engineer with 6+ years in ML, Gen AI, LLMs; strong Python, TensorFlow/PyTorch; AWS, Docker, Kubernetes; expert in agentic AI and distributed systems.
Python, TensorFlow, PyTorch, AWS, Docker, Kubernetes
2mo
Save
Mark Applied
Hide
Lead Machine Learning Engineer - Agentic Models, LLM, RAG, GenAI
Santa Clara, California, United States
$193k-$258k/yr HybridFull Time
Eightfold.ai
Eightfold.ai: AI-native platform for talent management and workforce optimization.
5+ YOESenior ML engineer with expertise in AI agents, LLMs, distributed systems; 5-7+ years of experience; strong Python and ML frameworks; AWS; Docker/Kubernetes; RAG/GenAI experience.
Python, TensorFlow, PyTorch, AWS, Docker, Kubernetes, Kafka, AWS SQS, LangGraph, CrewAI, AutoGen, Pinecone, pgvector, vLLM, TensorRT-LLM
1mo
Save
Mark Applied
Hide
Software Engineer 2 - LLM Inference
Vancouver or San Jose or Durham or Mexico City or Bengaluru or Pune or Hoofddorp or Belgrade or Barcelona or Singapore or Sydney or Tokyo
$129k-$193k/yr HybridFull Time
Nutanix
NutanixNASDAQ: NTNX: Sells cloud software and hyperconverged infrastructure for enterprises.
2+ YOE2–5 years product development experience; strong programming fundamentals; experience with Docker, Kubernetes, Go, Python, CI/CD, distributed systems, datacenter design, OS internals, and high-performance system tuning; Bachelor's/Masters in CS or equivalent.
Docker, Kubernetes, Go, Python, CI/CD, TensorFlow, PyTorch, Nutanix Cloud Platform for AI, GPUs
2mo
Save
Mark Applied
Hide
Lead AI Engineer (AI Foundations, LLM Customization and Finetuning)
Cambridge or McLean or San Jose or New York City
$197k-$225k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's in CS/AI/EE/CE with 4+ years or Master's with 2+ years in AI/ML; 4+ years Python/Go/Scala/Java
AWS, Hugging Face, VectorDBs, Nemo Guardrails, PyTorch, Python, Go, Scala, Java, C++
1mo
Save
Mark Applied
Hide
Staff Design Engineer, Insights
Minnesota or San Francisco or San Jose
$159k-$302k/yr RemoteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years at the intersection of design and engineering; hands-on experience building and shipping LLM-powered apps, internal tools, and high-fidelity prototypes; strong craft, collaboration, and product judgment.
LLM, Vega, D3, Adobe Acrobat Studio, Adobe Express, Adobe Firefly, Creative Cloud, Adobe Experience Platform, Adobe Experience Manager, GenStudio
3w
Save
Mark Applied
Hide
Staff Engineer, Agents
Santa Clara, California, United States
$160k-$200k/yr HybridFull Time
LeanData
LeanData: Automates lead-to-account matching and routing for B2B revenue teams.
4+ YOE4+ years building production systems (2+ years shipping LLM/agent systems), strong Python (TypeScript/Go a plus), experience with agent frameworks, RAG/memory, distributed-systems fundamentals, and shipping customer-facing LLM products.
Python, TypeScript, Go, LangGraph, agent SDKs, LLM, RAG, Salesforce APIs, Bulk 2.0, Composite, Pub/Sub, MCP (Model Context Protocol), Promptfoo, Braintrust, Langfuse, Inngest, Temporal, Postgres, RLS
2mo
Save
Mark Applied
Hide
Research Engineer - The Diffusion LLM Team
Sunnyvale, California, United States
OnsiteFull Time
Institute of Foundation Models
Institute of Foundation Models: Develops open-source frontier-class AI foundation models and research.
MSc/PhD in ML or CS; hands-on large-model training; transformer architectures; diffusion models knowledge a plus; publications/open-source contributions; independent yet collaborative.
PyTorch, TensorFlow, JAX, Distributed Training
1mo
Save
Mark Applied
Hide
DFx Mechanical Engineer III
Santa Clara, California, United States
$111k-$152k/yr OnsiteFull Time
Applied Materials
Applied MaterialsNASDAQ: AMAT: Manufacturers of equipment for semiconductor and display production.
DFx engineering experience with GD&T and ASME Y14.5, tolerance stack-up analysis, manufacturability/cost-reduction focus, familiarity with AI/LLM tools (Copilot, ChatGPT), and ability to participate in cross-functional design and packaging reviews.
Microsoft Copilot, ChatGPT, LLM
4w
Save
Mark Applied
Hide
Applied AI Engineer, Silicon Engineering
San Jose, California, United States
$150k-$275k/yr OnsiteFull Time
Etched
Etched: Designs specialized AI chips optimized for transformer architectures.
Experience building and shipping LLM-based agents and AI tooling, strong Python proficiency, agent orchestration and tool integration, eval-driven approach, ability to work across stacks and embed with hardware teams.
LLM, Python, C++, Rust, Docker, Slurm, Ray, Tcl, SystemVerilog, UVM, FPGA, EDA, CI/CD, Claude Code, MCP
2mo
Save
Mark Applied
Hide
Senior AI Engineer
San Jose or California
$187k-$215k/yr HybridFull Time
KlearNow.AI
KlearNow.AI: AI platform for digital customs brokerage and logistics management.
4+ YOE4+ years software engineering and infrastructure experience, production LLM/agent experience, Python and TypeScript proficiency, Docker/Kubernetes, cloud/distributed systems, ability to work hybrid in San Jose and authorized to work in the U.S.
Agno, LangChain, LlamaIndex, Semantic Kernel, Claude Code, Cursor, AugmentCode, Python, TypeScript, Docker, Kubernetes, CI/CD, LLMs, RAG
1mo
Save
Mark Applied
Hide
Security Engineer
San Francisco or San Jose or Bellevue
$266k-$395k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years security engineering experience; hands-on Linux; experience building detections, incident response, vulnerability remediation; production-grade coding in Python or Go; cross-domain security experience and collaboration skills.
SIEM, EDR, SOAR, IDS/IPS, Linux, Python, Go, KVM, Hyper-V, Xen, LLM
1mo
Save
Mark Applied
Hide
Motion Planning Engineer
San Jose, California, United States
OnsiteFull Time
Imagry
Imagry: Mapless autonomous driving software for vehicle manufacturers.
3+ YOEM.Sc. (or equivalent) in a quantitative field, 3+ years algorithm engineering experience, expertise in motion planning/decision making/optimization/Reinforcement Learning or LLM fine-tuning, and proficiency in Python and C++.
Python, C++
3w
Save
Mark Applied
Hide
Staff AI Engineer - Conversational & Agentic AI
Santa Clara, California, United States
$176k-$308k/yr HybridFull Time
ServiceNow
ServiceNowNYSE: NOW: Provides a cloud platform for automating enterprise digital workflows.
7+ YOE7+ years building production software, hands-on experience shipping generative AI products, deep knowledge of LLM behavior, prompt engineering, eval engineering, distributed systems, reliability, and production observability.
LangChain, LlamaIndex
4d
Save
Mark Applied
Hide
Software Engineer
Santa Clara, California, United States
$149k-$224k/yr OnsiteFull Time
Pure Storage
Pure StorageNYSE: PSTG: Provides all-flash enterprise data storage and management solutions.
3+ YOE3+ years software engineering building backend or distributed systems, strong Python and REST API skills, experience with LLMs, RAG, prompt engineering, AWS deployments, security and observability.
Python, REST APIs, LLM, RAG, embeddings, vector database, Pinecone, Weaviate, OpenSearch, AWS, EC2, S3, Lambda, RDS, Docker, Kubernetes, CI/CD, ServiceNow, Snowflake, Datadog, Prometheus
1w
Save
Mark Applied
Hide
SASE Cloud Software Engineer - AI & Backend
San Jose, California, United States
$137k-$277k/yr HybridFull Time
Hewlett Packard Enterprise
Hewlett Packard EnterpriseNYSE: HPE: Provides global edge-to-cloud technology solutions and IT infrastructure services.
5+ YOE5+ years Java/Python backend development, 5+ years designing REST APIs/microservices, 2+ years building or integrating AI/LLM applications, experience with SQL/NoSQL and cloud platforms, familiarity with JavaScript/React/Node.js, BS in a technical discipline.
Java, Python, REST APIs, LLM, Retrieval-Augmented Generation (RAG), SQL, NoSQL, AWS, Azure, JavaScript, React, Node.js