199 ml systems engineer jobs at 55 companies in Manteca, CA

2mo
Save
Mark Applied
Hide
ML Systems Engineer
San Jose or Santa Barbara
$150k-$350k/yr OnsiteFull Time
Alpha Design AI
Alpha Design AI: AI-native EDA platform for semiconductor design and verification.
Experience with large-scale ML systems and GPU computing; strong Python and C++/CUDA skills; familiarity with vLLM, PyTorch, SGLang, Ray; experience deploying and optimizing LLMs, profiling and benchmarking inference.
Python, C++, CUDA, SGLang, vLLM, PyTorch, Ray
5d
Save
Mark Applied
Hide
ML Systems Research Engineer, RL / Inference / Agent Systems
Santa Clara, California, United States
HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Experienced ML systems engineer with strong Python and ML framework skills, experience in RL/inference systems, distributed experimentation, and GPU/infrastructure workflows; advanced degree preferred.
Python, PyTorch, JAX, TensorFlow, Kubernetes, Ray, Slurm, ROCm, HIP, CUDA
2mo
Save
Mark Applied
Hide
Lead ML Inference Engineer, Advertising
San Jose or Austin
$247k-$486k/yr HybridFull Time
Roku
RokuNASDAQ: ROKU: Operates a TV streaming platform and sells streaming hardware.
10+ YOE5+ MgmtLead the design and development of a state-of-the-art inference platform; 10+ years in distributed systems; ML serving; leadership experience.
High-performance languages, ML frameworks, GPU acceleration, HPC, Distributed systems, Inference platforms, Monitoring tooling
2w
Save
Mark Applied
Hide
Senior Principal Applied ML Engineer
San Jose, California, United States
$154k-$286k/yr OnsiteFull Time
Cadence Design Systems
Cadence Design SystemsNASDAQ: CDNS: Develops computational software and hardware for electronic system design.
5+ YOEAdvanced ML and software engineering experience building production-quality agentic systems, LLM deployment, evaluation frameworks, retrieval pipelines, and observability for large-scale EDA workflows.
Xcelium, Jasper, Palladium, Protium, Helium, EDA, Large language models (LLMs), Retrieval-augmented generation (RAG)
1mo
Save
Mark Applied
Hide
Senior AI Systems Engineer
San Jose, California, United States
$160k-$180k/yr OnsiteFull Time
Archer
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
3+ YOEBS/MS/PhD in CS or related, 3+ years building AI/ML systems or HPC/ML infrastructure, experience with multi-cloud (AWS), Docker, Kubernetes, MLflow, vLLM/SGLang, SkyPilot, SQL/NoSQL and Parquet.
MLflow, Nebius AI Cloud, Docker, Kubernetes, SkyPilot, vLLM, SGLang, SQL, NoSQL, Parquet, AWS
1mo
Save
Mark Applied
Hide
Staff Agentic ML Engineer - Photoshop
San Jose or Waltham or San Francisco or Seattle or Los Angeles or New York City or California
$190k-$346k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
8+ YOEMaster's or PhD in CS/ML/Data Science, 8+ years ML experience, expertise in LLMs, fine-tuning (SFT, RLHF, DPO, PEFT/LoRA), agent system development, evaluation frameworks; proficiency with PyTorch and agentic frameworks; strong communication and problem-solving.
PyTorch, LangChain, LangGraph, MCP, A2A, Agent Development Kit (ADK), Claude Code, Codex, Cursor
1mo
Save
Mark Applied
Hide
AI/ML Lead Engineer
Stamford or San Ramon or San Mateo
$180k-$212k/yr HybridFull Time
Franklin Templeton
Franklin TempletonNYSE: BEN: Global investment firm providing asset and wealth management services.
5+ YOE5+ years software engineering experience including 2+ years deploying LLM/GenAI or agent-based systems; expert Python; experience with LangChain/OpenAI tool calling, vector DBs (Pinecone, FAISS), RAG, observability, and distributed services.
LangChain, OpenAI, Python, Pinecone, FAISS, Docker, Kubernetes, AWS, Azure, GCP
1mo
Save
Mark Applied
Hide
AI/ML Engineering Manager
Milpitas, California, United States
$170k-$210k/yr OnsiteFull Time
Corsair
CorsairNASDAQ: CRSR: Sells high-performance gaming peripherals and PC hardware components.
6+ YOE1+ Mgmt6+ years in AI/ML engineering with 1+ years technical lead experience; strong Python and SQL; experience with LLM frameworks, workflow orchestration, data engineering, and MLOps; ability to architect agentic AI systems and lead engineers.
Python, SQL, LangChain, LangGraph, CrewAI, Strands, AWS Bedrock, Azure AI Foundry, Microsoft Fabric, Temporal, Airflow, Step Functions, dbt, Dagster, Spark, MLflow, W&B, Docker, Kubernetes, CI/CD, LoRA, PEFT
3w
Save
Mark Applied
Hide
Edge ML Software Engineer (System Modeling-PICO) - San Jose
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
3+ YOEBachelor's in CS/EE/CE or equivalent, 3+ years in computer architecture/system modeling, strong C/C++ and System C skills, knowledge of memory, cache, DMA, tiling, and ML accelerator modeling.
C, C++, System C
2h
Save
Mark Applied
Hide
ML Infra Engineer Graduate (Ads Infra) - 2027 Start
San Jose, California, United States
$128k-$256k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Bachelor's or Master's in computer science or related field; proficiency in C++, Python, Go or Java; strong knowledge of data structures, algorithms, ML, distributed systems and large-scale data processing.
C++, Python, Go, Java
2w
Save
Mark Applied
Hide
AI Systems Performance Engineer - New Graduate
San Jose, California, United States
$135k-$165k/yr OnsiteFull Time
SambaNova Systems
SambaNova Systems: Develops custom AI hardware and software for enterprise computing.
Bachelor's degree in CS/EE/CE or related, strong Python/C++ skills, foundations in algorithms/architecture/OS/parallel computing, familiarity with deep learning and at least one ML framework, ability to analyze and optimize model performance.
Python, C++, PyTorch, TensorFlow, JAX, CUDA, Triton, OpenCL, vLLM, DeepSpeed, Megatron, TensorRT, SambaNova Suite
1mo
Save
Mark Applied
Hide
Software Engineer, ML Engineering
San Jose or Hong Kong
RemoteFull Time
Nex
Nex: Gaming system that turns body movement into interactive play.
3+ YOE3+ years building production ML systems or training infrastructure; proficiency in Python and one systems language (C++, C#, Java, Rust, or Go); experience with PyTorch or TensorFlow; building training pipelines, data workflows, and model deployment.
Python, C++, C#, Java, Rust, Go, PyTorch, TensorFlow, Docker
1w
Save
Mark Applied
Hide
Sr. Engineer, AI/ML Apps Engineering (AI2525)
San Jose, California, United States
$121k-$193k/yr OnsiteFull Time
SiMa.ai
SiMa.ai: Designs low-power machine learning chips for edge computing applications.
6+ YOE6+ years in customer-facing applications/field engineering for AI, embedded, robotics, or edge systems; strong Python and C++; experience with PyTorch, TensorFlow, ONNX Runtime, OpenCV, and embedded Linux; customer-facing and presentation skills.
SiMa Modalix platform, Python, C++, PyTorch, TensorFlow, ONNX Runtime, OpenCV, Embedded Linux
3w
Save
Mark Applied
Hide
HPE Labs – AI/ML Research Scientist III
Milpitas, California, United States
$137k-$277k/yr OnsiteFull Time
Hewlett Packard Enterprise
Hewlett Packard EnterpriseNYSE: HPE: Provides edge-to-cloud IT infrastructure and platform services.
PhD in CS/EE or related with strong ML research background, expertise in LLMs and reinforcement learning, proficiency in PyTorch and Python, GPU/system optimization experience, strong publication record and mentoring ability.
PyTorch, Python
2d
Save
Mark Applied
Hide
Software Engineer III, AI/ML, Enterprise Engineering, Core
San Jose, California, United States
$147k-$210k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
2+ YOEBachelor's degree or equivalent,2 years programming in Python or C++,1 year applied ML,experience with distributed systems and data structures;concurrency and accessibility experience preferred.
Python, C++
1mo
Save
Mark Applied
Hide
ML Researcher
San Jose, California, United States
$150k-$290k/yr OnsiteFull Time
Rivet Industries
Rivet Industries: Building integrated task systems for frontline industrial and defense operators.
5+ YOEBS + 5+ years (or MS + 2+ years) in ML research or applied ML engineering; proficiency in Python and C++; experience with PyTorch/TensorFlow, ML pipelines, model deployment, and model optimization for edge/embedded systems.
Python, C++, PyTorch, TensorFlow, TensorFlow Lite, CUDA, AWS, GCP, Azure, Docker, Kubernetes
1mo
Save
Mark Applied
Hide
Lead System Engineer
San Jose, California, United States
$132k-$194k/yr HybridFull Time
Sony
SonyNYSE: SONY: Sells consumer electronics, video games, and entertainment media.
1+ YOEMaster's in Electrical Engineering (or equivalent) plus 1 year firmware/applications experience; skills in IoT architecture, edge vision AI, computer vision, low-level C/C++, C++/Python OOP, firmware/embedded engineering, sensor drivers, and ML fundamentals.
C/C++, C++, Python
5d
Save
Mark Applied
Hide
Senior Lead AI Engineer (MLXT)
McLean or San Francisco or Cambridge or San Jose or New York
$230k-$286k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's in CS/AI/EE/CE+6yrs or Master's+4yrs experience in AI/ML, 6+ years programming in Python/Go/Scala/Java, cloud AI deployment experience, leadership and systems engineering skills.
Python, Go, Scala, Java, C++, C#, Golang, AWS, Google Cloud, Azure
3mo
Save
Mark Applied
Hide
Staff Machine Learning Engineer - Agentic Models, LLM, RAG, GenAI
Santa Clara, California, United States
$232k-$310k/yr HybridFull Time
Eightfold
Eightfold: Global AI-native talent intelligence platform provider.
6+ YOELead AI/ML engineer with 6+ years in ML, Gen AI, LLMs; strong Python, TensorFlow/PyTorch; AWS, Docker, Kubernetes; expert in agentic AI and distributed systems.
Python, TensorFlow, PyTorch, AWS, Docker, Kubernetes
1mo
Save
Mark Applied
Hide
Distinguished Engineer - AI (San Jose, CA, US, 95128)
San Jose, California, United States
$266k-$396k/yr OnsiteFull Time
NetApp
NetAppNASDAQ: NTAP: Sells enterprise data storage and cloud management software.
15+ YOE15+ years building low-latency, fault-tolerant distributed systems and AI/ML inference platforms; expertise with inference engines, model optimization, storage for AI, RDMA/DPDK, and Kubernetes-based orchestration.
TensorRT, vLLM, ONNX Runtime, Triton, RDMA, DPDK, Kubernetes