765 ml systems engineer jobs at 176 companies in Gilroy, CA

PromotedHiringCafe
ML Engineer - Inference & Model Deployment
Cupertino, CA, US
$250k-$310k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Turn powerful AI and ML models into fast, reliable production systems. Own inference latency, throughput, model-serving architecture, multi-GPU systems, and production deployment for millions of users.
Python, PyTorch, vLLM, SGLang, TensorRT, LLMs
PromotedHiringCafe
Founding Machine Learning / AI Search Engineer
Cupertino, CA, US
$160k-$310k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Build the ML and AI search behind HiringCafe — ranking, recommenders, retrieval, and LLM agents that surface jobs people would never find on their own.
Python, PyTorch, Elasticsearch, LLMs
PromotedHiringCafe
Founding Backend / Infra Engineer
Cupertino, CA, US
$160k-$300k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Own the crawlers, pipelines, and infrastructure powering a real-time job search engine. Strong Node.js and Python fundamentals; bonus points for security and reverse-engineering chops.
Node.js, Python, Elasticsearch, Redis
2mo
Save
Mark Applied
Hide
ML Systems Engineer
Menlo Park, California, United States
$300k-$400k/yr OnsiteFull Time
Periodic Labs
Periodic Labs: Builds autonomous laboratories for AI-driven scientific discovery.
Experience with large-scale ML systems, GPU clusters, distributed training, and performance profiling; proficiency in RDMA, CUDA, and scheduling across Ray, Slurm, or Kubernetes.
CUDA, RDMA, NVLink, Ray, Slurm, Kubernetes, SGLang, Megatron-LM, vLLM
1mo
Save
Mark Applied
Hide
ML Systems Engineer
San Jose or Santa Barbara
$150k-$350k/yr OnsiteFull Time
Alpha Design AI
Alpha Design AI: AI-native EDA platform for semiconductor design and verification.
Experience with large-scale ML systems and GPU computing; strong Python and C++/CUDA skills; familiarity with vLLM, PyTorch, SGLang, Ray; experience deploying and optimizing LLMs, profiling and benchmarking inference.
Python, C++, CUDA, SGLang, vLLM, PyTorch, Ray
1mo
Save
Mark Applied
Hide
ML and Agentic Systems Engineer
Santa Clara, California, United States
$224k-$431k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
12+ YOE12+ years building ML systems and software platforms; expert Python and PyTorch; experience with ML pipelines, evaluation, developer tooling, and LLM/agentic systems; BS/MS in CS, Engineering, or equivalent experience.
Python, PyTorch, LLM, AuthN, AuthZ, IAM
1mo
Save
Mark Applied
Hide
Staff ML Engineer
Palo Alto or Seattle or Paris
$205k-$330k/yr RemoteFull Time
Docker
Docker: Provides a platform for building, sharing, and running containerized applications.
8+ YOE8+ years professional software engineering experience, 5+ years applied ML experience, bachelor's in CS/Engineering or equivalent, experience shipping ML systems, LLM/agent experience, on-call participation possible.
Docker Desktop, Docker Hub, Docker Scout, Agentic Platform, MCP
2mo
Save
Mark Applied
Hide
Software Engineer, Systems ML
Sunnyvale or Menlo Park
$184k-$257k/yr OnsiteFull Time
Meta
MetaNASDAQ: META: Develops social networking platforms and virtual reality technologies.
2+ YOEBachelor's degree in Computer Science or related field; experience ML infra domains; experience in C/C++ or Python
C/C++, Python, PyTorch, Machine Learning Frameworks, High Performance Computing
1w
Save
Mark Applied
Hide
Senior ML Engineer, Perception
Palo Alto, California, United States
$179k-$223k/yr OnsiteFull Time
Rivian
RivianNASDAQ: RIVN: Designs and manufactures electric vehicles and charging networks.
5+ YOE5+ years building and scaling ML solutions for auto-labeling and AV perception; strong Python, perception pipeline and system engineering experience; BS/MS/PhD in CS/Robotics/Electrical Engineering or related.
Python
3w
Save
Mark Applied
Hide
Principal AI/ML Engineer
Sunnyvale or Austin or Detroit or Warren or Milford or Mountain View or United States
$296k-$424k/yr HybridFull Time
General Motors
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
MS or PhD in CS/Robotics/ML or related; experience leading technical teams delivering production ML systems; deep expertise in robotics planning/control, imitation or reinforcement learning, generative models, and large-scale ML; strong software engineering skills in Python and C++.
Python, C++
3d
Save
Mark Applied
Hide
ML Infrastructure Engineer
Palo Alto, California, United States
$180k-$440k/yr OnsiteFull Time
xAI
xAI: Develops advanced artificial intelligence systems to understand the universe.
2+ YOE2+ years building large-scale production systems or ML infrastructure; degree in CS or related field or equivalent experience; strong Python and compiled-language skills; experience with GPU and distributed systems.
Python, C++, Rust, JAX, PyTorch, NVIDIA drivers, CUDA, Linux, Slurm, Puppet, Ansible
1w
Save
Mark Applied
Hide
Senior ML Infra Engineer
Santa Clara, California, United States
OnsiteFull Time
MaxInsights
MaxInsights: Provides robot data collection for physical AI development.
Experience building production ML training and deployment systems, strong software engineering and infra fundamentals, PyTorch experience, HPC/GPU knowledge, and strong communication and product sense.
Python, PyTorch, Docker, CI/CD, GPUs, CUDA, vLLM, TensorRT, Triton
2d
Save
Mark Applied
Hide
ML Systems Engineer, Large-Scale Model Training & RL Infrastructure
Palo Alto, California, United States
$195k-$262k/yr OnsiteFull Time
Nebius
NebiusNasdaq: NBIS: Builds cloud infrastructure and software for artificial intelligence development.
Strong Python and PyTorch skills, hands-on distributed model training and GPU cluster experience, debugging across NCCL/CUDA/PyTorch/Ray, and quantitative reasoning about throughput, utilization, memory, and cost.
Python, PyTorch, Megatron-LM, DeepSpeed, PyTorch FSDP/DTensor, Ray, verl, slime, AReaL, OpenRLHF, NCCL, CUDA, Triton, Nsight, InfiniBand, RDMA, RoCE, Slurm, Kubernetes
2w
Save
Mark Applied
Hide
Founding Engineer - ML Research
Mountain View, California, United States
$220k-$300k/yr OnsiteFull Time
Clera
Clera: AI talent agent matching professionals with high-growth startup roles
3+ YOE3+ years ML research or ML systems experience; strong Python and PyTorch/JAX/TensorFlow skills; experience with Transformers, diffusion models, RLHF, and building reproducible research-to-production pipelines.
Python, PyTorch, JAX, TensorFlow
3w
Save
Mark Applied
Hide
Principal Engineer - AI/ML
Frisco or San Jose or Newport Beach or New York or Toronto or Waterloo
$172k-$283k/yr HybridFull Time
McAfee
McAfee: Provides digital security and privacy software for consumers.
10+ YOE10+ years software development with technical leadership; cloud-native, agentic AI/ML, full-stack and distributed systems experience; Kubernetes, cloud platforms, and strong programming language proficiency.
AWS, Azure, GCP, Python, Java, Go, Rust, C++, JavaScript, TypeScript, Dart, React, Flutter, gRPC, protobuf, Docker, Kubernetes, Helm, ArgoCD, Calico, OPA, Istio, Linkerd, eBPF, Cilium, GitHub Copilot, Claude, ChatGPT, Data Bricks, Snowflake, Big Query, Spark, Hadoop, Kafka
2mo
Save
Mark Applied
Hide
Inference Optimization ML Engineer
Palo Alto, California, United States
OnsiteFull Time
Rhoda AI
Rhoda AI: Developing generalist robotic intelligence for real-world industrial automation.
3+ YOE3+ years in inference optimization, ML systems; strong PyTorch; experience with quantization, pruning, distillation; familiarity with Triton/TensorRT; CUDA knowledge.
PyTorch, JAX, TensorRT, Triton, CUDA, XLA, TorchServe, vLLM
2mo
Save
Mark Applied
Hide
Sr. / Staff ML Engineer, FM Training Integration - ML Compute
Santa Clara, California, United States
HybridFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
5+ YOE5+ years in software/ML infrastructure; Python; PyTorch or JAX; cloud infra and distributed systems; BS in CS/Engineering.
Python, PyTorch, JAX, Docker, Kubernetes, NVIDIA Nsight
1w
Save
Mark Applied
Hide
ML Systems Integration Engineer
Sunnyvale or Toronto
OnsiteFull Time
Cerebras Systems
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
BS/MS in CS/CE/EE or related, strong Python and/or C++, Linux and OS fundamentals, excellent debugging, analytical and communication skills; experience with distributed systems, telemetry, and hardware-adjacent software preferred.
Python, C++, Linux
3w
Save
Mark Applied
Hide
Senior Systems Engineer – Sensors
Santa Clara, California, United States
$142k-$213k/yr OnsiteFull Time
Qualcomm
QualcommNASDAQ: QCOM: Designs and manufactures semiconductors and wireless telecommunications products.
1+ YOEDegree in engineering/CS or related field (BS/MS/PhD) with 1–2+ years systems engineering experience; DSP, sensor fusion, ML/AI knowledge; MATLAB, Python, C/C++; algorithm implementation and validation experience.
MATLAB, Python, C, C++
1mo
Save
Mark Applied
Hide
ML and Agentic Systems Engineer
Santa Clara, California, United States
$224k-$431k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
12+ YOE12+ years building ML systems and platforms; expert Python and PyTorch; experience with ML pipelines, evaluation, tooling, and software engineering; BS/MS in CS/Engineering or equivalent.
Python, PyTorch
1mo
Save
Mark Applied
Hide
ML Software Engineer 6 - AI for Member Systems (AIMS)
Los Gatos, California, United States
$600k-$1066k/yr OnsiteFull Time
Netflix
NetflixNASDAQ: NFLX: Global video streaming and media production service.
Significant experience building and operating production ML systems; deep Python expertise and proficiency in a JVM language (Scala or Java); experience integrating ML capabilities with serving, inference, and feature stores; prototyping and mentoring experience.
Python, JVM, Scala, Java, TensorFlow, PyTorch, JAX
2mo
Save
Mark Applied
Hide
Lead ML Inference Engineer, Advertising
San Jose or Austin
$247k-$486k/yr HybridFull Time
Roku
RokuNASDAQ: ROKU: Operates a TV streaming platform and sells streaming hardware.
10+ YOE5+ MgmtLead the design and development of a state-of-the-art inference platform; 10+ years in distributed systems; ML serving; leadership experience.
High-performance languages, ML frameworks, GPU acceleration, HPC, Distributed systems, Inference platforms, Monitoring tooling
3w
Save
Mark Applied
Hide
Staff ML Software Engineer (L6) — Platform Systems, AIMS Engineering
Los Gatos, California, United States
$600k-$1066k/yr OnsiteFull Time
Netflix
NetflixNASDAQ: NFLX: Provider of global streaming entertainment and video content.
5+ YOEExtensive experience designing, building, and operating large-scale production AI/ML systems; deep Python expertise and JVM (Scala/Java) proficiency; strong distributed systems, observability, cost-optimization, and cross-team collaboration skills.
Python, Scala, Java, LLM