896 ml systems engineer jobs at 348 companies in Larkspur, CA
2mo
Save
Mark Applied
Hide
2mo
ML Systems & Performance Engineer
San Francisco, California, United States
OnsiteFull Time
Engram: Developing persistent memory layers for enterprise AI systems.
5+ YOE5+ years building training/inference systems; strong engineering skills; experience with ML frameworks, GPUs, distributed systems; bachelor's degree or equivalent experience.
ML Systems Research Engineer, RL / Inference / Agent Systems
Santa Clara, California, United States
HybridFull Time
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Experienced ML systems engineer with strong Python and ML framework skills, experience in RL/inference systems, distributed experimentation, and GPU/infrastructure workflows; advanced degree preferred.
Python, PyTorch, JAX, TensorFlow, Kubernetes, Ray, Slurm, ROCm, HIP, CUDA
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
8+ YOERequires 8+ years building distributed platforms, AI-tool experience, and scalable full-stack coding with Python, TypeScript, Go, React, SQL, GraphQL, and related technologies.
Docker: Provides a platform for building, sharing, and running containerized applications.
8+ YOE8+ years professional software engineering experience, 5+ years applied ML experience, bachelor's in CS/Engineering or equivalent, experience shipping ML systems, LLM/agent experience, on-call participation possible.
Physical Intelligence: Creating foundation models for general-purpose robot intelligence.
Strong software engineering fundamentals with experience building distributed systems, large-scale data pipelines, object storage and batch/streaming systems; ownership mindset and performance focus.
Vancouver or San Francisco or San Jose or Cork or London
$160k-$180k/yrHybridFull Time
Tigera: Security and observability platform for Kubernetes clusters.
5+ YOERequires 5+ years of ML engineering experience, including 2+ years deploying production ML systems, strong classical ML and anomaly detection skills, LLM application experience, Python, ML frameworks, telemetry data, and communication skills.
Sciforium: Building multimodal AI models and high-performance model serving infrastructure.
5+ YOE5+ years of ML/AI software engineering experience, production ML systems expertise, and a BS, MS, or PhD in a technical field or equivalent experience. Requires generative AI and distributed ML framework knowledge.
RivianNASDAQ: RIVN: Designs and manufactures electric vehicles and charging networks.
5+ YOE5+ years building and scaling ML solutions for auto-labeling and AV perception; strong Python, perception pipeline and system engineering experience; BS/MS/PhD in CS/Robotics/Electrical Engineering or related.
New York City or San Francisco or London or Sydney
$240k-$270k/yrOnsiteFull Time
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
10+ YOEBachelor's in CS/Engineering/Mathematics, 10+ years building and deploying production AI/ML systems, expertise in ML/DL, full ML lifecycle experience, foundation-model adaptation experience.
MakerMaker: Small San Francis-based team building autonomous ML systems
6+ YOESenior ML engineer with 6+ years building production-grade ML systems; strong Python; distributed systems experience; familiar with Ray, Kubernetes, and experimentation infrastructure.
xAI: Develops advanced artificial intelligence systems to understand the universe.
2+ YOE2+ years building large-scale production systems or ML infrastructure; degree in CS or related field or equivalent experience; strong Python and compiled-language skills; experience with GPU and distributed systems.
Databricks: A unified platform for data analytics and artificial intelligence.
4+ YOEMaster's in ML/data science or related field, strong production ML experience, cloud/distributed systems familiarity, proficiency in Python/Scala/Java; PhD preferred; 4+ years ML engineering experience preferred.
Python, Scala, Java, Apache Spark, Delta Lake, MLflow
Clera: AI talent agent matching professionals with high-growth startup roles
5+ YOERequires 5+ years building production ML inference or model-serving systems, experience scaling for latency and reliability, Docker, Kubernetes, distributed systems, observability tools, cloud platforms, and Python, Go, Rust, C++, or Java.
ML Systems Engineer, Large-Scale Model Training & RL Infrastructure
Palo Alto, California, United States
$195k-$262k/yrOnsiteFull Time
NebiusNasdaq: NBIS: Builds cloud infrastructure and software for artificial intelligence development.
Strong Python and PyTorch skills, hands-on distributed model training and GPU cluster experience, debugging across NCCL/CUDA/PyTorch/Ray, and quantitative reasoning about throughput, utilization, memory, and cost.
Sodalis: AI operating system for specialty pharmacy workflows and automation.
5+ YOE5+ years building production ML systems, experience in information extraction/NLP/LLMs, eval and data-pipeline ownership, startup 0→1 experience, pragmatic modeling and clinical domain curiosity.
Experienced in biomolecular modeling and ML for drug discovery; ability to fine-tune protein/structure models, build ML systems that integrate with wet-lab workflows, and evaluate model impact on drug development.
Rhoda AI: Developing generalist robotic intelligence for real-world industrial automation.
3+ YOE3+ years in inference optimization, ML systems; strong PyTorch; experience with quantization, pruning, distillation; familiarity with Triton/TensorRT; CUDA knowledge.