217 ml platform engineer jobs at 85 companies in Watsonville, CA
1mo
Save
Mark Applied
Hide
1mo
AI/ML Platform Engineer
Santa Clara, California, United States
$179k-$306k/yrOnsiteFull Time
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Experienced systems engineer to build and operate distributed ML platform infrastructure, GPU cluster scheduling, experiment tracking, and production developer tooling. Bachelor's required or equivalent; Master's/PhD preferred.
Principal Engineer - AI /ML Platform (Remote Eligible)
Brooklyn Park or Sunnyvale
$168k-$356k/yrHybridFull Time
TargetNYSE: TGT: General merchandise retailer operating physical stores and e-commerce.
Extensive experience designing and delivering large-scale cloud-native ML platforms, expertise with MLOps, Kubernetes, observability, and platform automation; MS in related field preferred.
Vertex AI, Kubeflow, MLflow, Kubernetes, Terraform, GitOps, service mesh
Woven by Toyota: Developing software-defined mobility platforms and autonomous driving systems
5+ YOEBSc in CS/ML or equivalent experience,5+ years software engineering,2+ years Linux/Python/PyTorch or TensorFlow,experience across MLOps (data curation, distributed training, deployment),Docker and CI experience,strong engineering skills.
TargetNYSE: TGT: Operates a chain of general merchandise stores and supermarkets.
MS or equivalent preferred; extensive experience designing and operating large-scale cloud-native ML platforms, Kubernetes-based infrastructure, MLOps, model governance, observability, and platform automation.
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOE8+ years building and operating production platform/backend infrastructure; 5+ years ML infrastructure; strong Python and a compiled language; experience with job queues, sandboxed execution, and production reliability.
RokuNASDAQ: ROKU: Operates a TV streaming platform and sells streaming hardware.
10+ YOE5+ MgmtLead the design and development of a state-of-the-art inference platform; 10+ years in distributed systems; ML serving; leadership experience.
New York City or Milwaukee or Dallas or Columbus or Kirkland or Cincinnati or Cleveland or Oklahoma City or Austin or Albany or Chicago or St. Petersburg or Hartford or Pittsburgh or St. Louis or Miami or Sacramento or Raleigh or Minneapolis or Mountain View or Scottsdale or San Francisco or Morristown or Denver or Boston or Philadelphia or Des Moines or Overland Park or Los Angeles or Charlotte or Walnut Creek or Carmel or Seattle or Houston or Arlington or Atlanta or Redmond or Bentonville or Beaverton or Nashville or Detroit or San Diego
$80k-$294k/yrHybridFull Time
AccentureNYSE: ACN: Global provider of management consulting and technology services.
6+ YOEBachelor's degree or equivalent experience, 6+ years developing and deploying AI/ML solutions, programming proficiency, experience with ML frameworks, big data, databases, and cloud platforms.
Python, R, Java, TensorFlow, PyTorch, Hadoop, Spark, AWS, Microsoft Azure, Google Cloud, D3.js, ggplot
Waymo: Autonomous driving technology for ride-hailing and logistics.
4+ YOEBS in CS or equivalent, 4+ years backend development in Java/C++, Python ML experience with TensorFlow/PyTorch/Keras; building backend platforms and ML workflows.
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
8+ YOE8+ years designing and operating production platform/backend infrastructure,5+ years ML infrastructure,BS/MS/PhD or equivalent in CS/EE/CE,strong Python and compiled-language skills,experience with Kubernetes Jobs and job queues.
United States or California or San Jose or Seattle or Portland or Boston or Chicago
$250k-$289k/yrRemoteFull Time
WEXNYSE: WEX: Provides global payment processing and business information management services.
12+ YOE12+ years in software or ML engineering, including 5+ years in applied AI/ML research. Requires advanced expertise in model architecture, algorithms, Python, deep learning frameworks, distributed AI, cloud, and ML platforms.
KLANASDAQ: KLAC: Provides process control and yield management for semiconductor manufacturing.
5+ YOEDegree in CS/CE or related, 5+ years systems/DevOps/ML infrastructure experience, hands-on AI/GPU cluster and Linux/Kubernetes expertise, Python/Bash scripting, and knowledge of storage, networking, and security.
Exact SciencesNASDAQ: EXAS: Provides molecular diagnostic tests for early cancer detection.
3+ YOE3+ MgmtDesign and deploy AI/ML and NLP solutions (including LLM/Gen AI) with leadership experience, strong Python skills, AI/ML knowledge, ethical AI awareness, and 3+ years project leadership experience.
Wayve: Develops end-to-end artificial intelligence for autonomous driving systems.
10+ YOE10+ years building large-scale distributed systems or ML infrastructure, 3+ years at staff/principal level, experience with Spark, Ray, Kubernetes, Airflow, MLflow, web frameworks, reliability engineering, and mentoring engineers.
Hewlett Packard EnterpriseNYSE: HPE: Providing global edge-to-cloud infrastructure and IT solutions for businesses.
4+ YOESenior IC with 4-7 years experience; design, build, operate production-grade agentic platform; hybrid work model; Bachelor’s degree in CS/Engineering; Master’s preferred.
Quest Global: Global engineering services for product development and lifecycle management.
3+ YOEExperience designing and delivering NLP/LLM systems using Python or Typescript, working with LLM SDKs/APIs (Claude, Gemini), cloud platforms (AWS/Cloudbase), vector databases, and containerization; 3+ years experience.
Sunnyvale or Washington or Austin or Mountain View or Warren
$171k-$261k/yrHybridFull Time
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
3+ YOE3+ years building and operating production cloud systems; experience with scalable services, Go or Python, Unix/Linux, networking, and mentoring; BS in CS or equivalent.
Nuro: Builds autonomous driving software and electric delivery robots.
3+ YOE3+ years in ML infrastructure/backend platform or distributed systems. Experience with Terraform/Pulumi/Crossplane, Kubernetes/Ray/Slurm/Volcano schedulers, Apache Spark/Beam, feature stores (Feast/Hopsworks/Redis), and systems design for HPC.
ML Infra Engineer Intern (Ads Infra) - 2027 Summer
San Jose, California, United States
OnsiteInternship
TikTok: Global short-form video hosting and social media platform.
Currently pursuing bachelor's or master's in CS/AI/ML or related; proficiency in C++/Python/Go/Java; knowledge of data structures, algorithms, ML, and distributed systems; familiarity with recommendation systems and ML platforms.