59 ml infrastructure engineer jobs at 22 companies in Manteca, CA
1w
Save
Mark Applied
Hide
1w
ML Systems Research Engineer, RL / Inference / Agent Systems
Santa Clara, California, United States
HybridFull Time
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Experienced ML systems engineer with strong Python and ML framework skills, experience in RL/inference systems, distributed experimentation, and GPU/infrastructure workflows; advanced degree preferred.
Python, PyTorch, JAX, TensorFlow, Kubernetes, Ray, Slurm, ROCm, HIP, CUDA
PayPalNASDAQ: PYPL: Digital platform for sending money and processing online payments.
10+ YOE10+ years relevant experience and a Bachelor’s degree (or equivalent); deep expertise in databases, data pipelines, messaging, caching, performance and resilience engineering; experience with real-time analytics and AI/ML infrastructure; strong technical leadership and mentoring.
8+ YOEBachelor's degree or equivalent experience; 8 years C++ programming; 5 years testing/launching software; 5 years building large-scale infrastructure or storage/compute; 3 years software design/architecture.
TikTok: Global short-form video hosting and social media platform.
3+ YOE3+ years building scalable ML systems, strong CS fundamentals, coding skills, experience with causal inference/uplift/deep learning, project management and communication skills.
Altera: Manufacturer of field-programmable gate arrays and programmable logic devices.
10+ YOEBachelor's degree, 10+ years ML engineering/MLOps experience, strong Python, cloud ML platforms (AWS/GCP/Azure), Docker/Kubernetes, CI/CD, ML frameworks (PyTorch/TensorFlow/JAX), experience with MLflow/W&B and HPC schedulers.
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
5+ YOEBachelor's/Master's in CS or related,5+ years software engineering,3+ years ML infrastructure,expertise in Python/C++/Java,PyTorch/Tensorflow,Kubernetes,Slurm,CUDA and GPU cluster management.
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
Build and scale data infrastructure for data engineering/ML; Kubernetes, Trino, Ray, MLflow; on-call; Go or Python; BS in CS/Engineering or equivalent.
Tech Lead Software Engineer - AI Compute Infrastructure
San Jose, California, United States
OnsiteFull Time
ByteDance: Developing AI-driven content platforms and mobile applications.
5+ YOE5+ years experience building cloud/ML infrastructure, strong knowledge of large-model inference, distributed systems, scheduling, and container orchestration; proficiency in Go/Rust/Python/C++.
Lucid MotorsNASDAQ: LCID: Designs and manufactures high-performance luxury electric vehicles.
Design and operate cloud and hybrid infrastructure for ADAS/autonomous systems; lead an in-house DevOps/infrastructure org; build CI/CD and ML pipelines; implement IaC (Terraform/GitOps); manage AWS/OCI/Azure; ensure security and regulatory compliance (e.g., GDPR).
Nex: Gaming system that turns body movement into interactive play.
3+ YOE3+ years building production ML systems or training infrastructure; proficiency in Python and one systems language (C++, C#, Java, Rust, or Go); experience with PyTorch or TensorFlow; building training pipelines, data workflows, and model deployment.
Circuit Check: Designs and manufactures automated electronic test systems and fixtures.
10+ YOE10+ years in strategic account management or business development for AI/ML infrastructure; engineering degree; executive-level engagement experience.
BlackLineNASDAQ: BL: Provides cloud-based financial close and accounting automation software.
Extensive experience building and operating ML/AI production systems, designing scalable data pipelines, distributed training, model deployment, CI/CD, observability, and cloud infrastructure.
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
3+ YOE3+ years in software engineering, backend infrastructure, data systems, ML infrastructure; distributed systems; APIs/data pipelines; Python/Java/C/C++/Go; cloud; ML systems familiarity.
Python, Java, C++, Go, Kafka, Spark, Flink, distributed data frameworks
KLANASDAQ: KLAC: Provides process control and yield management for semiconductor manufacturing.
5+ YOEDegree in CS/CE or related, 5+ years systems/DevOps/ML infrastructure experience, hands-on AI/GPU cluster and Linux/Kubernetes expertise, Python/Bash scripting, and knowledge of storage, networking, and security.
Sr. Staff AI Engineer - On-Prem AI Infrastructure & Agentic Systems
San Jose, California, United States
$140k-$165k/yrOnsiteFull Time
SK hynix Memory Solutions AmericaKorea Exchange: 000660: Develops semiconductor controllers and firmware for enterprise data storage.
2+ YOE2+ years in AI/ML engineering with on-prem/private-cloud deployment, experience building agentic AI and RAG pipelines, model fine-tuning (LoRA/QLoRA), Python/Linux/Docker/Kubernetes proficiency, and familiarity with vector DBs and AI serving frameworks.
PayPalNASDAQ: PYPL: Global digital payments platform for consumers and merchants.
5+ YOEBachelor's degree in computer science, computer engineering, electrical engineering, or related field, plus 5 years' experience. Requires Java, Python, Golang, GCP, Kubernetes, MLOps, and production ML infrastructure expertise.
Java, Python, Golang, GCP, Kubernetes, Prometheus, Datadog, Jupyter Notebook, Spark, MLFlow, ONNX, MySQL, MLOps, ML Model Serving, CRDs/controllers
Senior Director of Engineering - Next-generation Network OS for Switching and routing Infrastructure-10393
San Jose, California, United States
$250k-$310k/yrHybridFull Time
Extreme NetworksNASDAQ: EXTR: Provides cloud-managed networking hardware and software solutions.
20+ YOE7+ Mgmt20+ years engineering experience with leadership of large global teams; deep networking infrastructure expertise; hands-on adoption of AI/ML tools; cloud-native, microservices, DevSecOps and CI/CD experience.
San Francisco or Atlanta or New York City or Denver or San Jose or Los Angeles or Seattle or Chicago or Austin or Dallas or Arlington
$154k-$239k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
8+ YOE8+ years technical experience, 3+ years designing and consulting on applications/infrastructure, experience in sales-facing technical roles, deep AI/ML knowledge, strong math/stat and communication skills.