66 distribution modeling engineer jobs at 23 companies in Salinas, CA

2mo
Save
Mark Applied
Hide
Distinguished Engineer - AI
San Jose, California, United States
$266k-$396k/yr OnsiteFull Time
NetApp
NetAppNasdaq: NTAP: Provides intelligent data infrastructure for hybrid cloud environments.
15+ YOEExpert in AI inferencing and distributed systems at scale with 15+ years experience; hands-on with inference engines, model optimization, GPU/TPU orchestration, Kubernetes, RDMA/DPDK; strong architecture, communication, and mentorship skills.
TensorRT, vLLM, ONNX Runtime, Triton, RDMA, DPDK, Kubernetes, GPU, TPU
3d
Save
Mark Applied
Hide
Staff Machine Learning Engineer
Santa Clara, California, United States
$176k-$308k/yr OnsiteFull Time
ServiceNow
ServiceNowNYSE: NOW: Enterprise cloud platform for digital workflow automation.
6+ YOERequires 6+ years building reliable, scalable production software; experience shipping generative AI products, prompt and evaluation engineering, distributed systems, API design, testing, model cost and latency optimization.
Kubernetes, OpenShift, CMDB, Workflow Data Fabric, Knowledge Graph, RAG, Anthropic, Google, OpenAI, Microsoft ServiceNow
2mo
Save
Mark Applied
Hide
Distinguished Engineer - AI (San Jose, CA, US, 95128)
San Jose, California, United States
$266k-$396k/yr OnsiteFull Time
NetApp
NetAppNASDAQ: NTAP: Sells enterprise data storage and cloud management software.
15+ YOE15+ years building low-latency, fault-tolerant distributed systems and AI/ML inference platforms; expertise with inference engines, model optimization, storage for AI, RDMA/DPDK, and Kubernetes-based orchestration.
TensorRT, vLLM, ONNX Runtime, Triton, RDMA, DPDK, Kubernetes
2mo
Save
Mark Applied
Hide
Senior AI / Machine Learning Engineer
San Francisco or New York City or San Jose or Seattle or Austin or Boston
$115k-$200k/yr HybridFull Time
Absentia Labs
Absentia Labs: AI-native toxicology platform accelerating drug safety and discovery.
5+ YOE5+ years ML industry experience, proven production-scale model training, expertise with LLMs/diffusion/GNNs, strong PyTorch skills, distributed training and data-pipeline experience, and solid software engineering practices.
PyTorch, GitHub
2mo
Save
Mark Applied
Hide
Distributed Systems Engineer 4 - Content & Business Products
Los Gatos or United States
$250k-$413k/yr RemoteFull Time
Netflix
NetflixNASDAQ: NFLX: Provider of global streaming entertainment and video content.
2+ YOE2+ years working on distributed systems; proficiency in Java or C# and OO design; experience with multithreading, microservices, data modeling, API design; participate in on-call rotation and lead incident reviews.
Java, C#, gRPC, GraphQL, S3
1mo
Save
Mark Applied
Hide
Tech Lead Software Engineer - AI Compute Infrastructure
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
5+ YOE5+ years experience building cloud/ML infrastructure, strong knowledge of large-model inference, distributed systems, scheduling, and container orchestration; proficiency in Go/Rust/Python/C++.
vLLM, SGLang, TensorRT-LLM, Kubernetes, Ray, Docker, CUDA, AWS, Azure, GCP, SageMaker, Azure ML, Vertex AI, DeepSpeed, PyTorch, Go, Rust, Python, C++
1mo
Save
Mark Applied
Hide
Systems Design/Architecture Engineer 5
Seattle or San Jose
$149k-$293k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
15+ YOE15+ years systems and architecture experience, 10+ years in distributed computing, expertise in IAM, cryptography, network/cloud/application security, threat modeling, and strong communication; bachelor’s or equivalent.
AWS, Azure, GCP
1mo
Save
Mark Applied
Hide
Machine Learning Engineer, TikTok - Business Governance
San Jose, California, United States
$156k-$388k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Strong ML/DL knowledge with Transformer/LLM familiarity, hands-on Python and PyTorch, distributed training and large-scale data processing, experience productionizing models for content safety and cross-team collaboration.
Python, PyTorch
3d
Save
Mark Applied
Hide
Research Engineer
San Jose or New York City
$200k-$300k/yr OnsiteFull Time
Tessera Labs
Tessera Labs: Automates complex enterprise workflows with multi-agent AI systems.
Significant language-model training or post-training experience, RL tuning, Python and PyTorch or JAX proficiency, distributed GPU training, empirical experimentation, and strong software engineering and writing skills.
SAP, Salesforce, Workday, Oracle, Snowflake, MuleSoft, Python, PyTorch, JAX, vLLM, SGLang, Triton, DeepSpeed, Ray, Megatron, TRL
2mo
Save
Mark Applied
Hide
Senior Systems Software Engineer, Accelerated Kubernetes Performance and Scale - DGX Cloud
Santa Clara or Seattle
$184k-$357k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOE8+ years experience in computer architecture, networking, storage or accelerator platforms; expertise in Kubernetes, distributed systems, performance modeling; proficiency in Golang and/or Python; experience with major public cloud providers.
Kubernetes, GPU Operator, Network Operator, node-feature-discovery, topograph, dra-driver-nvidia-gpu, nvsentinel, Grove, gateway-api-inference-extension, DSX, Golang, Python, AWS, Azure, GCP, OCI
3mo
Save
Mark Applied
Hide
Staff Machine Learning Engineer - Agentic Models, LLM, RAG, GenAI
Santa Clara, California, United States
$232k-$310k/yr HybridFull Time
Eightfold
Eightfold: Global AI-native talent intelligence platform provider.
6+ YOELead AI/ML engineer with 6+ years in ML, Gen AI, LLMs; strong Python, TensorFlow/PyTorch; AWS, Docker, Kubernetes; expert in agentic AI and distributed systems.
Python, TensorFlow, PyTorch, AWS, Docker, Kubernetes
2mo
Save
Mark Applied
Hide
ML Engineer - Inference & Model Deployment
Cupertino, California, United States
$250k-$310k/yr OnsiteFull Time
Hiring.Cafe
Hiring.Cafe: An AI-powered job search engine and aggregator.
Experience deploying and optimizing deep learning models in production, multi-GPU inference, profiling/benchmarking model performance, inference optimization techniques, and cloud/distributed systems familiarity.
vLLM, TensorRT, SGLang, GPU
1mo
Save
Mark Applied
Hide
Senior Systems Software Engineer, Accelerated Kubernetes Performance and Scale - DGX Cloud
Santa Clara or Seattle or United States
$152k-$242k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOEBachelor's degree or equivalent experience, 5+ years in computer architecture, networking, storage or accelerator platforms; expertise in Kubernetes, distributed systems, performance modeling, Golang/Python, and major public cloud providers.
Kubernetes, GPU Operator, Network Operator, node-feature-discovery, topograph, dra-driver-nvidia-gpu, nvsentinel, Grove, gateway-api-inference-extension, DSX, Golang, Python, AWS, Azure, GCP, OCI, CI/CD
3mo
Save
Mark Applied
Hide
Lead Machine Learning Engineer - Agentic Models, LLM, RAG, GenAI
Santa Clara, California, United States
$193k-$258k/yr HybridFull Time
Eightfold.ai
Eightfold.ai: AI-native platform for talent management and workforce optimization.
5+ YOESenior ML engineer with expertise in AI agents, LLMs, distributed systems; 5-7+ years of experience; strong Python and ML frameworks; AWS; Docker/Kubernetes; RAG/GenAI experience.
Python, TensorFlow, PyTorch, AWS, Docker, Kubernetes, Kafka, AWS SQS, LangGraph, CrewAI, AutoGen, Pinecone, pgvector, vLLM, TensorRT-LLM
1mo
Save
Mark Applied
Hide
Sr./Staff ML Infrastructure Engineer, Compute (TPU Scheduling) - Foundation Model
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience building schedulers, resource managers, or orchestration systems for distributed workloads; experience with TPU/GPU accelerator infrastructure, distributed ML training/inference, and frameworks such as JAX, PyTorch, TensorFlow, Ray, Pathways; MS/PhD preferred.
TPU, GPU, JAX, PyTorch, TensorFlow, Ray, Pathways
3d
Save
Mark Applied
Hide
Senior Engineer, Platform & Data
Santa Clara, California, United States
$150k-$200k/yr HybridFull Time
LeanData
LeanData: Automates lead-to-account matching and routing for B2B revenue teams.
5+ YOERequires 5+ years building production backend and distributed systems, strong Python, distributed-systems expertise, cloud-native development, relational data modeling and SQL, and independent service ownership.
Python, TypeScript, AWS, SQL, Postgres, Salesforce, Bulk 2.0, REST, Pub/Sub, CDC, Postgres + RLS, Temporal, Inngest, Langfuse, Braintrust
1w
Save
Mark Applied
Hide
Machine Learning Engineer - Ads Core and Commerce Ads
San Jose or Los Angeles County
$137k-$360k/yr OnsiteFull Time
TikTok USDS Joint Venture
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
5+ YOERequires SQL and Python, data manipulation, Hadoop or Spark, distributed computing, machine learning, deep learning, feature engineering, model evaluation, optimization, and 5+ years of relevant experience preferred.
SQL, Python, Hadoop, Spark
3mo
Save
Mark Applied
Hide
Distributed Systems Engineer 4 - Content & Business Products
Los Gatos, California, United States
$250k-$413k/yr OnsiteFull Time
Netflix
NetflixNASDAQ: NFLX: Global video streaming and media production service.
2+ YOE2+ years in distributed systems; proficient in Java or C#; strong knowledge of multithreading, observability; experience with microservices, data modeling, API design; on-call experience; good cross-functional communication.
Java, C#, OO design, Microservices, API design, gRPC, GraphQL, observability
2mo
Save
Mark Applied
Hide
Senior Principal Software Engineer
Seattle or Santa Clara or United States
$135k-$306k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
6+ YOEMS in Computer Science with formal verification focus; 5+ years building concurrent/distributed systems; proficiency in C/C++, Java, GoLang, or Rust; experience with TLA+, model checkers (TLC, Apalache) and proof tools (TLAPS).
C, C++, Java, GoLang, Rust, TLA+, TLC, Apalache, TLAPS, Slack, QEMU, SQL, NoSQL
2mo
Save
Mark Applied
Hide
Senior Machine Learning Engineer
Austin or San Jose or United States or Canada or Mexico
HybridFull Time
Roku
RokuNASDAQ: ROKU: Operates a TV streaming platform and sells streaming hardware.
5+ YOE5+ years applied ML experience; strong software development skills in Spark, Python, or Java; experience with distributed ML frameworks, low-latency model evaluation, experimentation, statistics, and large-scale data systems.
Spark, Spark-MLlib, Python, Java, TensorFlow, Hive, Aerospike, ScyllaDB