214 distribution modeling engineer jobs at 119 companies in Santa Clara, CA
3mo
Save
Mark Applied
Hide
3mo
Principal Python Engineer (Modeling & Simulation)
San Francisco or Boston or United States
HybridFull Time
Code Metal: Verifiable AI-powered code translation for mission-critical industries.
7+ YOE7+ years Python software development, experience with scalable APIs, Docker, cloud distributed systems, stakeholder-driven design, SDLC best practices; eligible for U.S. Top Secret clearance.
People.ai: AI platform providing revenue intelligence and sales productivity solutions.
5+ YOE5+ years building backend systems; 2+ years in Python or Scala; experience with distributed systems, large-scale data processing (Spark/Hive/Hadoop/MapReduce), stream processing (Kafka/Apache Samza/Apache Storm); strong analytical and ownership skills.
CoupangNYSE: CPNG: Provides online retail, grocery delivery, and video streaming services.
8+ YOE3+ Mgmt8+ years data engineering experience, 3+ years people management, proficiency in data modeling, ETL, Python, SQL, distributed processing and storage, and ETL schedulers.
NetAppNasdaq: NTAP: Provides intelligent data infrastructure for hybrid cloud environments.
15+ YOEExpert in AI inferencing and distributed systems at scale with 15+ years experience; hands-on with inference engines, model optimization, GPU/TPU orchestration, Kubernetes, RDMA/DPDK; strong architecture, communication, and mentorship skills.
Sequen AI: AI-native ranking engine for enterprise search and recommendations.
4+ YOEMinimum 4+ years MLOps or ML/platform engineering; expertise with low-latency model serving, Python and PyTorch; cloud (AWS/GCP/Azure), Docker, Kubernetes, MLflow; strong distributed systems and pipeline experience.
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
Senior production engineer with distributed systems experience, hands-on with large language models, SRE mindset, and strong programming skills (Python, Go, Java, or C++), Kubernetes knowledge, and collaborative skills.
8+ YOE3+ MgmtBachelor's degree or equivalent, 8 years programming experience, 5 years launching software and building large-scale infrastructure, 3 years software architecture, and distributed computing experience.
Harvey: AI platform for legal research and document analysis.
Hands-on post-training/model-training experience (SFT, RLHF, reward modeling, distillation) with open-weight models, strong Python, ability to run and debug GPU/distributed experiments, and self-manage applied research projects.
Otter.ai: AI-powered meeting transcription and automated note-taking platform.
5+ YOE5+ years building LLM/AI-agent products, strong backend or distributed-systems engineering, experience with evaluation and improving AI quality, debugging models/pipelines, and product-focused execution.
ServiceNowNYSE: NOW: Enterprise cloud platform for digital workflow automation.
6+ YOERequires 6+ years building reliable, scalable production software; experience shipping generative AI products, prompt and evaluation engineering, distributed systems, API design, testing, model cost and latency optimization.
Kubernetes, OpenShift, CMDB, Workflow Data Fabric, Knowledge Graph, RAG, Anthropic, Google, OpenAI, Microsoft ServiceNow
Rhoda AI: Developing generalist robotic intelligence for real-world industrial automation.
B.S. in CS, Systems Engineering, or Robotics; experience designing AI safety for electromechanical products; STPA and V&V experience; understanding of distributional shift and model uncertainty; ability to define AI safety requirements and lead verification.
STPA, FMEA, ISO/IEC TS 22440, ISO/PAS 8800, IEC 61508, ISO 26262, time of flight cameras, laser sensors
Plenful: AI-powered workflow automation platform for healthcare and pharmacy operations.
12+ YOE12+ years building data architecture for large-scale production systems with relational DBs, distributed systems, Python, observability, schema evolution, and strong judgment on data modeling.
Rippling: Unified platform managing workforce HR, IT, and finance operations
8+ YOE8+ years software engineering experience, distributed systems ownership, experience training/deploying LLMs, model inference optimization, backend skills in Python/Go/Java, and cloud-native infrastructure (Kubernetes).
San Francisco or New York City or Los Angeles or Seattle
$180k-$260k/yrHybridFull Time
Whatnot: Social marketplace for buying and selling via live streams
3+ YOE3+ years building data warehouses or distributed/event-driven systems; expertise in data modeling, pipelines, observability, cloud warehouses, Python/SQL; experience with payments/finance stakeholders.
Kafka, Debezium, dbt, Spark, Flink, Dagster, Airflow, Monte Carlo, Great Expectations, Snowflake, BigQuery, Redshift, Python, SQL, CI/CD
Cloud Engineer III, Falcon Exposure Management (Hybrid)
Sunnyvale or New York City or Austin or Redmond
$120k-$180k/yrHybridFull Time
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
5+ YOE5+ years building production systems at scale; proficiency in Go, Python, or Java; distributed systems expertise; production LLM and agentic systems experience; frontier-model SDK experience; strong engineering judgment.
Shipt: Provides same-day delivery services from local retailers via app.
5+ YOE5+ years of machine learning and backend software engineering; backend in Go/Java and Python; embeddings, similarity search, ranking models; ML pipelines; distributed systems; SQL/NoSQL; API serving; A/B testing.
Distinguished Engineer - AI (San Jose, CA, US, 95128)
San Jose, California, United States
$266k-$396k/yrOnsiteFull Time
NetAppNASDAQ: NTAP: Sells enterprise data storage and cloud management software.
15+ YOE15+ years building low-latency, fault-tolerant distributed systems and AI/ML inference platforms; expertise with inference engines, model optimization, storage for AI, RDMA/DPDK, and Kubernetes-based orchestration.
Machine Learning Engineer, Model Evaluations (Speech LLM) - San Francisco
San Francisco, California, United States
$180k-$270k/yrHybridFull Time
Plaud: Develops AI-powered voice recorders and automated transcription software.
Python software engineering; building distributed systems, data pipelines, and evaluation harnesses at scale; partner with ML researchers to define benchmarks; build dashboards and monitor model health; debug mid-training anomalies; communicate results clearly.
San Francisco or New York City or San Jose or Seattle or Austin or Boston
$115k-$200k/yrHybridFull Time
Absentia Labs: AI-native toxicology platform accelerating drug safety and discovery.
5+ YOE5+ years ML industry experience, proven production-scale model training, expertise with LLMs/diffusion/GNNs, strong PyTorch skills, distributed training and data-pipeline experience, and solid software engineering practices.