48 distribution modeling engineer jobs at 16 companies in Pacific Grove, CA

13h
Save
Mark Applied
Hide
Senior Virtual Platform Functional Modeling Engineer
San Jose, California, United States
$175k-$300k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Preferred bachelor's degree in a technical field and master's degree. Requires C++ development, functional modeling, virtual platforms, validation, distributed systems, computer architecture, and interconnect technology experience.
C++, CI/CD, PCIe, CXL, AXI, CHI, ACE, UCIe, HBM, DDR, SystemC/TLM-2.0, Arm Fast Models, Simics, QEMU
2mo
Save
Mark Applied
Hide
Distinguished Engineer - AI
San Jose, California, United States
$266k-$396k/yr OnsiteFull Time
NetApp
NetAppNasdaq: NTAP: Provides intelligent data infrastructure for hybrid cloud environments.
15+ YOEExpert in AI inferencing and distributed systems at scale with 15+ years experience; hands-on with inference engines, model optimization, GPU/TPU orchestration, Kubernetes, RDMA/DPDK; strong architecture, communication, and mentorship skills.
TensorRT, vLLM, ONNX Runtime, Triton, RDMA, DPDK, Kubernetes, GPU, TPU
2mo
Save
Mark Applied
Hide
Distinguished Engineer - AI (San Jose, CA, US, 95128)
San Jose, California, United States
$266k-$396k/yr OnsiteFull Time
NetApp
NetAppNASDAQ: NTAP: Sells enterprise data storage and cloud management software.
15+ YOE15+ years building low-latency, fault-tolerant distributed systems and AI/ML inference platforms; expertise with inference engines, model optimization, storage for AI, RDMA/DPDK, and Kubernetes-based orchestration.
TensorRT, vLLM, ONNX Runtime, Triton, RDMA, DPDK, Kubernetes
2mo
Save
Mark Applied
Hide
Senior AI / Machine Learning Engineer
San Francisco or New York City or San Jose or Seattle or Austin or Boston
$115k-$200k/yr HybridFull Time
Absentia Labs
Absentia Labs: AI-native toxicology platform accelerating drug safety and discovery.
5+ YOE5+ years ML industry experience, proven production-scale model training, expertise with LLMs/diffusion/GNNs, strong PyTorch skills, distributed training and data-pipeline experience, and solid software engineering practices.
PyTorch, GitHub
2mo
Save
Mark Applied
Hide
Distributed Systems Engineer 4 - Content & Business Products
Los Gatos or United States
$250k-$413k/yr RemoteFull Time
Netflix
NetflixNASDAQ: NFLX: Provider of global streaming entertainment and video content.
2+ YOE2+ years working on distributed systems; proficiency in Java or C# and OO design; experience with multithreading, microservices, data modeling, API design; participate in on-call rotation and lead incident reviews.
Java, C#, gRPC, GraphQL, S3
1mo
Save
Mark Applied
Hide
Tech Lead Software Engineer - AI Compute Infrastructure
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
5+ YOE5+ years experience building cloud/ML infrastructure, strong knowledge of large-model inference, distributed systems, scheduling, and container orchestration; proficiency in Go/Rust/Python/C++.
vLLM, SGLang, TensorRT-LLM, Kubernetes, Ray, Docker, CUDA, AWS, Azure, GCP, SageMaker, Azure ML, Vertex AI, DeepSpeed, PyTorch, Go, Rust, Python, C++
1mo
Save
Mark Applied
Hide
Systems Design/Architecture Engineer 5
Seattle or San Jose
$149k-$293k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
15+ YOE15+ years systems and architecture experience, 10+ years in distributed computing, expertise in IAM, cryptography, network/cloud/application security, threat modeling, and strong communication; bachelor’s or equivalent.
AWS, Azure, GCP
1mo
Save
Mark Applied
Hide
Machine Learning Engineer, TikTok - Business Governance
San Jose, California, United States
$156k-$388k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Strong ML/DL knowledge with Transformer/LLM familiarity, hands-on Python and PyTorch, distributed training and large-scale data processing, experience productionizing models for content safety and cross-team collaboration.
Python, PyTorch
4d
Save
Mark Applied
Hide
Research Engineer
San Jose or New York City
$200k-$300k/yr OnsiteFull Time
Tessera Labs
Tessera Labs: Automates complex enterprise workflows with multi-agent AI systems.
Significant language-model training or post-training experience, RL tuning, Python and PyTorch or JAX proficiency, distributed GPU training, empirical experimentation, and strong software engineering and writing skills.
SAP, Salesforce, Workday, Oracle, Snowflake, MuleSoft, Python, PyTorch, JAX, vLLM, SGLang, Triton, DeepSpeed, Ray, Megatron, TRL
2mo
Save
Mark Applied
Hide
ML Engineer - Inference & Model Deployment
Cupertino, California, United States
$250k-$310k/yr OnsiteFull Time
Hiring.Cafe
Hiring.Cafe: An AI-powered job search engine and aggregator.
Experience deploying and optimizing deep learning models in production, multi-GPU inference, profiling/benchmarking model performance, inference optimization techniques, and cloud/distributed systems familiarity.
vLLM, TensorRT, SGLang, GPU
1mo
Save
Mark Applied
Hide
Sr./Staff ML Infrastructure Engineer, Compute (TPU Scheduling) - Foundation Model
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience building schedulers, resource managers, or orchestration systems for distributed workloads; experience with TPU/GPU accelerator infrastructure, distributed ML training/inference, and frameworks such as JAX, PyTorch, TensorFlow, Ray, Pathways; MS/PhD preferred.
TPU, GPU, JAX, PyTorch, TensorFlow, Ray, Pathways
1w
Save
Mark Applied
Hide
Machine Learning Engineer - Ads Core and Commerce Ads
San Jose or Los Angeles County
$137k-$360k/yr OnsiteFull Time
TikTok USDS Joint Venture
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
5+ YOERequires SQL and Python, data manipulation, Hadoop or Spark, distributed computing, machine learning, deep learning, feature engineering, model evaluation, optimization, and 5+ years of relevant experience preferred.
SQL, Python, Hadoop, Spark
3mo
Save
Mark Applied
Hide
Distributed Systems Engineer 4 - Content & Business Products
Los Gatos, California, United States
$250k-$413k/yr OnsiteFull Time
Netflix
NetflixNASDAQ: NFLX: Global video streaming and media production service.
2+ YOE2+ years in distributed systems; proficient in Java or C#; strong knowledge of multithreading, observability; experience with microservices, data modeling, API design; on-call experience; good cross-functional communication.
Java, C#, OO design, Microservices, API design, gRPC, GraphQL, observability
2mo
Save
Mark Applied
Hide
Senior Machine Learning Engineer
Austin or San Jose or United States or Canada or Mexico
HybridFull Time
Roku
RokuNASDAQ: ROKU: Operates a TV streaming platform and sells streaming hardware.
5+ YOE5+ years applied ML experience; strong software development skills in Spark, Python, or Java; experience with distributed ML frameworks, low-latency model evaluation, experimentation, statistics, and large-scale data systems.
Spark, Spark-MLlib, Python, Java, TensorFlow, Hive, Aerospike, ScyllaDB
1d
Save
Mark Applied
Hide
Senior Fullstack Platform Go Engineer
Morgan Hill, California, United States
$137k-$198k/yr OnsiteFull Time
Specialized
Specialized: Designs and manufactures high-performance bicycles and cycling equipment.
8+ YOERequires Go, Java, or C++; distributed systems, REST API, AWS, Docker, data modeling, and RDBMS experience. Preferred: 8+ years relevant experience, 2+ years Go/Java/C++, PostgreSQL/MySQL, Terraform, GitHub Actions, and HTMX.
Go, Golang, Java, C++, AWS, API Gateway, Lambda, EventBridge, SQS, SNS, ECS, Docker, Postgres, MySQL, SQL, GitHub Actions, Terraform, HTML, CSS, JavaScript, HTMX, Git, GitHub, Atlassian Confluence, JIRA, JIRA Service Desk, Slack, Dropbox
1w
Save
Mark Applied
Hide
Algorithm Expert - Financial Foundation LLM
San Jose, California, United States
RemoteFull Time
DiDi Global
DiDi GlobalOTC Markets: DIDIY: Global technology platform providing mobility, delivery, and financial services.
3+ YOEMaster's degree in a relevant field, 3+ years of deep learning R&D experience, pretrained model development, Transformer expertise, distributed training, ablation studies, and strong model engineering skills.
GPT, BERT, FT-Transformer, MoE, PyTorch, Contrastive Learning, ELECTRA, TimeMixer, TrajGPT
1w
Save
Mark Applied
Hide
Inference Systems Performance Architect
San Jose, California, United States
$245k-$325k/yr OnsiteFull Time
SambaNova Systems
SambaNova Systems: Develops custom AI hardware and software for enterprise computing.
12+ YOERequires 12+ years in performance engineering, distributed-systems analysis, workload generation, simulation, modeling, technical leadership, cross-functional influence, mentoring, and delivering complex ambiguous projects.
LLM, GPU, RDU, Headspace, Gympass+, One Medical, Employee Assistance Program (EAP)
2mo
Save
Mark Applied
Hide
Software Development Manager, AWS Neuron SDK - Distributed Training
Cupertino, California, United States
$213k-$288k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
7+ YOE3+ MgmtExperience with PyTorch or JAX, distributed training at scale, 7+ years engineering experience, 3+ years engineering team management, 3+ years designing/architecting systems, and partnering with product teams; deep learning model training experience.
PyTorch, JAX, FSDP, DeepSpeed, Megatron, CUDA, Neuron SDK, Trainium