606 optimizer jobs at 222 companies in Tracy, CA

2mo
Save
Mark Applied
Hide
Software Engineer - Ride and Fleet Services
San Diego or Seattle or Foster City
$145k-$219k/yr HybridFull Time
Zoox
ZooxNASDAQ: AMZN: Developing autonomous robotaxis for urban ride-hailing services.
Design and build scalable fleet-dispatch systems; develop optimization algorithms; apply ML/OR techniques; collaborate across teams.
Java, Kotlin, Machine Learning, Operations Research, Optimization
2mo
Save
Mark Applied
Hide
Software Engineer, Marketplace Pricing
Mountain View or San Francisco
$175k-$215k/yr OnsiteFull Time
Waymo
Waymo: Autonomous driving technology for ride-hailing and logistics.
4+ YOEBS in CS or equivalent; 4+ years backend experience; experience building distributed backend systems; preferred C++, MS CS; experience with low-latency, large-scale distributed systems; ML/optimization infrastructure and production models.
C++, Python, Java, Distributed Systems, Machine Learning, Optimization
1w
Save
Mark Applied
Hide
AI Systems, Model Optimization
Palo Alto or United States
HybridFull Time
Unconventional
Unconventional: Developing novel computing hardware for efficient AI acceleration.
MS/PhD (or equivalent) in quantitative field, deep practical experience with ML stack and GPU performance optimization, proficiency in profiling and optimizing large ML codebases.
PyTorch, torch.compile, DDP, FSDP, CUDA, Triton, CUTLASS, Megatron-LM, DeepSpeed
1w
Save
Mark Applied
Hide
Director of Optimization Engineering
Oakland, California, United States
HybridFull Time
Adapture Renewables
Adapture Renewables: Develops and operates utility-scale solar and energy storage systems.
8+ YOE3+ MgmtBachelor's degree in engineering, 8+ years in utility-scale solar/BESS design and optimization, 3+ years leading technical teams, PV performance modeling (PVSyst), battery dispatch/LP optimization, and strong data-to-decision skills.
PVSyst, Python, SciPy, MATLAB, SQL, SCADA, EMS
3w
Save
Mark Applied
Hide
Senior Data Scientist, Cloud Gaming - Prescriptive Analytics and Optimization
Santa Clara or United States
$184k-$288k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOE6+ years experience (or PhD) in quantitative field, strong prescriptive analytics and optimization background, Python and SQL coding, experience with large-scale data platforms and ML/optimization tooling.
Python, SQL, Delta Lake, Apache Spark, Databricks, MLflow, Grafana, Elasticsearch, Google OR-Tools, Kubeflow
1mo
Save
Mark Applied
Hide
Performance & Capacity Engineering - Capacity Planning Optimization
Bellevue or Menlo Park or Boston or New York
$184k-$257k/yr OnsiteFull Time
Meta
MetaNASDAQ: META: Develops social networking platforms and virtual reality technologies.
8+ YOEBachelor's in CS/CE or equivalent; 8+ years experience in performance/software/optimization; expertise designing optimization models, LP solvers (Gurobi/Xpress), distributed systems, infrastructure operations, and coding (Python, R, Java, C/C++, PHP).
Python, R, Java, C, C++, PHP, Xpress, Gurobi
2d
Save
Mark Applied
Hide
Sr Data Scientist - Supply Chain Optimization (Middle Mile)
Sunnyvale or Brooklyn Park
$98k-$211k/yr HybridFull Time
Target
TargetNYSE: TGT: Operates a chain of general merchandise stores and supermarkets.
3+ YOEMS/PhD in quantitative field, 3+ years building large-scale optimization/ML solutions, strong coding (Python/R/Java), SQL/Hive experience, supply chain optimization expertise, strong communication.
Python, R, Java, SQL, Hive
2mo
Save
Mark Applied
Hide
Inference Optimization ML Engineer
Palo Alto, California, United States
OnsiteFull Time
Rhoda AI
Rhoda AI: Developing generalist robotic intelligence for real-world industrial automation.
3+ YOE3+ years in inference optimization, ML systems; strong PyTorch; experience with quantization, pruning, distillation; familiarity with Triton/TensorRT; CUDA knowledge.
PyTorch, JAX, TensorRT, Triton, CUDA, XLA, TorchServe, vLLM
2mo
Save
Mark Applied
Hide
HPC/AI Performance Specialist in Bay Area, California, United States
Berkeley, California, United States
$139k-$268k/yr HybridFull Time
Lawrence Berkeley National Laboratory
Lawrence Berkeley National Laboratory: Conducts multidisciplinary scientific research for the U.S. Department of Energy.
8+ YOEExperience optimizing HPC/AI workloads, performance analysis, and software optimization; strong collaboration with scientific teams.
HPC, AI, Performance Profiling, Python, C/C++, CUDA
2mo
Save
Mark Applied
Hide
Staff Advanced Concepts Optimization Engineer
San Jose, California, United States
$130k-$240k/yr OnsiteFull Time
Archer
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
5+ YOESenior/Staff level engineer with 5+ to 12+ years in multidisciplinary optimization, MDO, numerical methods, Python, HPC, and cloud workflows.
Python, NumPy, SciPy, OpenMDAO, JAX, CasADi, SMT, PyTorch
2mo
Save
Mark Applied
Hide
Senior Software Engineering Manager, Ads Auction & Marketplace Optimization
Austin or San Jose
HybridFull Time
Roku
RokuNASDAQ: ROKU: Operates a TV streaming platform and sells streaming hardware.
5+ YOE5+ MgmtLead teams in ads auction and marketplace, build real-time optimization systems, guide roadmap and cross-functional collaboration.
Python, Java
3w
Save
Mark Applied
Hide
Machine Learning Engineer - AI Compiler Optimization
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Proficient with AI compiler frameworks and GPU/NPU compilation optimization; experience with model import/conversion for PyTorch/TensorFlow and performance tuning for recommendation models.
Triton, MLIR, TVM, PyTorch, TensorFlow
1w
Save
Mark Applied
Hide
Research Scientist / Engineer – Performance Optimization
Redwood City, California, United States
OnsiteFull Time
Luma AI
Luma AI: Develops multimodal AI for video generation and creative production.
Expert GPU/CPU/accelerator optimization with Triton/CUDA, strong PyTorch and kernel development, profiling tools experience, deep transformer knowledge, and distributed deployment skills.
Triton, CUDA, PyTorch, NVIDIA Nsight, torch profiler, torch.compile, TensorRT, ONNX, XLA
1mo
Save
Mark Applied
Hide
Member of Research Staff, Optimization
Berkeley or New York City or United States
$250k-$275k/yr HybridFull Time
The Voleon Group
The Voleon Group: Quantitative investment management firm using machine learning strategies.
Ph.D.-level coursework required (Ph.D. preferred); strong background in optimization and numerical methods; applied research track record; production coding in Python and/or C++; strong math and communication skills.
Python, C++
1w
Save
Mark Applied
Hide
Senior Quant Research Engineer, Trading & Portfolio Optimization
Mountain View, California, United States
$110k-$300k/yr HybridFull Time
Arta Finance
Arta Finance: Digital wealth management platform for sophisticated individual investors.
5+ YOE5+ years experience near markets or portfolio management; strong portfolio theory, optimization, risk modeling, software engineering, and tax-aware investing knowledge.
CoderPad, Google Meet, AI coding tools
1mo
Save
Mark Applied
Hide
Software Engineer, ML Infrastructure, Optimization
Mountain View, California, United States
$160k-$241k/yr OnsiteFull Time
Nuro
Nuro: Builds autonomous driving software and electric delivery robots.
2+ YOE2+ years in ML optimization infrastructure; experience with quantization, pruning, ML compilers and GPU runtimes; proficient in Python, C++, CUDA and deep learning frameworks (PyTorch, JAX, TensorFlow, Keras).
Python, C++, CUDA, PyTorch, JAX, TensorFlow, Keras, FTL
1w
Save
Mark Applied
Hide
Sr. Inference Optimization Engineer (local / edge runtime)
Santa Clara or Hillsboro or Folsom or Phoenix
$195k-$361k/yr HybridFull Time
Intel
IntelNasdaq: INTC: Designs and manufactures microprocessors and semiconductor components.
8+ YOE8+ years software development; strong C++ and/or Python; experience with LLM inference, profiling and optimizing CPU/GPU performance; Linux and low-level debugging expertise.
C++, Python, llama.cpp, vLLM, ggml, Vulkan, SYCL, oneAPI, CUDA, Metal, SIMD, Linux, GGUF, AWQ, GPTQ
2w
Save
Mark Applied
Hide
Senior Software Engineer - GPU Kernel Authoring & Optimization
Sunnyvale or Bellevue
$182k-$242k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
5+ YOE5+ years building HPC/GPU software, hands-on CUDA kernel authoring and optimization, C++/Python coding, GPU profiling, and experience delivering performance at scale.
CUDA, Nsight Compute, Nsight Systems, C++, Python, Triton, Mojo, CuTe DSL, JAX, HIP, ROCm, NCCL, Kubernetes, SUNK, Slurm, MLPerf, vLLM, TensorRT-LLM, llm-d, SGLang, KNYFE, Pallas, CUTLASS
1mo
Save
Mark Applied
Hide
ML Engineer - Inference & Model Deployment
Cupertino, California, United States
$250k-$310k/yr OnsiteFull Time
Hiring.Cafe
Hiring.Cafe: An AI-powered job search engine and aggregator.
Experience deploying and optimizing deep learning models in production, multi-GPU inference, profiling/benchmarking model performance, inference optimization techniques, and cloud/distributed systems familiarity.
vLLM, TensorRT, SGLang, GPU
1w
Save
Mark Applied
Hide
Software Development Manager, LLM Inference Model Enablement, Neuron SDK
Cupertino, California, United States
$213k-$288k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
7+ YOE3+ MgmtManage engineering team to onboard and optimize LLMs for inference on Trainium; strong background in LLM architectures, model performance optimization, and inference techniques; experience with PyTorch and Neuron stack.
PyTorch, AWS Neuron, Neuron compiler, Neuron runtime