560 optimizer jobs at 195 companies in Gilroy, CA

2mo
Save
Mark Applied
Hide
Software Engineer, Marketplace Pricing
Mountain View or San Francisco
$175k-$215k/yr OnsiteFull Time
Waymo
Waymo: Autonomous driving technology for ride-hailing and logistics.
4+ YOEBS in CS or equivalent; 4+ years backend experience; experience building distributed backend systems; preferred C++, MS CS; experience with low-latency, large-scale distributed systems; ML/optimization infrastructure and production models.
C++, Python, Java, Distributed Systems, Machine Learning, Optimization
1w
Save
Mark Applied
Hide
AI Systems, Model Optimization
Palo Alto or United States
HybridFull Time
Unconventional
Unconventional: Developing novel computing hardware for efficient AI acceleration.
MS/PhD (or equivalent) in quantitative field, deep practical experience with ML stack and GPU performance optimization, proficiency in profiling and optimizing large ML codebases.
PyTorch, torch.compile, DDP, FSDP, CUDA, Triton, CUTLASS, Megatron-LM, DeepSpeed
3w
Save
Mark Applied
Hide
Senior Data Scientist, Cloud Gaming - Prescriptive Analytics and Optimization
Santa Clara or United States
$184k-$288k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOE6+ years experience (or PhD) in quantitative field, strong prescriptive analytics and optimization background, Python and SQL coding, experience with large-scale data platforms and ML/optimization tooling.
Python, SQL, Delta Lake, Apache Spark, Databricks, MLflow, Grafana, Elasticsearch, Google OR-Tools, Kubeflow
1mo
Save
Mark Applied
Hide
Performance & Capacity Engineering - Capacity Planning Optimization
Bellevue or Menlo Park or Boston or New York
$184k-$257k/yr OnsiteFull Time
Meta
MetaNASDAQ: META: Develops social networking platforms and virtual reality technologies.
8+ YOEBachelor's in CS/CE or equivalent; 8+ years experience in performance/software/optimization; expertise designing optimization models, LP solvers (Gurobi/Xpress), distributed systems, infrastructure operations, and coding (Python, R, Java, C/C++, PHP).
Python, R, Java, C, C++, PHP, Xpress, Gurobi
3d
Save
Mark Applied
Hide
Sr Data Scientist - Supply Chain Optimization (Middle Mile)
Sunnyvale or Brooklyn Park
$98k-$211k/yr HybridFull Time
Target
TargetNYSE: TGT: Operates a chain of general merchandise stores and supermarkets.
3+ YOEMS/PhD in quantitative field, 3+ years building large-scale optimization/ML solutions, strong coding (Python/R/Java), SQL/Hive experience, supply chain optimization expertise, strong communication.
Python, R, Java, SQL, Hive
2mo
Save
Mark Applied
Hide
Inference Optimization ML Engineer
Palo Alto, California, United States
OnsiteFull Time
Rhoda AI
Rhoda AI: Developing generalist robotic intelligence for real-world industrial automation.
3+ YOE3+ years in inference optimization, ML systems; strong PyTorch; experience with quantization, pruning, distillation; familiarity with Triton/TensorRT; CUDA knowledge.
PyTorch, JAX, TensorRT, Triton, CUDA, XLA, TorchServe, vLLM
2mo
Save
Mark Applied
Hide
Staff Advanced Concepts Optimization Engineer
San Jose, California, United States
$130k-$240k/yr OnsiteFull Time
Archer
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
5+ YOESenior/Staff level engineer with 5+ to 12+ years in multidisciplinary optimization, MDO, numerical methods, Python, HPC, and cloud workflows.
Python, NumPy, SciPy, OpenMDAO, JAX, CasADi, SMT, PyTorch
2mo
Save
Mark Applied
Hide
Senior Software Engineering Manager, Ads Auction & Marketplace Optimization
Austin or San Jose
HybridFull Time
Roku
RokuNASDAQ: ROKU: Operates a TV streaming platform and sells streaming hardware.
5+ YOE5+ MgmtLead teams in ads auction and marketplace, build real-time optimization systems, guide roadmap and cross-functional collaboration.
Python, Java
3w
Save
Mark Applied
Hide
Machine Learning Engineer - AI Compiler Optimization
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Proficient with AI compiler frameworks and GPU/NPU compilation optimization; experience with model import/conversion for PyTorch/TensorFlow and performance tuning for recommendation models.
Triton, MLIR, TVM, PyTorch, TensorFlow
1w
Save
Mark Applied
Hide
Research Scientist / Engineer – Performance Optimization
Redwood City, California, United States
OnsiteFull Time
Luma AI
Luma AI: Develops multimodal AI for video generation and creative production.
Expert GPU/CPU/accelerator optimization with Triton/CUDA, strong PyTorch and kernel development, profiling tools experience, deep transformer knowledge, and distributed deployment skills.
Triton, CUDA, PyTorch, NVIDIA Nsight, torch profiler, torch.compile, TensorRT, ONNX, XLA
1w
Save
Mark Applied
Hide
Senior Quant Research Engineer, Trading & Portfolio Optimization
Mountain View, California, United States
$110k-$300k/yr HybridFull Time
Arta Finance
Arta Finance: Digital wealth management platform for sophisticated individual investors.
5+ YOE5+ years experience near markets or portfolio management; strong portfolio theory, optimization, risk modeling, software engineering, and tax-aware investing knowledge.
CoderPad, Google Meet, AI coding tools
1mo
Save
Mark Applied
Hide
Software Engineer, ML Infrastructure, Optimization
Mountain View, California, United States
$160k-$241k/yr OnsiteFull Time
Nuro
Nuro: Builds autonomous driving software and electric delivery robots.
2+ YOE2+ years in ML optimization infrastructure; experience with quantization, pruning, ML compilers and GPU runtimes; proficient in Python, C++, CUDA and deep learning frameworks (PyTorch, JAX, TensorFlow, Keras).
Python, C++, CUDA, PyTorch, JAX, TensorFlow, Keras, FTL
1w
Save
Mark Applied
Hide
Sr. Inference Optimization Engineer (local / edge runtime)
Santa Clara or Hillsboro or Folsom or Phoenix
$195k-$361k/yr HybridFull Time
Intel
IntelNasdaq: INTC: Designs and manufactures microprocessors and semiconductor components.
8+ YOE8+ years software development; strong C++ and/or Python; experience with LLM inference, profiling and optimizing CPU/GPU performance; Linux and low-level debugging expertise.
C++, Python, llama.cpp, vLLM, ggml, Vulkan, SYCL, oneAPI, CUDA, Metal, SIMD, Linux, GGUF, AWQ, GPTQ
2w
Save
Mark Applied
Hide
Senior Software Engineer - GPU Kernel Authoring & Optimization
Sunnyvale or Bellevue
$182k-$242k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
5+ YOE5+ years building HPC/GPU software, hands-on CUDA kernel authoring and optimization, C++/Python coding, GPU profiling, and experience delivering performance at scale.
CUDA, Nsight Compute, Nsight Systems, C++, Python, Triton, Mojo, CuTe DSL, JAX, HIP, ROCm, NCCL, Kubernetes, SUNK, Slurm, MLPerf, vLLM, TensorRT-LLM, llm-d, SGLang, KNYFE, Pallas, CUTLASS
1mo
Save
Mark Applied
Hide
ML Engineer - Inference & Model Deployment
Cupertino, California, United States
$250k-$310k/yr OnsiteFull Time
Hiring.Cafe
Hiring.Cafe: An AI-powered job search engine and aggregator.
Experience deploying and optimizing deep learning models in production, multi-GPU inference, profiling/benchmarking model performance, inference optimization techniques, and cloud/distributed systems familiarity.
vLLM, TensorRT, SGLang, GPU
1w
Save
Mark Applied
Hide
Software Development Manager, LLM Inference Model Enablement, Neuron SDK
Cupertino, California, United States
$213k-$288k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
7+ YOE3+ MgmtManage engineering team to onboard and optimize LLMs for inference on Trainium; strong background in LLM architectures, model performance optimization, and inference techniques; experience with PyTorch and Neuron stack.
PyTorch, AWS Neuron, Neuron compiler, Neuron runtime
1mo
Save
Mark Applied
Hide
Power Systems Research Scientist
Cupertino, California, United States
$175k-$235k/yr OnsiteFull Time
Gridmatic
Gridmatic: AI-powered platform for optimizing energy trading and battery storage.
Advanced degree in EE/power systems, strong power systems modeling and optimization background, experience with power flow, transmission/congestion analysis, large-scale optimization, and Python programming.
Python, PSS/E, PowerWorld, PSLF, PyTorch, JAX
1mo
Save
Mark Applied
Hide
Optimization Engineer, Grid Systems
Redwood City, California, United States
$140k-$220k/yr HybridFull Time
GridCARE
GridCARE: Unlocking electrical grid capacity for AI data center power.
1+ YOEMaster's or equivalent in a quantitative field, 1+ years industry or applied research experience (1–5 years typical), production-quality Python or Julia coding, optimization solvers and simulation experience, strong math and software engineering.
Python, Julia
2w
Save
Mark Applied
Hide
Staff Computational Scientist, Grid Optimization, Tapestry
Mountain View, California, United States
$207k-$300k/yr HybridFull Time
Alphabet
AlphabetNASDAQ: GOOGL: Holding providing internet, software, and AI services.
6+ YOEMaster's or PhD in STEM, 6+ years in modeling and optimization, strong power systems knowledge, numerical methods and solver experience, ability to produce production-ready code, collaborative across disciplines.
Python, Java, C/C++, JAX, Julia, AWS, GCP, Azure
1w
Save
Mark Applied
Hide
Senior Machine Learning Engineer, LLM Inference Optimization
Palo Alto or California
$195k-$262k/yr OnsiteFull Time
Nebius
NebiusNasdaq: NBIS: Builds cloud infrastructure and software for artificial intelligence development.
Expert Python and PyTorch skills, hands-on LLM/VLM inference deployment and optimization, knowledge of modern inference stacks, quantitative reasoning about latency/throughput/cost, and strong communication.
Python, PyTorch, vLLM, SGLang, TensorRT-LLM, Triton Inference Server, NVIDIA Dynamo, Ray Serve, KServe, CUDA, FlashInfer, LMCache, Ray