567 gpu jobs at 75 companies in Hollister, CA

3w
Save
Mark Applied
Hide
GPU Performance Architect
Folsom or Santa Clara
$127k-$217k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Develop performance models and prototypes for GPU/SoC systems; analyze ML/HPC workloads; write and optimize GPU kernels; strong programming in C/C++/Python; knowledge of CUDA/OpenCL and ML frameworks.
TensorFlow, PyTorch, CUDA, OpenCL, C, C++, Python, Hip, Triton, MLIR, LLVM, RTL, System C
2mo
Save
Mark Applied
Hide
Senior GPU Memory Architect
Santa Clara, California, United States
$184k-$357k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOEMaster's degree or equivalent in Electrical Engineering, Computer Science, or Computer Engineering; 6+ years in GPU/CPU memory architecture; strong communication; mentoring experience.
GPU Architecture, Memory Subsystems, Memory Modeling, RTL Simulation, Performance Modeling
1mo
Save
Mark Applied
Hide
Senior GPU Capacity Planner
San Francisco or Sunnyvale or Bellevue
$160k-$195k/yr OnsiteFull Time
Crusoe
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
3+ YOE3+ years capacity planning or systems engineering experience; hyperscaler cloud experience; GPU topology knowledge (NVIDIA H100/B200); Bachelor’s or Master’s in quantitative field; strong cross-functional communication and modeling skills.
AWS, GCP, Azure, Oracle Cloud, NVIDIA H100, NVIDIA B200
3w
Save
Mark Applied
Hide
Senior Staff Engineer, GPU Architect, Machine Learning
San Jose, California, United States
$198k-$297k/yr OnsiteFull Time
Samsung Electronics
Samsung ElectronicsKorea Exchange: 005930: Develops and manufactures consumer electronics, semiconductors, and mobile devices.
11+ YOE11+ years with BS or 9+ with MS or 7+ with PhD; expertise in GPU architecture, ML, workload analysis; strong C++/Python skills; familiarity with TensorFlow and PyTorch.
C++, Python, TensorFlow, PyTorch
1mo
Save
Mark Applied
Hide
Sr. Staff/Principal Engineer — GPU Driver & Systems Software
San Diego or San Jose
$179k-$286k/yr OnsiteFull Time
MediaTek
MediaTekTaiwan Stock Exchange: 2454: Designs and develops system-on-chip solutions for electronic devices.
10+ YOE10+ years building GPU drivers or low-level systems software; fluent C/C++; deep knowledge of Vulkan/DirectX/OpenCL/OpenGL ES and GPU internals; experience with kernel-mode drivers, firmware, performance optimization.
C, C++, Vulkan, DirectX, OpenGL ES, OpenCL
1mo
Save
Mark Applied
Hide
GPU Performance Engineer
Sunnyvale or Bellevue
$120k-$160k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
2+ YOE2+ years software engineering experience; proficiency in Go and/or Python; production Kubernetes experience; develop performance tests, automation, and platform tooling; participate in on-call rotation.
Go, Python, Kubernetes
1w
Save
Mark Applied
Hide
GPU ML Engineer
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Work on high-performance data-parallel algorithms in linear algebra, image processing, and machine learning for Apple platforms.
iOS, macOS, Apple TV
4d
Save
Mark Applied
Hide
Staff Software Engineer, GPU Inference
Toronto or Sunnyvale
HybridFull Time
Cerebras Systems
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
8+ YOE8+ years software engineering experience, strong C++ and Python skills, GPU inference experience, Linux, containers and Kubernetes, benchmarking and production optimization for latency-sensitive services.
C++, Python, vLLM, PyTorch, ROCm, HIP, RCCL, rocprofiler, AMD SMI, AITER, hipBLASLt, Composable Kernel, CUDA, SGLang, TensorRT-LLM, Triton Inference Server, Kubernetes, Linux, RDMA
1mo
Save
Mark Applied
Hide
Sr. GPU/Accelerator Hardware Development Engineer
Austin or Cupertino
$159k-$248k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
5+ YOEBachelor's in EE/CE, 5+ years in high-speed system design and validation, experience with schematic/layout tools, lab test equipment, component selection, supply chain and full product lifecycle.
Schematic and layout tools, bench power supplies, high-speed oscilloscopes, logic analyzers, spectrum analyzers, VNA, thermal chambers
1w
Save
Mark Applied
Hide
System Engineer, GPU Server
San Jose, California, United States
$90k-$110k/yr OnsiteFull Time
Supermicro
SupermicroNASDAQ: SMCI: Designs and manufactures high-performance server and storage solutions.
2+ YOEBachelor's or Master's in CS/EE/Computer Engineering,2+ years system/server architecture experience,GPU and networking knowledge preferred,ability to diagnose failures,basic scripting,CRM familiarity.
CRM
3w
Save
Mark Applied
Hide
GPU/AI Application Platform Architect - San Jose
San Jose, California, United States
$137k-$360k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
3+ YOEMaster's degree in EE/CE/CS, deep GPU/AI system architecture knowledge, experience in GPU/AI application performance optimization and LLM requirements; passport and travel readiness.
1w
Save
Mark Applied
Hide
Principal Program Manager – GPU Platform & Infrastructure (OCI)
Santa Clara, California, United States
$102k-$210k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
7+ YOEBachelor's degree or equivalent,7+ years in supply/demand planning or supply chain leadership,expert Excel,program management and cross-functional communication,experience with datacenter/infrastructure preferred.
Microsoft Excel, OCI, AWS, Azure, Google Cloud
3w
Save
Mark Applied
Hide
Sr. Software Development Engineer - Video Rendering
Seattle or San Francisco or San Jose
$174k-$331k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years software engineering with deep GPU/graphics or rendering systems experience, strong modern C++, and knowledge of GPU APIs, shading languages, and performance debugging.
C++, DirectX 12, Metal, Vulkan, CUDA, OpenCL, HLSL, Slang, SPIR-V, DXIL, DXC, Nsight, PIX, TensorRT, ONNX Runtime
1mo
Save
Mark Applied
Hide
Principal Technical Product Marketing Manager
El Dorado Hills or San Jose
HybridFull Time
Blaize
BlaizeNASDAQ: BZAI: Designs AI processors and edge-to-cloud software platforms.
8+ YOE8+ years in product/technical marketing or technical GTM for semiconductors, AI infrastructure, cloud platforms, or developer tools; deep technical fluency across inference, accelerators, GPU/non-GPU architectures, and hybrid topologies; strong storytelling and partner GTM experience; BS/MS in CE/EE/CS or equivalent.
Blaize AI Services, Blaize GSP, GPU
2d
Save
Mark Applied
Hide
Senior Technical Program Manager II, AI/ML Systems
Sunnyvale, California, United States
$240k-$333k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
10+ YOE10+ MgmtBachelor's in a technical field or equivalent,10+ years program management, experience with ML/GPU/TPU systems, strong communication and roadmap skills.
GPU, TPUs, Vertex AI, Google Cloud
1w
Save
Mark Applied
Hide
Performance Engineer
San Jose, California, United States
$135k-$170k/yr OnsiteFull Time
Astera Labs
Astera LabsNASDAQ: ALAB: Designs connectivity solutions for cloud and AI infrastructure.
2+ YOEBachelor's in engineering/computer science, experience benchmarking and characterizing GPU cluster performance, proficiency with performance tooling and scripting (Python); strong systems and networking knowledge.
NVBandwidth, NCCL, Confluence, CUDA, MPI, Python, COSMOS, NVLink, PCIe, Ethernet, UALink, UEC
3w
Save
Mark Applied
Hide
Applied Scientist - LLM Training System as a Service - Global Frontier Tech Recruitment Program - 2027 Start (PhD)
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
PhD candidate in CS or related field with deep learning knowledge, Python and CUDA proficiency, experience or interest in distributed LLM training and GPU performance optimization.
Python, Pytorch, CUDA, FSDP, Deepspeed, JAX, Megatron-LM, Verl, TensorRT-LLM, ORCA, VLLM, SGLang
1mo
Save
Mark Applied
Hide
Inference Optimization Intern – Performance Modeling
Sunnyvale, California, United States
OnsiteInternship
Institute of Foundation Models
Institute of Foundation Models: Develops open-source frontier-class AI foundation models and research.
Currently pursuing a quantitative degree; experience or coursework in CUDA, GPU kernel development, performance modeling, Nsight profiling, C++, Python, and deep learning frameworks preferred.
CUDA, Nsight Systems, Nsight Compute, PTX, SASS, PyTorch, TensorFlow, C++, Python
1w
Save
Mark Applied
Hide
Sr. Inference Optimization Engineer (local / edge runtime)
Santa Clara or Hillsboro or Folsom or Phoenix
$195k-$361k/yr HybridFull Time
Intel
IntelNasdaq: INTC: Designs and manufactures microprocessors and semiconductor components.
8+ YOE8+ years software development; strong C++ and/or Python; experience with LLM inference, profiling and optimizing CPU/GPU performance; Linux and low-level debugging expertise.
C++, Python, llama.cpp, vLLM, ggml, Vulkan, SYCL, oneAPI, CUDA, Metal, SIMD, Linux, GGUF, AWQ, GPTQ
1mo
Save
Mark Applied
Hide
ML Engineer - Inference & Model Deployment
Cupertino, California, United States
$250k-$310k/yr OnsiteFull Time
Hiring.Cafe
Hiring.Cafe: An AI-powered job search engine and aggregator.
Experience deploying and optimizing deep learning models in production, multi-GPU inference, profiling/benchmarking model performance, inference optimization techniques, and cloud/distributed systems familiarity.
vLLM, TensorRT, SGLang, GPU