12 infrastructure hardware systems engineer jobs at 8 companies in Santa Rosa, CA

3mo
Save
Mark Applied
Hide
Infrastructure Engineer (GPU & Compute)
New York or San Francisco or Seattle
$180k-$200k/yr RemoteFull Time
Lightning AI
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
NVIDIA DCGM, PXE, IPMI, Redfish, iDRAC, LiveCD, InfiniBand, NVLink, Linux, Python
3mo
Save
Mark Applied
Hide
Infrastructure Software Engineer
Zurich or United Kingdom or Germany or San Francisco or New York
HybridFull Time
Namespace
Namespace: Cloud infrastructure platform for faster software builds and tests.
Experience engineering large-scale infrastructure; strong network architecture and performance tuning; hands-on hardware deployment; proficient with orchestration and automation tools; track record of reliable systems and efficiency gains.
Go, Kubernetes, Terraform, Ansible
1d
Save
Mark Applied
Hide
Systems Integration Manager | Consumer Devices
San Francisco, California, United States
$290k-$365k/yr HybridFull Time
OpenAI
OpenAI: Develops artificial intelligence models and generative AI software services.
5+ MgmtExperience leading engineering teams in systems integration, test infrastructure, developer tooling, device quality, or hardware-software validation; strong software, embedded, test automation, or systems engineering background; extensive consumer hardware experience.
hardware-in-the-loop, automated test frameworks
1mo
Save
Mark Applied
Hide
Data Center Compute Infrastructure
San Francisco, California, United States
$230k-$490k/yr OnsiteFull Time
OpenAI
OpenAI: Develops artificial intelligence models and generative AI software services.
Experience building, scaling, or operating complex technical systems across software, hardware, supply chain, and data center domains; strong technical judgment and cross-disciplinary collaboration skills.
GPU
3mo
Save
Mark Applied
Hide
Tech Lead, Deployment & Operations — Custom Infrastructure
San Francisco, California, United States
$342k-$445k/yr HybridFull Time
OpenAI
OpenAI: Develops artificial intelligence models and generative AI software services.
8+ YOE8+ years in hardware systems, data center deployment, or related areas; strong leadership, technical depth, and cross-functional collaboration.
hardware deployment, data center operations, infrastructure automation, observability, fleet monitoring
1mo
Save
Mark Applied
Hide
Member of Technical Staff — Inference Infrastructure
San Francisco, California, United States
OnsiteFull Time
Causal Labs
Causal Labs: Building physics-based causal AI models for predictive weather intelligence.
Experience building/optimizing inference and serving systems, distributed compute and GPU parallelism knowledge, familiarity with PyTorch/JAX, hardware-aware optimization, and strong engineering/debugging skills.
TensorRT, Kubernetes, Ray, Slurm, PyTorch, JAX, vLLM, SGLang, Triton
3d
Save
Mark Applied
Hide
Software Engineer, Model Runtime
San Francisco, California, United States
$266k-$445k/yr HybridFull Time
OpenAI
OpenAI: Develops artificial intelligence models and generative AI software services.
Strong systems programming in C++, Rust, or Python; experience with runtimes, distributed systems, compilers, kernels, or serving infrastructure; knowledge of LLM inference and hardware-software performance optimization.
C++, Rust, Python, vLLM, SGLang
2mo
Save
Mark Applied
Hide
Test Software Engineer (Lead)
San Francisco, California, United States
$130k-$160k/yr OnsiteFull Time
Droyd
Droyd: Builds autonomous robotic systems and the software and hardware infrastructure to automate manual work in production environments.
Strong software engineering fundamentals; experience building internal tools, scripts, and test infrastructure; comfortable working close to hardware and debugging system-level issues; ability to set technical direction and lead a test team.
Python
5d
Save
Mark Applied
Hide
Engineering Manager, Fleet Engineering
San Francisco or San Jose or Bellevue
$297k-$440k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
3+ Mgmt3+ years managing engineers in AI/ML infrastructure or large-scale compute; production systems and SLA ownership; Linux, hardware, networking, technical design, troubleshooting, and team-building expertise.
Linux, TCP/IP, PXE, Redfish, IPMI, BMC, DHCP, DNS, InfiniBand, NetBox, Microsoft?, GPU, APIs
1mo
Save
Mark Applied
Hide
Fluidstack Labs Lead
Austin or New York or San Francisco or Seattle
$275k-$327k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience leading hardware qualification, systems engineering, or lab teams for compute/infrastructure; running lab operations; engaging vendors; making deploy decisions based on data.
1mo
Save
Mark Applied
Hide
Software Engineering, Commissioning Automation
Austin or New York City or San Francisco or Seattle
$269k-$317k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building test automation/orchestration for hardware or infrastructure, integrating with control systems, designing auditable systems, and shipping production-quality tooling.
Python, Go, TypeScript, BMS, EPMS
1mo
Save
Mark Applied
Hide
Senior Manager, Technical Support Engineering - Cloud
Livingston or New York or Sunnyvale or San Francisco or Bellevue
$198k-$264k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
5+ YOE5+ Mgmt5+ years leading infrastructure or data center support, Linux administration, hardware diagnostics, GPU infrastructure familiarity, incident/escalation management, ticket systems experience, and metrics-driven ops.
Linux, Jira, Zendesk, NVIDIA A100/H100s