745 ai infrastructure engineer jobs at 301 companies in San Bruno, CA

1w
Save
Mark Applied
Hide
AI Infrastructure Engineer
Boston or San Francisco or Scottsdale or Seattle
$134k-$247k/yr OnsiteFull Time
Axon
AxonNASDAQ: AXON: Develops public safety technologies, devices, and cloud software.
4+ YOERequires 4+ years in platform, DevOps, infrastructure, internal tools, automation, or software engineering; cloud infrastructure, CI/CD, infrastructure-as-code, production support, and AI application knowledge.
Vercel, Azure, AWS, GitHub Actions, Terraform, Bicep, Pulumi, CloudFormation, Python, TypeScript, JavaScript, Node.js, GCP, SSO, OAuth, OIDC, Entra ID, Azure AD, RBAC, Slack, Jira, Confluence, Quip, Microsoft 365, Salesforce, Snowflake, ServiceNow, React
3mo
Save
Mark Applied
Hide
AI Infrastructure Engineer
San Francisco, California, United States
$190k-$270k/yr OnsiteFull Time
Together AI
Together AI: Research-driven AI cloud infrastructure provider offering inference, fine-tuning, GPU clusters, and model training to developers and enterprises.
5+ YOE5+ years in AI infrastructure or related roles; BS in CS or equivalent; knowledge of Ansible, Terraform, Kubernetes; programming/scripting; monitoring/observability; cloud services; collaborative work
Ansible, Terraform, Kubernetes
3w
Save
Mark Applied
Hide
AI Infrastructure Engineer
San Francisco, California, United States
$150k-$220k/yr OnsiteFull Time
Sciforium
Sciforium: AI infrastructure building multimodal models and high-efficiency serving software for developers and teams.
5+ YOE5+ years in systems or infrastructure engineering with GPU, HPC, or ML infrastructure experience; technical bachelor's or master's degree; Linux, Kubernetes, schedulers, configuration management, Python, Bash, containers, GPUs, and RDMA expertise.
Ansible, SaltStack, Git, Python, Bash, Kubernetes, NVIDIA GPU Operator, Slurm, Run:AI, enroot, pyxis, Docker, containerd, NVIDIA Container Toolkit, CUDA, cuDNN, NCCL, Fabric Manager, ROCm, RCCL, DKMS, GPUDirect RDMA, GPUDirect Storage, MOFED, DOCA, PyTorch, JAX, DCGM exporter, Prometheus, Grafana, PXE, MaaS, Packer, Foreman, Terraform, Lustre, GPFS, Weka, vLLM, Triton Inference Server, TensorRT-LLM, Nsight Systems, Nsight Compute, rocprof, perf, eBPF, EMR
2w
Save
Mark Applied
Hide
AI Infrastructure Engineer
Los Altos, California, United States
OnsiteFull Time
Palona AI
Palona AI: AI platform helping restaurants capture demand, convert revenue, and manage operations through voice, text, and visual agents.
3+ YOE3+ years in a relevant technical domain, distributed systems and software engineering experience, cloud platform expertise, infrastructure automation, production debugging, and Python or another modern programming language.
Python, Docker, AWS, Azure, ECS, Lambda, API Gateway, OpenTofu, Terraform, Datadog, CI/CD
1mo
Save
Mark Applied
Hide
AI Infrastructure Engineer
Fremont, California, United States
OnsiteFull Time
AMAX Engineering Corporation
AMAX Engineering CorporationTaiwan Stock Exchange: 6933: Public AI infrastructure engineering serving enterprises with servers, HPC systems, and turnkey data-center solutions.
Experience with on-prem/datacenter operations, Infrastructure-as-Code, networking (VLANs/routing/firewalls), container orchestration, scripting, Git; comfortable with hands-on hardware tasks.
Terraform, Terragrunt, Ansible, Vault, Boundary, Keycloak, Prometheus, Grafana, Alertmanager, Docker, Kubernetes, Git, Jira, Confluence
2mo
Save
Mark Applied
Hide
AI Infrastructure Engineer
San Jose, California, United States
$192k-$250k/yr OnsiteFull Time
NIO
NIONYSE: NIO: Global smart electric vehicle and battery technology provider.
5+ YOERequires 5+ years building AI inference systems, LLM/VLM internals, performance engineering, GPU/NPU programming, C/C++, systems programming, and a BS/MS in computer science, computer engineering, or a related field.
Large Language Models (LLMs), Vision-Language Models (VLMs), GPU, NPU, DSP, CUDA, PyTorch, TensorFlow, C, C++, AIOS
2w
Save
Mark Applied
Hide
Senior/Staff AI Infrastructure Engineer
San Francisco, California, United States
$180k-$250k/yr OnsiteFull Time
Echelon
Echelon: AI business-operations platform connecting enterprise HR, finance, and operations data for decision-makers.
5+ YOERequires 5+ years building production infrastructure, distributed systems, developer platforms, or execution runtimes; multi-tenant workloads; strong Linux and programming skills; Kubernetes or comparable scheduler; AWS or Azure experience.
Linux, Kubernetes, AWS, Azure, Go, Rust, TypeScript, Python, Firecracker, gVisor, Kata Containers, eBPF, E2B, Modal, Fly Machines, FUSE, Kafka, NATS, Inngest, Temporal
2w
Save
Mark Applied
Hide
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization
San Jose or Mountain View
$170k-$351k/yr OnsiteFull Time
DiDi Autonomous Driving
DiDi Autonomous Driving: Chinese autonomous-driving developing Level 4 self-driving technology, robotaxis, and autonomous trucking logistics for mobility-fleet applications.
3+ YOEMaster's degree or higher in a technical field; 3+ years in HPC, AI infrastructure, model optimization, or embedded deployment; C++, Python, CUDA, OpenMP, inference engines, GPU architectures, and system profiling expertise.
C++, Python, CUDA, OpenMP, TensorRT, ONNX Runtime, vLLM, SGLang, TensorRT-LLM, NVIDIA Hopper, NVIDIA Thor, TGI, LightLLM, PyTorch, INT8, FP8, AWQ, LLaMA, Qwen, GPT
3w
Save
Mark Applied
Hide
Infrastructure Engineer, Applied AI
San Francisco, California, United States
$250k-$400k/yr OnsiteFull Time
Paradigm
Paradigm: Private technology investment firm building and investing in crypto, AI, robotics, and other frontier technologies.
Experienced engineer with infra, security, and AI model experience; familiarity with durable control planes, runtimes, secrets, observability, and self-hosted deployment.
Reth, Foundry, EVMBench, OpenAI, Centaur, Kubernetes
3w
Save
Mark Applied
Hide
AI Infrastructure Engineer
Santa Clara or Hillsboro or Folsom or Austin
$171k-$315k/yr HybridFull Time
Intel
IntelNasdaq: INTC: Designs and manufactures microprocessors and semiconductor components.
3+ YOEBachelor's degree plus 4+ years, master's plus 3+ years, or PhD; 3+ years in GPU computing, AI systems, or HPC; proficiency in modern C++ and Python.
C++, Python, vLLM, SGLang, PyTorch, Triton, SYCL, CUDA, CUTLASS, llama.cpp
1mo
Save
Mark Applied
Hide
Staff AI Infrastructure Engineer
Redwood City, California, United States
HybridFull Time
Luma AI
Luma AI: AI is a private creative AI platform generating video and images for creators and teams.
Deep Linux and distributed systems expertise, experience operating GPU/accelerator clusters, Kubernetes fluency, debugging across hardware/kernel/runtime/orchestration, coding and automation skills, technical leadership and hiring experience.
Linux, Kubernetes
2mo
Save
Mark Applied
Hide
Gen AI Infrastructure Engineer
Woodbridge or New York City or Atlanta or Boston or Chicago or Dallas or Delaware or Denver or Garden City or Cayman Islands or Greenwich or Houston or Los Angeles or Miami or Naples or Nevada or Palm Beach or San Diego or San Francisco or Seattle or Stuart or Washington
$160k-$200k/yr HybridFull Time
Bessemer Investors
Bessemer Investors: Privately owned family office providing investment management, wealth planning, and family-office services to high-net-worth individuals and families.
7+ YOE7+ years in DevOps/platform/cloud infrastructure engineering; strong AWS (IAM, CloudFormation, Lambda, API Gateway, VPC, CloudWatch, SSM, Secrets Manager, ECR); AWS CDK/CloudFormation/Terraform; CI/CD (Bitbucket/GitHub Actions, OIDC); container and datastore operations; security fundamentals.
AWS Bedrock, AgentCore, Lambda, API Gateway, AWS CDK, CloudFormation, Terraform, Bitbucket Pipelines, GitHub Actions, OIDC, IAM, VPC, CloudWatch, SSM, Secrets Manager, ECR, Neo4j, Neptune, Redis, Milvus, VectorDB
2mo
Save
Mark Applied
Hide
AI Infrastructure Engineer
San Jose, California, United States
$192k-$250k/yr OnsiteFull Time
NIO
NIONYSE: NIO: Designs and manufactures premium smart electric vehicles and technology
5+ YOE5+ years building and optimizing large-scale LLM/VLM inference systems; strong C/C++ and performance engineering skills; GPU/NPU programming (CUDA), PyTorch/TensorFlow, and BS/MS in CS/CE or related field required.
CUDA, PyTorch, TensorFlow, C/C++, AIOS
2mo
Save
Mark Applied
Hide
Infrastructure Engineer
Menlo Park, California, United States
OnsiteFull Time
Shakudo
Shakudo: Sovereign AI for critical business operations
8+ YOE8+ years engineering experience, 5+ years Kubernetes operation, proficiency in Rust, experience with production infrastructure (physical servers, GPU/DGX clusters), CI/CD, security hardening, observability, and LLM/AI infrastructure.
Kubernetes, Rust, CI/CD, DGX, GPU, LLM, ETL
3w
Save
Mark Applied
Hide
Sr. Cloud AI Infrastructure Engineer
Palo Alto, California, United States
$145k-$273k/yr OnsiteFull Time
Tencent Cloud
Tencent Cloud: Cloud is a Chinese cloud-computing provider serving businesses with infrastructure, databases, AI, and digital-transformation services.
Master's or PhD in related field, expertise in GPGPU/AI accelerator architectures, low-level operator development (CUDA,Triton), distributed systems knowledge, and experience optimizing large-scale accelerator clusters and DL frameworks.
CUDA, Triton, PyTorch, TensorFlow
1mo
Save
Mark Applied
Hide
AI Infrastructure Engineer, Sandbox Platform
San Francisco or Seattle or New York City
$180k-$225k/yr OnsiteFull Time
Scale AI
Scale AI: Develops data and infrastructure for AI systems.
4+ YOE4+ years building high-performance systems software; deep Linux internals, containerization/virtualization, systems programming (Go/Rust/C/C++); strong debugging and API/SDK design skills.
Docker, Firecracker, gVisor, QEMU, Kata Containers, Go, Rust, C/C++, Kubernetes, OpenHands, Agent2Agent, MCP, CRIU
3w
Save
Mark Applied
Hide
Senior Applied AI and AI Infrastructure Engineer - Chip Design and DFX
Santa Clara, California, United States
$200k-$380k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
6+ YOERequires BSEE with 12+, MSEE with 10+, or PhD with 6+ years in AI infrastructure, applied machine learning, and generative AI; Python, C++, SQL, ETL, data modeling, and cloud platform expertise.
SQL, ETL, AWS, Microsoft Azure, Google Cloud Platform (GCP), Python, C++
1w
Save
Mark Applied
Hide
AI & HPC Infrastructure Engineer
Albany or Arlington or Atlanta or Austin or Beaverton or Bentonville or Boston or Carmel or Charlotte or Chicago or Cincinnati or Cleveland or Columbus or Culver City or Denver or Des Moines or Detroit or Hartford or Houston or Irvine or Irving or Kirkland or Miami or Milwaukee or Minneapolis or Morristown or Mountain View or Nashville or New York City or Oklahoma City or Overland Park or Philadelphia or Pittsburgh or Raleigh or Redmond or Sacramento or San Diego or San Francisco or Scottsdale or Seattle or St. Louis or St. Petersburg or Walnut Creek or United States
$80k-$266k/yr HybridFull Time
Accenture
AccentureNYSE: ACN: Global professional services delivering 360° value.
5+ YOERequires 5+ years designing AI infrastructure, accelerated computing, clusters, orchestration, and automation; bachelor's degree or equivalent experience. Strong Kubernetes, Python, Terraform, and cloud expertise required.
Slurm, Run:ai, Kubernetes, NVIDIA Base Command Manager (BCM), NVIDIA NGC, NCCL, NVLink, CUDA, TensorRT-LLM, vLLM, SGLang, Triton Inference Server, NVIDIA Dynamo, llm-d, MLPerf, NCCL tests, fio, iperf, InfiniBand, Ethernet, SONiC, NVMe, NVMe-oF, VAST, Weka, DDN, AWS, Azure, GCP, VMware, Nutanix, Python, Terraform, Ansible, REST, OpenAPI, JSON, YAML, TensorFlow, PyTorch, JAX, Jupyter notebooks, Google Colab, MCP
1mo
Save
Mark Applied
Hide
Senior AI Infrastructure Engineer - Model Training
Mountain View, California, United States
$190k-$260k/yr OnsiteFull Time
Kodiak Robotics
Kodiak RoboticsNasdaq: KDK: Public autonomous-vehicle technology serving commercial trucking, industrial trucking, defense, and public-sector customers.
2+ YOEDegree in CS or related field,2+ years ML systems experience,expertise in distributed training,high-performance data pipelines,GPU performance and profiling,Python and PyTorch skills.
PyTorch, PyTorch DDP/FSDP, DeepSpeed, Megatron, NCCL, WebDataset, MosaicML Streaming, MDS, Nsight, PyTorch Profiler, Python, C++, CUDA, Triton, NVLink, InfiniBand
4w
Save
Mark Applied
Hide
Infrastructure engineer
New York City or Seattle or San Francisco or London
$140k-$274k/yr HybridFull Time
Writer
Writer: Enterprise generative AI platform that helps businesses build and supervise secure AI agents.
5+ YOE5+ years infrastructure/DevOps experience, production Kubernetes, Helm, Terraform/Pulumi, major cloud (AWS preferred), Python or Go, observability stacks (Prometheus, Grafana, ELK), and daily AI-assisted workflows.
Python, Go, AWS, GCP, Azure, Kubernetes, Helm, Terraform, Pulumi, Claude Code, Droid, Codex, Prometheus, Grafana, ELK