745 ai infrastructure engineer jobs at 301 companies in San Bruno, CA
1w
Save
Mark Applied
Hide
1w
AI Infrastructure Engineer
Boston or San Francisco or Scottsdale or Seattle
$134k-$247k/yrOnsiteFull Time
AxonNASDAQ: AXON: Develops public safety technologies, devices, and cloud software.
4+ YOERequires 4+ years in platform, DevOps, infrastructure, internal tools, automation, or software engineering; cloud infrastructure, CI/CD, infrastructure-as-code, production support, and AI application knowledge.
Together AI: Research-driven AI cloud infrastructure provider offering inference, fine-tuning, GPU clusters, and model training to developers and enterprises.
5+ YOE5+ years in AI infrastructure or related roles; BS in CS or equivalent; knowledge of Ansible, Terraform, Kubernetes; programming/scripting; monitoring/observability; cloud services; collaborative work
Sciforium: AI infrastructure building multimodal models and high-efficiency serving software for developers and teams.
5+ YOE5+ years in systems or infrastructure engineering with GPU, HPC, or ML infrastructure experience; technical bachelor's or master's degree; Linux, Kubernetes, schedulers, configuration management, Python, Bash, containers, GPUs, and RDMA expertise.
Palona AI: AI platform helping restaurants capture demand, convert revenue, and manage operations through voice, text, and visual agents.
3+ YOE3+ years in a relevant technical domain, distributed systems and software engineering experience, cloud platform expertise, infrastructure automation, production debugging, and Python or another modern programming language.
AMAX Engineering CorporationTaiwan Stock Exchange: 6933: Public AI infrastructure engineering serving enterprises with servers, HPC systems, and turnkey data-center solutions.
Experience with on-prem/datacenter operations, Infrastructure-as-Code, networking (VLANs/routing/firewalls), container orchestration, scripting, Git; comfortable with hands-on hardware tasks.
NIONYSE: NIO: Global smart electric vehicle and battery technology provider.
5+ YOERequires 5+ years building AI inference systems, LLM/VLM internals, performance engineering, GPU/NPU programming, C/C++, systems programming, and a BS/MS in computer science, computer engineering, or a related field.
Large Language Models (LLMs), Vision-Language Models (VLMs), GPU, NPU, DSP, CUDA, PyTorch, TensorFlow, C, C++, AIOS
Echelon: AI business-operations platform connecting enterprise HR, finance, and operations data for decision-makers.
5+ YOERequires 5+ years building production infrastructure, distributed systems, developer platforms, or execution runtimes; multi-tenant workloads; strong Linux and programming skills; Kubernetes or comparable scheduler; AWS or Azure experience.
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization
San Jose or Mountain View
$170k-$351k/yrOnsiteFull Time
DiDi Autonomous Driving: Chinese autonomous-driving developing Level 4 self-driving technology, robotaxis, and autonomous trucking logistics for mobility-fleet applications.
3+ YOEMaster's degree or higher in a technical field; 3+ years in HPC, AI infrastructure, model optimization, or embedded deployment; C++, Python, CUDA, OpenMP, inference engines, GPU architectures, and system profiling expertise.
Paradigm: Private technology investment firm building and investing in crypto, AI, robotics, and other frontier technologies.
Experienced engineer with infra, security, and AI model experience; familiarity with durable control planes, runtimes, secrets, observability, and self-hosted deployment.
IntelNasdaq: INTC: Designs and manufactures microprocessors and semiconductor components.
3+ YOEBachelor's degree plus 4+ years, master's plus 3+ years, or PhD; 3+ years in GPU computing, AI systems, or HPC; proficiency in modern C++ and Python.
Luma AI: AI is a private creative AI platform generating video and images for creators and teams.
Deep Linux and distributed systems expertise, experience operating GPU/accelerator clusters, Kubernetes fluency, debugging across hardware/kernel/runtime/orchestration, coding and automation skills, technical leadership and hiring experience.
Woodbridge or New York City or Atlanta or Boston or Chicago or Dallas or Delaware or Denver or Garden City or Cayman Islands or Greenwich or Houston or Los Angeles or Miami or Naples or Nevada or Palm Beach or San Diego or San Francisco or Seattle or Stuart or Washington
$160k-$200k/yrHybridFull Time
Bessemer Investors: Privately owned family office providing investment management, wealth planning, and family-office services to high-net-worth individuals and families.
7+ YOE7+ years in DevOps/platform/cloud infrastructure engineering; strong AWS (IAM, CloudFormation, Lambda, API Gateway, VPC, CloudWatch, SSM, Secrets Manager, ECR); AWS CDK/CloudFormation/Terraform; CI/CD (Bitbucket/GitHub Actions, OIDC); container and datastore operations; security fundamentals.
NIONYSE: NIO: Designs and manufactures premium smart electric vehicles and technology
5+ YOE5+ years building and optimizing large-scale LLM/VLM inference systems; strong C/C++ and performance engineering skills; GPU/NPU programming (CUDA), PyTorch/TensorFlow, and BS/MS in CS/CE or related field required.
Shakudo: Sovereign AI for critical business operations
8+ YOE8+ years engineering experience, 5+ years Kubernetes operation, proficiency in Rust, experience with production infrastructure (physical servers, GPU/DGX clusters), CI/CD, security hardening, observability, and LLM/AI infrastructure.
Tencent Cloud: Cloud is a Chinese cloud-computing provider serving businesses with infrastructure, databases, AI, and digital-transformation services.
Master's or PhD in related field, expertise in GPGPU/AI accelerator architectures, low-level operator development (CUDA,Triton), distributed systems knowledge, and experience optimizing large-scale accelerator clusters and DL frameworks.
Scale AI: Develops data and infrastructure for AI systems.
4+ YOE4+ years building high-performance systems software; deep Linux internals, containerization/virtualization, systems programming (Go/Rust/C/C++); strong debugging and API/SDK design skills.
Senior Applied AI and AI Infrastructure Engineer - Chip Design and DFX
Santa Clara, California, United States
$200k-$380k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
6+ YOERequires BSEE with 12+, MSEE with 10+, or PhD with 6+ years in AI infrastructure, applied machine learning, and generative AI; Python, C++, SQL, ETL, data modeling, and cloud platform expertise.
SQL, ETL, AWS, Microsoft Azure, Google Cloud Platform (GCP), Python, C++
Albany or Arlington or Atlanta or Austin or Beaverton or Bentonville or Boston or Carmel or Charlotte or Chicago or Cincinnati or Cleveland or Columbus or Culver City or Denver or Des Moines or Detroit or Hartford or Houston or Irvine or Irving or Kirkland or Miami or Milwaukee or Minneapolis or Morristown or Mountain View or Nashville or New York City or Oklahoma City or Overland Park or Philadelphia or Pittsburgh or Raleigh or Redmond or Sacramento or San Diego or San Francisco or Scottsdale or Seattle or St. Louis or St. Petersburg or Walnut Creek or United States
$80k-$266k/yrHybridFull Time
AccentureNYSE: ACN: Global professional services delivering 360° value.
5+ YOERequires 5+ years designing AI infrastructure, accelerated computing, clusters, orchestration, and automation; bachelor's degree or equivalent experience. Strong Kubernetes, Python, Terraform, and cloud expertise required.
Senior AI Infrastructure Engineer - Model Training
Mountain View, California, United States
$190k-$260k/yrOnsiteFull Time
Kodiak RoboticsNasdaq: KDK: Public autonomous-vehicle technology serving commercial trucking, industrial trucking, defense, and public-sector customers.
2+ YOEDegree in CS or related field,2+ years ML systems experience,expertise in distributed training,high-performance data pipelines,GPU performance and profiling,Python and PyTorch skills.
New York City or Seattle or San Francisco or London
$140k-$274k/yrHybridFull Time
Writer: Enterprise generative AI platform that helps businesses build and supervise secure AI agents.
5+ YOE5+ years infrastructure/DevOps experience, production Kubernetes, Helm, Terraform/Pulumi, major cloud (AWS preferred), Python or Go, observability stacks (Prometheus, Grafana, ELK), and daily AI-assisted workflows.