259 ai infrastructure engineer jobs at 163 companies in Petaluma, CA

1w
Save
Mark Applied
Hide
AI Infrastructure Engineer
Boston or San Francisco or Scottsdale or Seattle
$134k-$247k/yr OnsiteFull Time
Axon
AxonNASDAQ: AXON: Develops weapons and software for law enforcement and safety.
4+ YOERequires 4+ years in platform, DevOps, infrastructure, internal tools, automation, or software engineering; cloud infrastructure, CI/CD, infrastructure-as-code, production support, and AI application knowledge.
Vercel, Azure, AWS, GitHub Actions, Terraform, Bicep, Pulumi, CloudFormation, Python, TypeScript, JavaScript, Node.js, GCP, SSO, OAuth, OIDC, Entra ID, Azure AD, RBAC, Slack, Jira, Confluence, Quip, Microsoft 365, Salesforce, Snowflake, ServiceNow, React
2w
Save
Mark Applied
Hide
AI Infrastructure Engineer
San Francisco, California, United States
$150k-$220k/yr OnsiteFull Time
Sciforium
Sciforium: Building multimodal AI models and high-performance model serving infrastructure.
5+ YOE5+ years in systems or infrastructure engineering with GPU, HPC, or ML infrastructure experience; technical bachelor's or master's degree; Linux, Kubernetes, schedulers, configuration management, Python, Bash, containers, GPUs, and RDMA expertise.
Ansible, SaltStack, Git, Python, Bash, Kubernetes, NVIDIA GPU Operator, Slurm, Run:AI, enroot, pyxis, Docker, containerd, NVIDIA Container Toolkit, CUDA, cuDNN, NCCL, Fabric Manager, ROCm, RCCL, DKMS, GPUDirect RDMA, GPUDirect Storage, MOFED, DOCA, PyTorch, JAX, DCGM exporter, Prometheus, Grafana, PXE, MaaS, Packer, Foreman, Terraform, Lustre, GPFS, Weka, vLLM, Triton Inference Server, TensorRT-LLM, Nsight Systems, Nsight Compute, rocprof, perf, eBPF, EMR
3mo
Save
Mark Applied
Hide
AI Infrastructure Engineer
San Francisco, California, United States
$190k-$270k/yr OnsiteFull Time
Together AI
Together AI: Cloud platform for training and deploying artificial intelligence models.
5+ YOE5+ years in AI infrastructure or related roles; BS in CS or equivalent; knowledge of Ansible, Terraform, Kubernetes; programming/scripting; monitoring/observability; cloud services; collaborative work
Ansible, Terraform, Kubernetes
2mo
Save
Mark Applied
Hide
Cloud Infrastructure and AI Efficiency Engineer
San Francisco, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Engineering role focused on cloud infrastructure and AI efficiency; specific experience and certifications not provided in the description.
3w
Save
Mark Applied
Hide
Infrastructure Engineer, Applied AI
San Francisco, California, United States
$250k-$400k/yr OnsiteFull Time
Paradigm
Paradigm: Venture capital firm focused on crypto and frontier technologies.
Experienced engineer with infra, security, and AI model experience; familiarity with durable control planes, runtimes, secrets, observability, and self-hosted deployment.
Reth, Foundry, EVMBench, OpenAI, Centaur, Kubernetes
2mo
Save
Mark Applied
Hide
Gen AI Infrastructure Engineer
Woodbridge or New York City or Atlanta or Boston or Chicago or Dallas or Delaware or Denver or Garden City or Cayman Islands or Greenwich or Houston or Los Angeles or Miami or Naples or Nevada or Palm Beach or San Diego or San Francisco or Seattle or Stuart or Washington
$160k-$200k/yr HybridFull Time
Bessemer Trust
Bessemer Trust: Wealth management and family office services for affluent clients.
7+ YOE7+ years in DevOps/platform/cloud infrastructure engineering; strong AWS (IAM, CloudFormation, Lambda, API Gateway, VPC, CloudWatch, SSM, Secrets Manager, ECR); AWS CDK/CloudFormation/Terraform; CI/CD (Bitbucket/GitHub Actions, OIDC); container and datastore operations; security fundamentals.
AWS Bedrock, AgentCore, Lambda, API Gateway, AWS CDK, CloudFormation, Terraform, Bitbucket Pipelines, GitHub Actions, OIDC, IAM, VPC, CloudWatch, SSM, Secrets Manager, ECR, Neo4j, Neptune, Redis, Milvus, VectorDB
1mo
Save
Mark Applied
Hide
AI Infrastructure Engineer, Sandbox Platform
San Francisco or Seattle or New York City
$180k-$225k/yr OnsiteFull Time
Scale AI
Scale AI: Provides data and infrastructure for training artificial intelligence models.
4+ YOE4+ years building high-performance systems software; deep Linux internals, containerization/virtualization, systems programming (Go/Rust/C/C++); strong debugging and API/SDK design skills.
Docker, Firecracker, gVisor, QEMU, Kata Containers, Go, Rust, C/C++, Kubernetes, OpenHands, Agent2Agent, MCP, CRIU
1w
Save
Mark Applied
Hide
AI & HPC Infrastructure Engineer
Albany or Arlington or Atlanta or Austin or Beaverton or Bentonville or Boston or Carmel or Charlotte or Chicago or Cincinnati or Cleveland or Columbus or Culver City or Denver or Des Moines or Detroit or Hartford or Houston or Irvine or Irving or Kirkland or Miami or Milwaukee or Minneapolis or Morristown or Mountain View or Nashville or New York City or Oklahoma City or Overland Park or Philadelphia or Pittsburgh or Raleigh or Redmond or Sacramento or San Diego or San Francisco or Scottsdale or Seattle or St. Louis or St. Petersburg or Walnut Creek or United States
$80k-$266k/yr HybridFull Time
Accenture
AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
5+ YOERequires 5+ years designing AI infrastructure, accelerated computing, clusters, orchestration, and automation; bachelor's degree or equivalent experience. Strong Kubernetes, Python, Terraform, and cloud expertise required.
Slurm, Run:ai, Kubernetes, NVIDIA Base Command Manager (BCM), NVIDIA NGC, NCCL, NVLink, CUDA, TensorRT-LLM, vLLM, SGLang, Triton Inference Server, NVIDIA Dynamo, llm-d, MLPerf, NCCL tests, fio, iperf, InfiniBand, Ethernet, SONiC, NVMe, NVMe-oF, VAST, Weka, DDN, AWS, Azure, GCP, VMware, Nutanix, Python, Terraform, Ansible, REST, OpenAPI, JSON, YAML, TensorFlow, PyTorch, JAX, Jupyter notebooks, Google Colab, MCP
3w
Save
Mark Applied
Hide
Infrastructure engineer
New York City or Seattle or San Francisco or London
$140k-$274k/yr HybridFull Time
Writer
Writer: Platform for building and deploying enterprise generative AI agents.
5+ YOE5+ years infrastructure/DevOps experience, production Kubernetes, Helm, Terraform/Pulumi, major cloud (AWS preferred), Python or Go, observability stacks (Prometheus, Grafana, ELK), and daily AI-assisted workflows.
Python, Go, AWS, GCP, Azure, Kubernetes, Helm, Terraform, Pulumi, Claude Code, Droid, Codex, Prometheus, Grafana, ELK
1mo
Save
Mark Applied
Hide
Lead Infrastructure Engineer-Network Engineer
San Francisco or Seattle
$143k-$185k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years infrastructure engineering experience, formal training/certification, deep cloud and network knowledge, scripting and automation experience, familiarity with security/segmentation and AI-assisted engineering, strong problem-solving and mentoring skills.
Cisco, Juniper, Arista, Juniper Mist, Cisco/Viptela, Fortinet, BGP, OSPF, MPLS, EVPN/VXLAN, SD-WAN, Wi Fi 6E, Wi Fi 7, 5G, Palo Alto, Zscaler, IPsec, TLS, ZTNA, Python, Ansible, Terraform, Git, GitHub Actions, Jenkins, ThousandEyes, Splunk, Grafana, Wireshark
2mo
Save
Mark Applied
Hide
Forward Deployed Infrastructure Engineer
San Francisco, California, United States
OnsiteFull Time
Sirius Technology: AI-powered retention platform for subscription-based businesses.
1+ YOE1+ year in infrastructure/DevOps/platform engineering or solutions architecture with customer-facing deployment experience; deep experience with AWS, Terraform, container orchestration, and cloud networking (VPC, IAM, DNS); strong operational and stakeholder communication skills.
AWS, Terraform, VPC, IAM, DNS, container orchestration, CRM, AI/ML, LLM
3mo
Save
Mark Applied
Hide
AI Engineer, Agent Infrastructure
San Francisco, California, United States
OnsiteFull Time
Zed
Zed: AI-native neobank providing premium credit services to young professionals.
Experience shipping production LLM/agent systems; strong backend/infrastructure; familiarity with workflows, observability, and production readiness.
Python, Go, REST APIs, Docker, Kubernetes, LLMs, Observability, Monitoring
5d
Save
Mark Applied
Hide
Staff Infrastructure Engineer
United States or San Francisco or New York City or Chicago
$187k-$235k/yr RemoteFull Time
Komodo Health
Komodo Health: Provides AI-driven healthcare data and patient journey analytics software.
8+ YOERequires 8+ years in infrastructure, cloud, or platform engineering; deep AWS experience; Terraform and Kubernetes ownership; AI-assisted engineering fluency; and security/compliance experience in regulated environments.
Amazon Web Services (AWS), Kubernetes, Terraform, ArgoCD, GitHub Actions, Claude, OpenAI Codex, Cursor, GitHub Copilot, AWS Bedrock, Okta, OpenID Connect (OIDC), Security Assertion Markup Language (SAML), Envoy, Python, Go, Bash, Snowflake, IAM, RBAC
2mo
Save
Mark Applied
Hide
Founding Engineer, AI Infra
San Francisco, California, United States
HybridFull Time
Phenix Space
Phenix Space: Builds robotic systems for on-orbit satellite assembly and upgrades.
5+ YOE5+ years building or operating ML infrastructure; deep GPU and distributed training knowledge; experience with PyTorch/DeepSpeed/Megatron/Ray, inference stacks (vLLM, TGI, Triton), Python and C++/Rust/Go, Kubernetes and IaC, and observability tooling.
FlashAttention, CUDA, Triton, PyTorch, DeepSpeed, Megatron, Ray, vLLM, SGLang, TGI, Python, C++, Rust, Go, Kubernetes, Terraform, Pulumi, Prometheus, Grafana, OpenTelemetry, Llama 3, Qwen, DeepSeek
3mo
Save
Mark Applied
Hide
Infrastructure Engineer
New York or San Francisco or United States
$165k-$200k/yr HybridFull Time
Roboflow
Roboflow: Platform for building and deploying custom computer vision models.
Kubernetes production experience; IaC (Terraform/Helm); cloud (AWS/GCP); Python/Node.js; CI/CD (GitHub Actions/Spacelift); security and ML/AI infrastructure familiarity.
Kubernetes, Terraform, Helm, Python, Node.js, GitHub Actions, Spacelift, AWS, GCP, PyTorch, TensorFlow, Bash
3mo
Save
Mark Applied
Hide
AI Engineer
San Francisco, California, United States
OnsiteFull Time
Emanate
Emanate: A technology building AI-powered revenue infrastructure.
Backend/AI engineer role focusing on AI infrastructure, LLM integration, data pipelines, and scalable autonomous workflows.
2mo
Save
Mark Applied
Hide
Founding Infrastructure Engineer
San Francisco, California, United States
$130k-$200k/yr OnsiteFull Time
Virio
Virio: A fast-growing technology startup focused on scalable infrastructure for AI-driven workloads.
Build and scale distributed infrastructure, own core services, design for reliability and performance, and collaborate with product and AI teams.
3w
Save
Mark Applied
Hide
Senior Lead AI Engineer (GenAI Platform, Agentic Infrastructure)
New York City or San Francisco or San Jose or Cambridge or McLean or Plano
$209k-$286k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Provides credit card, banking, and auto loan services.
6+ YOEBachelor's degree plus 6 years or master's degree plus 4 years developing AI/ML technologies; 6 years programming with Python, Go, Scala, or Java. Leadership and cloud AI experience preferred.
AWS Ultraclusters, Hugging Face, VectorDBs, Nemo Guardrails, PyTorch, Python, Go, Scala, Java, Google Cloud, Azure, C++, C#, Golang
3mo
Save
Mark Applied
Hide
Staff Software Engineer - AI Research Infrastructure
San Francisco or New York City
$199k-$270k/yr HybridFull Time
Databricks
Databricks: A unified platform for data analytics and artificial intelligence.
5+ YOEBS/MS or PhD in computer science; 5+ years of software engineering experience in distributed systems or infrastructure; strong experience with GPUs, clusters, and cloud platforms; proficient in systems languages and large-scale job orchestration.
C++, Rust, Go, Java, Scala, Kubernetes, Slurm, Ray, GPU, Cloud computing, Distributed systems
2mo
Save
Mark Applied
Hide
Senior Lead AI Engineer (GenAI Platform, Agentic Infrastructure)
New York or San Francisco or McLean or Cambridge or San Jose or Plano
$209k-$286k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's in CS/AI/EE/CE +6 years or Master's +4 years; 6+ years programming with Python/Go/Scala/Java; experience deploying scalable AI on cloud; LLM, inference, similarity search, VectorDBs, guardrails, model evaluation, and optimization experience; leadership and research literacy.
Python, Go, Scala, Java, C++, C#, Golang, AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, AWS, Google Cloud, Azure