206 ai infrastructure engineer jobs at 70 companies in Patterson, CA

1mo
Save
Mark Applied
Hide
AI Infrastructure Engineer
Fremont, California, United States
OnsiteFull Time
AMAX
AMAXTaiwan Stock Exchange: 6933: Designs and manufactures GPU-accelerated AI and HPC computing infrastructure.
Experience with on-prem/datacenter operations, Infrastructure-as-Code, networking (VLANs/routing/firewalls), container orchestration, scripting, Git; comfortable with hands-on hardware tasks.
Terraform, Terragrunt, Ansible, Vault, Boundary, Keycloak, Prometheus, Grafana, Alertmanager, Docker, Kubernetes, Git, Jira, Confluence
8h
Save
Mark Applied
Hide
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization
San Jose or Mountain View
$170k-$351k/yr OnsiteFull Time
DiDi Autonomous Driving
DiDi Autonomous Driving: Develops Level 4 autonomous driving technology for shared mobility.
3+ YOEMaster's degree or higher in a technical field; 3+ years in HPC, AI infrastructure, model optimization, or embedded deployment; C++, Python, CUDA, OpenMP, inference engines, GPU architectures, and system profiling expertise.
C++, Python, CUDA, OpenMP, TensorRT, ONNX Runtime, vLLM, SGLang, TensorRT-LLM, NVIDIA Hopper, NVIDIA Thor, TGI, LightLLM, PyTorch, INT8, FP8, AWQ, LLaMA, Qwen, GPT
6d
Save
Mark Applied
Hide
AI Infrastructure Engineer
Santa Clara or Hillsboro or Folsom or Austin
$171k-$315k/yr HybridFull Time
Intel
IntelNasdaq: INTC: Designs and manufactures microprocessors and semiconductor components.
3+ YOEBachelor's degree plus 4+ years, master's plus 3+ years, or PhD; 3+ years in GPU computing, AI systems, or HPC; proficiency in modern C++ and Python.
C++, Python, vLLM, SGLang, PyTorch, Triton, SYCL, CUDA, CUTLASS, llama.cpp
1mo
Save
Mark Applied
Hide
AI Infrastructure Engineer
San Jose, California, United States
$192k-$250k/yr OnsiteFull Time
NIO
NIONYSE: NIO: Designs and manufactures premium smart electric vehicles and technology
5+ YOE5+ years building and optimizing large-scale LLM/VLM inference systems; strong C/C++ and performance engineering skills; GPU/NPU programming (CUDA), PyTorch/TensorFlow, and BS/MS in CS/CE or related field required.
CUDA, PyTorch, TensorFlow, C/C++, AIOS
1mo
Save
Mark Applied
Hide
Senior Applied AI and AI Infrastructure Engineer - Chip Design and DFX
Santa Clara, California, United States
$200k-$380k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
6+ YOEAdvanced degree or equivalent experience in engineering with 6+–12+ yrs experience in AI infrastructure, applied ML/GenAI; experience with agents, SQL/ETL/data modeling, cloud (AWS/Azure/GCP), Python and C++; strong communication.
SQL, ETL, AWS, Azure, GCP, Python, C++
2mo
Save
Mark Applied
Hide
Senior Applied AI and AI Infrastructure Engineer - Chip Design and DFX
Santa Clara, California, United States
$200k-$380k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOEBSEE/MSEE/PhD with 12+/10+/6+ years experience in AI infrastructure, applied ML and Gen AI. Experience with SQL, ETL, data modeling, cloud (AWS/Azure/GCP), Python, C++, distributed systems, and mentoring.
SQL, ETL, Python, C++, AWS, Azure, GCP
1mo
Save
Mark Applied
Hide
Lead AI Infrastructure Engineer, Reinforcement Learning
Santa Clara, California, United States
$179k-$306k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Design and operate distributed RL training infrastructure at scale; strong systems experience in ML platforms; proficiency with PyTorch/JAX, NCCL/MPI-style distributed training, C++/Python performance tuning; Bachelor's degree required.
PyTorch, JAX, NCCL, MPI, C++, Python
3w
Save
Mark Applied
Hide
Senior AI Platform & Agentic Infrastructure Engineer
San Jose, California, United States
$178k-$321k/yr OnsiteFull Time
OKX
OKX: Global cryptocurrency exchange and Web3 digital wallet developer.
7+ YOE7+ years building and operating resilient backend or platform systems; proven brownfield migrations; fluent Python, strong SQL, and TypeScript/Node or Go; cloud (AWS or GCP), Terraform, CI/CD, observability; data provenance, security, and agentic runtime engineering.
Claude, Anthropic, OpenAI, Google, Google Workspace, JIRA, Lark, Apps Script, Drive API, Docs API, Sheets API, Slides API, Claude Code, Claude Agent SDK, Terraform, Python, SQL, TypeScript, Node, Go, Model Context Protocol (MCP)
1w
Save
Mark Applied
Hide
AI Infrastructure Engineer Intern (Algorithm Infrastructure) - 2027 Start (PhD)
San Jose, California, United States
OnsiteInternship
TikTok
TikTok: Global short-form video hosting and social media platform.
PhD candidate in CS or related field with system design, large-model familiarity, performance optimization, and production engineering skills; experience with distributed scheduling and CUDA/Triton.
CUDA, Triton
1mo
Save
Mark Applied
Hide
Senior MLOps & AI Infrastructure Engineer
San Jose, California, United States
$149k-$216k/yr OnsiteFull Time
Altera
Altera: Manufacturer of field-programmable gate arrays and programmable logic devices.
10+ YOEBachelor's degree, 10+ years ML engineering/MLOps experience, strong Python, cloud ML platforms (AWS/GCP/Azure), Docker/Kubernetes, CI/CD, ML frameworks (PyTorch/TensorFlow/JAX), experience with MLflow/W&B and HPC schedulers.
PyTorch, TensorFlow, JAX, Hugging Face, scikit-learn, XGBoost, MLflow, Kubeflow, Airflow, Weights & Biases, DVC, Feast, AWS SageMaker, GCP Vertex AI, Azure ML, Terraform, CloudFormation, Docker, Kubernetes, Slurm, LSF, Python, Bash, Go, SQL, Prometheus, Grafana, ELK Stack, Evidently AI, Arize
2w
Save
Mark Applied
Hide
Senior Lead AI Engineer (GenAI Platform, Agentic Infrastructure)
New York City or San Francisco or San Jose or Cambridge or McLean or Plano
$209k-$286k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Provides credit card, banking, and auto loan services.
6+ YOEBachelor's degree plus 6 years or master's degree plus 4 years developing AI/ML technologies; 6 years programming with Python, Go, Scala, or Java. Leadership and cloud AI experience preferred.
AWS Ultraclusters, Hugging Face, VectorDBs, Nemo Guardrails, PyTorch, Python, Go, Scala, Java, Google Cloud, Azure, C++, C#, Golang
1mo
Save
Mark Applied
Hide
Senior Lead AI Engineer (GenAI Platform, Agentic Infrastructure)
New York or San Francisco or McLean or Cambridge or San Jose or Plano
$209k-$286k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's in CS/AI/EE/CE +6 years or Master's +4 years; 6+ years programming with Python/Go/Scala/Java; experience deploying scalable AI on cloud; LLM, inference, similarity search, VectorDBs, guardrails, model evaluation, and optimization experience; leadership and research literacy.
Python, Go, Scala, Java, C++, C#, Golang, AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, AWS, Google Cloud, Azure
2mo
Save
Mark Applied
Hide
Principal Engineer, Cloud Data Infrastructure
San Jose or Austin
$218k-$324k/yr HybridFull Time
PayPal
PayPalNASDAQ: PYPL: Digital platform for sending money and processing online payments.
10+ YOE10+ years relevant experience and a Bachelor’s degree (or equivalent); deep expertise in databases, data pipelines, messaging, caching, performance and resilience engineering; experience with real-time analytics and AI/ML infrastructure; strong technical leadership and mentoring.
3mo
Save
Mark Applied
Hide
Senior AI Engineer
San Jose or California
$187k-$215k/yr HybridFull Time
KlearNow.AI
KlearNow.AI: AI platform for digital customs brokerage and logistics management.
4+ YOE4+ years software engineering and infrastructure experience, production LLM/agent experience, Python and TypeScript proficiency, Docker/Kubernetes, cloud/distributed systems, ability to work hybrid in San Jose and authorized to work in the U.S.
Agno, LangChain, LlamaIndex, Semantic Kernel, Claude Code, Cursor, AugmentCode, Python, TypeScript, Docker, Kubernetes, CI/CD, LLMs, RAG
2mo
Save
Mark Applied
Hide
Senior AI Data Infrastructure/Pipeline Engineer
Santa Clara or Mountain View or United States
$175k-$296k/yr OnsiteFull Time
XPeng
XPengNew York Stock Exchange: XPEV: Designs and manufactures smart electric vehicles and autonomous technology.
3+ YOEBachelor's in CS or related, 3+ years in large-scale data processing or data platform development, proficiency in Python/Go/Java, experience with distributed systems, message queues, data lakes, databases, Kubernetes/Docker, and strong cross-team communication.
Python, Go, Java, Kafka, Pulsar, RabbitMQ, Apache Iceberg, Lance, MySQL, PostgreSQL, Redis, MongoDB, Kubernetes, Docker, GitHub
2mo
Save
Mark Applied
Hide
Helix AI Engineer, Backend Infrastructure
San Jose, California, United States
$150k-$400k/yr OnsiteFull Time
Figure
Figure: Develops autonomous humanoid robots for commercial and residential tasks.
5+ YOESenior backend engineer skilled in high-throughput, real-time streaming, cloud systems, and ML-serving integration.
Go, C++, Python, Rust, AWS, GCP, Azure, Docker, Kubernetes, Service Mesh, CI/CD
1mo
Save
Mark Applied
Hide
Sr. Staff AI Engineer - On-Prem AI Infrastructure & Agentic Systems
San Jose, California, United States
$140k-$165k/yr OnsiteFull Time
SK hynix Memory Solutions America
SK hynix Memory Solutions AmericaKorea Exchange: 000660: Develops semiconductor controllers and firmware for enterprise data storage.
2+ YOE2+ years in AI/ML engineering with on-prem/private-cloud deployment, experience building agentic AI and RAG pipelines, model fine-tuning (LoRA/QLoRA), Python/Linux/Docker/Kubernetes proficiency, and familiarity with vector DBs and AI serving frameworks.
vLLM, TGI, Triton, Milvus, Qdrant, FAISS, Kubernetes, Helm, Docker, LangGraph, AutoGen, LoRA, QLoRA, Model Control Protocols (MCP), Python, Linux, Pinecone, Ollama, ONNX, TensorRT, GGUF, BabyAGI, LangSmith, Weights & Biases, Prometheus, Grafana, CI/CD
1mo
Save
Mark Applied
Hide
AI Platform Engineer
Milpitas, California, United States
$136k-$232k/yr OnsiteFull Time
KLA
KLANASDAQ: KLAC: Provides process control and yield management for semiconductor manufacturing.
5+ YOEDegree in CS/CE or related, 5+ years systems/DevOps/ML infrastructure experience, hands-on AI/GPU cluster and Linux/Kubernetes expertise, Python/Bash scripting, and knowledge of storage, networking, and security.
Linux, Kubernetes, Docker, Slurm, Ray, TensorFlow, PyTorch, GPU, TPU, Bash, Python, CI/CD, Infrastructure as Code (IaC), TPM
1mo
Save
Mark Applied
Hide
Principal Infrastructure Engineer
Pleasanton, California, United States
$170k-$230k/yr HybridFull Time
Blackhawk Network
Blackhawk Network: Provider of global branded payment and gift card solutions.
12+ YOEBA plus 12+ years IAM experience; mastery of IAM (lifecycle, RBAC, IGA, PAM), protocols (SAML, OAuth, OIDC, SCIM, LDAP, Kerberos), directory services (Active Directory, Azure AD), Okta, ServiceNow, compliance frameworks, and AI tool fluency.
SAML, OAuth, OIDC, SCIM, LDAP, Kerberos, Microsoft Active Directory, Microsoft Azure AD, Okta, ServiceNow, AI tools
1w
Save
Mark Applied
Hide
AI Infrastructure Engineer Intern (Compute Efficiency & Scheduling) - 2027 Summer
San Jose, California, United States
OnsiteInternship
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Currently pursuing undergraduate or master’s in CS or related field; able to commit to a 12-week Summer 2027 internship; coding experience in Go,C++,Python,Java,C#; interest in systems/distributed computing.
Godel, Go, C++, Python, Java, C#, Linux, containers, Kubernetes