4 sre intern jobs at 3 companies in San Rafael, CA

3mo
Save
Mark Applied
Hide
Infrastructure Intern
Sunnyvale, California, United States
HybridFull Time, Internship
Meshy
Meshy: AI-powered platform for rapid 3D model and asset generation.
Pursuing Bachelor's, Master's, or PhD in CS/Software Eng or related field; strong Python/Go skills; Linux and cloud familiarity; Kubernetes, Terraform, and IaC exposure; interest in infra, DevOps, SRE, and AI infrastructure.
Python, Go, Kubernetes, Docker, Terraform, Linux, Cloud (AWS GCP Azure)
1w
Save
Mark Applied
Hide
IT Systems Engineer - Internal Platforms & SRE
San Francisco or San Jose
$206k-$275k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
Experience with system design, scalable cloud infrastructure, configuration management, programming in Python or Go, distributed systems, automation, documentation, and cross-functional collaboration.
AWS, GCP, Azure, Chef, Ansible, Terraform, GitHub Actions, Python, Go
2w
Save
Mark Applied
Hide
Senior Production Engineer, Storage
Sunnyvale, California, United States
$170k-$205k/yr OnsiteFull Time
Crusoe
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
5+ YOEBachelor's or equivalent experience,5+ years storage SRE/systems experience,enterprise storage knowledge,programming (Go,Python,Java,C),IaC (Terraform/Ansible/Puppet),Linux internals,Kubernetes and cloud experience.
Pure Storage, EMC, Go, Python, Java, C, Terraform, Ansible, Puppet, Linux, NFS, SMB, iSCSI, NVMe-oF, Kubernetes, Docker, AWS, GCP, Azure, Ceph, GlusterFS, OpenEBS, Vast, Lightbits
3mo
Save
Mark Applied
Hide
Staff Software Engineer - Managed Kubernetes
Bellevue or San Francisco or San Jose
$314k-$465k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
10+ YOE10+ years in software/platform engineering or SRE; 5+ years Kubernetes at scale; strong Go and Python; deep Kubernetes internals; GPU orchestration; multi-tenant infrastructure; distributed systems; observability at scale; Linux networking; IaC and GitOps.
Go, Python, Kubernetes, NVIDIA GPU Operator, NCCL, DCGM, GPUDirect, CNI, InfiniBand, RDMA, GaP? note: ignore invalid, GitOps