1,061 site engineer jobs at 558 companies in California
1w
Save
Mark Applied
Hide
1w
Site Reliability Engineer Intern (Data Infra) - 2027 Fall
San Jose, California, United States
OnsiteInternship
ByteDance: Global technology specializing in AI-powered content platforms.
Currently pursuing a bachelor's degree in computer science or related technical discipline; programming experience in C, C++, Java, Python, Go, or Rust; knowledge of Unix/Linux internals, networking, and distributed systems.
Northwood Space: Northwood is an end-to-end ground infrastructure provider for space missions, delivering hardware, software, and network services.
5+ YOERequires 5+ years of relevant experience, a bachelor's degree in mechanical, civil, electrical, or related engineering, and willingness to travel domestically and internationally. Construction, facilities, operations, and dashboard experience preferred.
Washington or Salt Lake City or Albuquerque or Elgin or Cleveland or Denver or Houston or Indianapolis or Los Angeles or Dallas-Fort Worth or Seattle or Minneapolis or Oakland
$65k-$95k/yrFieldFull Time
P17 Solutions: 8(a)-certified professional-services contractor providing information technology, engineering, and program-management services to government agencies.
7+ years in telecommunications network engineering, circuit integration, or infrastructure planning; expertise in BGP, TCP/IP, MPLS, IPv6, fiber, servers, and TDM-to-IP migration. U.S. citizenship and Public Trust eligibility required.
BGP, TCP/IP, MPLS, IPv6, TDM, AWS, Microsoft Azure, Google Cloud Platform (GCP)
3+ YOEBachelor's degree in computer science or related field; 3+ years in site reliability engineering; 2+ years with AWS and cloud automation; Kubernetes, Linux, Terraform, networking, GitOps, monitoring, and customer support experience.
AWS, Kubernetes, Helm, Linux, Terraform, GitOps, Prometheus, Grafana, Bazel, CueLang, Version Control, Okta, Snowflake, Google
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
BrazeNASDAQ: BRZE: Customer engagement platform for cross-channel marketing and analytics.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
ShopifyNasdaq: SHOP: Provides internet infrastructure and tools for commerce.
Experienced SRE/engineer with on-call experience, ability to build resilient production tooling, improve observability, respond to alerts, and collaborate across engineering teams.
BAE SystemsLSE: BA.: Global defense, aerospace, and security technology.
4+ YOERequires 4–6+ years of site reliability engineering, Juniper networking, cloud technologies, automation, storage, virtualization, and security clearance eligibility; Security+ required or obtainable within 90 days.
Thinking Machines Lab: Private AI research and product building customizable multimodal systems for researchers and the wider public.
Experience in distributed systems/cloud/site reliability, software automation for reliability, incident response and postmortems, strong communication and coordination skills.
Arena Intelligence: AI model evaluation platform serving enterprises, AI labs, and independent researchers in real-world workflows.
6+ YOE6+ years backend engineering with distributed systems, proficiency in Go or Rust, experience with LLM provider APIs, cloud (AWS/GCP), Kubernetes, Terraform, Postgres, and Redis.
STN Incorporated: U.S.-based IT infrastructure provider delivering managed cloud, cybersecurity, and GPU compute services to enterprises and AI teams.
5+ YOE5+ years in SRE/DevOps or production engineering; strong Go and/or Python skills; Kubernetes at scale; observability with Prometheus, Grafana, Datadog, OpenTelemetry; incident management and on-call experience.
Runloop AI: Runloop AI provides AI infrastructure, secure code sandboxes, and evaluation tools for developers building software-engineering agents.
5+ YOE5+ years software engineering experience with 3+ years in SRE/DevOps, strong Python or Go skills, containerization, cloud infra, monitoring, networking, Linux administration, on‑call and incident management.
SpaceXNasdaq: SPCX: Designing, manufacturing, and launching advanced rockets and spacecraft.
1+ YOE1+ years hands-on experience with client/server hardware, networking, Linux/Windows, scripting and automation; bachelor's in CS/engineering/math or 2+ years software experience in lieu; HPC and systems engineering experience preferred.
Charm Industrial: U.S. carbon-removal converts agricultural and forest biomass into bio-oil and injects it underground for permanent storage.
8+ YOEBachelor's degree in engineering, Colorado civil PE license, and 8+ years designing or executing heavy industrial facilities. Requires multidisciplinary coordination, industrial layouts, site infrastructure, and CAD proficiency.
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
8+ YOE8+ years supporting live-site production environments; BS/MS or equivalent; strong Kubernetes, AWS, Python; Akamai/CDN and SRE on-call experience required.
Cerebras SystemsNasdaq Global Select Market: CBRS: Designs processors and systems for AI training and inference.
15+ YOE15+ years in SRE/infrastructure/platform engineering with large-scale fleets; experience in capacity management, orchestration, observability, SLOs/SLIs, incident response, and cross-team architecture.
IXL Learning: Private educational technology providing personalized K–12 learning software and resources for students, teachers, and families.
6+ YOEBachelor's degree,6+ years SRE/software engineering,experience with OO and scripting languages,cloud (AWS/GCP),Docker/Kubernetes,monitoring,on-call availability,strong troubleshooting and communication skills.
Green Dot CorporationNYSE: GDOT: Public U.S. fintech bank holding providing banking and payment services to consumers and businesses.
7+ YOE7+ years in release/reliability engineering, cloud platform experience (AWS/Azure/GCP), automated deployment and observability proficiency, scripting with PowerShell/Bash/Python, excellent troubleshooting and communication skills.
SalesforceNYSE: CRM: The #1 AI CRM driving customer success together.
5+ YOE5+ years systems and software engineering experience for large-scale internet services; expertise in SRE principles, containers, observability, incident management, Python and Go, and applying AI/ML to operations.