98 network reliability engineer jobs at 55 companies in California

1mo
Save
Mark Applied
Hide
Senior Network & Site Reliability Engineer
San Francisco, California, United States
$210k-$240k/yr OnsiteFull Time
Alembic
Alembic: AI-powered marketing attribution and revenue forecasting platform.
8+ YOE8+ years in network or infrastructure engineering (5+ years datacenter ops); strong network security and architecture skills; hands-on with BGP, QoS, MPLS, IPsec, EVPN/VXLAN, ECMP; IaC (Ansible, Terraform, Nornir); NetBox/Infoblox; Kubernetes networking; Linux; monitoring stacks; Python/Bash.
NVIDIA DGX SuperPOD, Grace Blackwell, BGP, VPNs, WAN, QoS, MPLS, IPsec, EVPN, VXLAN, ECMP, Ansible, Terraform, Nornir, NetBox, Infoblox, Kubernetes, Prometheus, Grafana, Datadog, ELK, OpenTelemetry, Python, Bash, Cumulus Linux, InfiniBand, Spectrum-X, BlueField, Spark, Airflow, Kafka, NFS, LustreFS, iSCSI, Linux
3w
Save
Mark Applied
Hide
Senior Network Systems Reliability Engineer
Lake Buena Vista or Burbank
$135k-$181k/yr HybridFull Time
The Walt Disney Company
The Walt Disney CompanyNYSE: DIS: Produces movies, operates theme parks, and provides streaming services.
5+ YOE5+ years supporting network infrastructure/IT operations; familiarity with Python and SQL; experience with network technologies, asset/license management, ticketing systems, Excel/Pivot reporting; strong analytical, vendor management, and documentation skills.
Python, SQL, LogicMonitor, Cisco DNA Center, Aruba Central, Power BI, ServiceNow, Jira Service Management, Microsoft Excel
2d
Save
Mark Applied
Hide
Site Reliability Engineer (Network)
San Francisco or Golden
$157k-$239k/yr OnsiteFull Time
Loft Orbital
Loft Orbital: Deploy and operate satellite missions for organizations and governments.
4+ YOE4–5 years network engineering experience, hands-on SDN and public-cloud networking (ideally GCP), Kubernetes/Docker familiarity, IaC (Terraform) and GitOps, SRE mindset (SLOs, observability), degree or equivalent experience.
GCP, k8s, Docker, Terraform, Grafana, ArgoCD, FluxCD, Cockpit, GitOps
1mo
Save
Mark Applied
Hide
Senior Network Reliability Engineer - DGX Cloud
Santa Clara or United States
$136k-$265k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOE5+ years network operations experience; deep knowledge of TCP/IP, BGP, OSPF, MPLS, EVPN/VxLAN and related protocols; experience in CSPs (AWS, Microsoft Azure, GCP, OCI); scripting/automation and vendor familiarity (Arista, Juniper, Fortinet).
TCP/IP, BGP, OSPF, MPLS, IS-IS, VxLAN, EVPN, QoS, GRE, IPsec, DNS, MACsec, AWS, Microsoft Azure, GCP, OCI, Arista, Fortinet, Juniper, Mellanox, Cumulus OS, NetBox, Nautobot, Prometheus, Grafana, Panoptes, Python, Shell
3w
Save
Mark Applied
Hide
Senior Network Systems Reliability Engineer
Lake Buena Vista or Burbank
$135k-$181k/yr HybridFull Time
The Walt Disney Company
The Walt Disney CompanyNYSE: DIS: Produces media content and operates global theme parks.
5+ YOE5+ years supporting network infrastructure/IT operations, familiarity with Python and SQL, strong Excel and reporting skills, experience with network technologies, asset/license lifecycle management, vendor management, and ticketing systems.
Python, SQL, Microsoft Excel, LogicMonitor, Cisco DNA Center, Aruba Central, Power BI, ServiceNow, Jira Service Management
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineering - Network
Palo Alto or Columbus
$152k-$215k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE10+ MgmtFormal network engineering training, 5+ years applied experience, 10+ years leading technologists, advanced network reliability skills, SD-WAN and cloud (AWS, Azure) proficiency, major network vendor experience, observability tooling and incident leadership.
SD-WAN, AWS, Azure, Palo Alto, Juniper, F5, Broadcom, Arista, Cisco, Grafana, SevOne, Prometheus, Kibana, ThousandEyes, Splunk, Jenkins, GitLab, Terraform, eBPF, TCP/IP, HTTPS, BGP
2mo
Save
Mark Applied
Hide
Senior Network Reliability Engineer - DGX Cloud
Santa Clara or United States
$136k-$265k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOE5+ years in network operations; strong TCP/IP, BGP, OSPF, MPLS, EVPN; experience with AWS/Azure/GCP; hands-on with automation; Bachelor’s in CS or related field.
TCP/IP, BGP, OSPF, MPLS, EVPN, VxLAN, GRE, IPsec, DNS, MACsec, Arista, Fortinet, Juniper, Mellanox, Cumulus OS, Infiniband, Netbox, Nautobot, Prometheus, Grafana, Python, Shell, AWS, Azure, GCP, OCI
2mo
Save
Mark Applied
Hide
Senior Director, Network Reliability
San Ramon or United States
$136k-$448k/yr HybridFull Time
Five9
Five9NASDAQ: FIVN: Provides cloud-based software for enterprise contact center operations.
10+ YOE10+ MgmtLeads global production networks; builds and mentors teams; drives engineering-led operations and modernization.
OSPF, EVPN, VXLAN, Cisco Nexus, NX-OS, Nexus, GCP, Terraform, Ansible, Python, Git, CI/CD, Kubernetes, ThousandEyes, Grafana, Loki, Netflow, Netscout, IaC
2mo
Save
Mark Applied
Hide
Senior Systems Reliability Engineer
California or Oregon
$150k-$225k/yr RemoteFull Time
IEX
IEX: Operates a stock exchange using proprietary market integrity technology.
Automation experience with Ansible or similar tools; Linux, Python, Bash, Git; experience with large distributed systems; networking and data center knowledge; able to troubleshoot across hardware, software, and network.
Ansible, Linux, Python, Bash, Git, Arista, Cisco, Corvil, Solarflare, Mellanox, TCP/IP, Networking, Packet Analysis
1w
Save
Mark Applied
Hide
Site Reliability Engineer
San Francisco, California, United States
HybridFull Time
Runloop
Runloop: Provides infrastructure and secure sandboxes for AI agents.
5+ YOE5+ years software engineering experience with 3+ years in SRE/DevOps, strong Python or Go skills, containerization, cloud infra, monitoring, networking, Linux administration, on‑call and incident management.
AWS, GCP, Azure, Grafana, Prometheus, Datadog, Python, Go, Docker, Kubernetes, Terraform, Pulumi, Sentry, RUM, CI/CD
1w
Save
Mark Applied
Hide
Site Reliability Engineer (Raptor)
Hawthorne, California, United States
$125k-$175k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
1+ YOE1+ years hands-on experience with client/server hardware, networking, Linux/Windows, scripting and automation; bachelor's in CS/engineering/math or 2+ years software experience in lieu; HPC and systems engineering experience preferred.
Infiniband, ANSYS, StarCCM+, Bash, Python, Puppet, Ansible, Kubernetes, Docker, Linux, Windows
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Palo Alto or Pittsburgh
$179k-$269k/yr OnsiteFull Time
Latitude AI
Latitude AI: Developing automated driving technology for next-generation Ford vehicles.
4+ YOEBachelor's degree in engineering/computer science (or higher) with 4+ years experience (or equivalent), strong Linux, networking, Go/Python development, cloud (AWS/GCP), Kubernetes, IaC, monitoring and SLO experience.
Go, Python, AWS, GCP, Terraform, CloudFormation, Kubernetes, Prometheus, Elasticsearch, Loki, Jaeger, Tempo, Linux
2w
Save
Mark Applied
Hide
Site Reliability Engineer
San Francisco, California, United States
OnsiteFull Time
Specter: Building a software-defined perception engine for the physical world.
Strong Linux administration, experience with edge/on‑prem hardware and cloud (AWS), networking fundamentals, scripting in Python/Go/Bash, containerization (Docker, Kubernetes) and embedded/firmware familiarity; on‑call participation.
AWS, Bash, C, Docker, Go, Kubernetes, Linux, Python, Rust, SSH, DNS, VPN, IAM
1w
Save
Mark Applied
Hide
Site Reliability Engineer
Research Triangle Park or San Jose or Milpitas or Richardson or Santa Clara
$127k-$182k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOE5+ years SRE/Cloud Ops experience, Docker and Kubernetes proficiency, scripting in Python/Go/Bash, monitoring and incident response experience, Linux and networking knowledge, CI/CD and IaC familiarity, bachelor’s degree or equivalent.
Docker, Kubernetes, Python, Go, Bash, Git
1w
Save
Mark Applied
Hide
Site Reliability Engineer
Santa Clara, California, United States
$230k-$250k/yr OnsiteFull Time
Forward Networks
Forward Networks: Provides a digital twin platform for enterprise network management.
6+ YOE6+ years SRE/DevOps experience in SaaS/cloud, strong networking fundamentals, Kubernetes, observability (Prometheus/Grafana/Datadog/Splunk), Python/Bash automation, cloud and IaC (AWS/GCP/Azure, Terraform/Ansible), and incident response ownership.
Kubernetes, Prometheus, Grafana, Datadog, Splunk, Python, Bash, AWS, GCP, Azure, Terraform, Ansible
3mo
Save
Mark Applied
Hide
Site Reliability Engineer — Human Engineering
Cupertino or California or United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
5+ YOEBS in CS/Engineering with 5+ years distributed systems; Kubernetes in production; AWS experience; IaC; backend language; CI/CD; networking; strong communication; automation experience.
Kubernetes, AWS, Terraform, Helm, Python, Go, CI/CD, GitOps, Kafka, PostgreSQL, Redis, Elasticsearch, Django, Envoy, Celery, Gunicorn
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer - SDN
San Francisco or San Jose or Bellevue
$240k-$312k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
Kubernetes, SmartNICs, Python, Ansible, Go, C, Helm, Terraform, GitOps, CI/CD, Linux, OpenStack Neutron, OVN, OVS, DPDK, SR-IOV
2w
Save
Mark Applied
Hide
Senior Systems Reliability Engineer
Pune or San Jose or Durham or Mexico City or Bangalore or Hoofddorp or Belgrade or Barcelona or Singapore or Sydney or Tokyo
HybridFull Time
Nutanix
NutanixNASDAQ: NTNX: Sells cloud software and hyperconverged infrastructure for enterprises.
7+ YOE7+ years SRE experience with networking, virtualization (VMware ESXi), Linux, cloud and strong customer-facing troubleshooting and communication skills.
VMware ESXi, VMware, Linux, DevOps, Cloud, Citrix, Microsoft
3mo
Save
Mark Applied
Hide
Observability Lead - Cloud SRE & Network Reliability
Fremont, California, United States
$114k-$253k/yr HybridFull Time
Lam Research
Lam ResearchNASDAQ: LRCX: Designs and manufactures wafer fabrication equipment for the semiconductor industry.
12+ YOE6+ MgmtSenior SRE leader with 12+ years infrastructure/SRE/DevOps experience and 6+ years leading teams; deep multi-cloud networking, observability, DR/BCP, backup/restore, automation (Ansible/Terraform/Python), Kubernetes, and incident management experience.
Prometheus, Grafana, Datadog, PagerDuty, ThousandEyes, Azure Monitor, CloudWatch, Google Cloud Operations, Splunk, Ansible, Terraform, Python, Kubernetes, AKS, EKS, GKE, ServiceNow
1mo
Save
Mark Applied
Hide
Senior Manager, Networ Reliability Engineering
Santa Clara or Seattle or United States
$133k-$306k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
5+ YOE3+ Mgmt5+ years network reliability engineering, 3+ years engineering/operations management, strong cloud networking and distributed systems expertise, proven people leadership, excellent communication and organizational skills.
OCI

Explore Jobs

Expand Your Job Search