43 network reliability engineer jobs at 18 companies in Soquel, CA

1w
Save
Mark Applied
Hide
Senior Network Reliability Engineer, Incident Management
Mountain View or Espoo or Bengaluru
HybridFull Time
Skylo
Skylo: Provides direct-to-device satellite connectivity for mobile and IoT devices.
5+ YOE5+ years telecom or network operations experience, incident/outage management expertise, observability and Kubernetes literacy, ticketing and on-call tool proficiency, strong written and verbal communication.
Grafana, Prometheus, Loki, kubectl, Jira, ServiceNow, PagerDuty, Python, Bash
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineering - Network
Palo Alto or Columbus
$152k-$215k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE10+ MgmtFormal network engineering training, 5+ years applied experience, 10+ years leading technologists, advanced network reliability skills, SD-WAN and cloud (AWS, Azure) proficiency, major network vendor experience, observability tooling and incident leadership.
SD-WAN, AWS, Azure, Palo Alto, Juniper, F5, Broadcom, Arista, Cisco, Grafana, SevOne, Prometheus, Kibana, ThousandEyes, Splunk, Jenkins, GitLab, Terraform, eBPF, TCP/IP, HTTPS, BGP
2mo
Save
Mark Applied
Hide
Senior Network Reliability Engineer - DGX Cloud
Santa Clara or United States
$136k-$265k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOE5+ years in network operations; strong TCP/IP, BGP, OSPF, MPLS, EVPN; experience with AWS/Azure/GCP; hands-on with automation; Bachelor’s in CS or related field.
TCP/IP, BGP, OSPF, MPLS, EVPN, VxLAN, GRE, IPsec, DNS, MACsec, Arista, Fortinet, Juniper, Mellanox, Cumulus OS, Infiniband, Netbox, Nautobot, Prometheus, Grafana, Python, Shell, AWS, Azure, GCP, OCI
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Palo Alto or Pittsburgh
$179k-$269k/yr OnsiteFull Time
Latitude AI
Latitude AI: Developing automated driving technology for next-generation Ford vehicles.
4+ YOEBachelor's degree in engineering/computer science (or higher) with 4+ years experience (or equivalent), strong Linux, networking, Go/Python development, cloud (AWS/GCP), Kubernetes, IaC, monitoring and SLO experience.
Go, Python, AWS, GCP, Terraform, CloudFormation, Kubernetes, Prometheus, Elasticsearch, Loki, Jaeger, Tempo, Linux
2w
Save
Mark Applied
Hide
Site Reliability Engineer
Santa Clara, California, United States
$230k-$250k/yr OnsiteFull Time
Forward Networks
Forward Networks: Provides a digital twin platform for enterprise network management.
6+ YOE6+ years SRE/DevOps experience in SaaS/cloud, strong networking fundamentals, Kubernetes, observability (Prometheus/Grafana/Datadog/Splunk), Python/Bash automation, cloud and IaC (AWS/GCP/Azure, Terraform/Ansible), and incident response ownership.
Kubernetes, Prometheus, Grafana, Datadog, Splunk, Python, Bash, AWS, GCP, Azure, Terraform, Ansible
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer - SDN
San Francisco or San Jose or Bellevue
$240k-$312k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
Kubernetes, SmartNICs, Python, Ansible, Go, C, Helm, Terraform, GitOps, CI/CD, Linux, OpenStack Neutron, OVN, OVS, DPDK, SR-IOV
3w
Save
Mark Applied
Hide
Senior Systems Reliability Engineer
Pune or San Jose or Durham or Mexico City or Bangalore or Hoofddorp or Belgrade or Barcelona or Singapore or Sydney or Tokyo
HybridFull Time
Nutanix
NutanixNASDAQ: NTNX: Sells cloud software and hyperconverged infrastructure for enterprises.
7+ YOE7+ years SRE experience with networking, virtualization (VMware ESXi), Linux, cloud and strong customer-facing troubleshooting and communication skills.
VMware ESXi, VMware, Linux, DevOps, Cloud, Citrix, Microsoft
1mo
Save
Mark Applied
Hide
Senior Manager, Networ Reliability Engineering
Santa Clara or Seattle or United States
$133k-$306k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
5+ YOE3+ Mgmt5+ years network reliability engineering, 3+ years engineering/operations management, strong cloud networking and distributed systems expertise, proven people leadership, excellent communication and organizational skills.
OCI
3w
Save
Mark Applied
Hide
Site Reliability Engineer, Compute Platform
San Jose, California, United States
$156k-$388k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Bachelor's in CS/Engineering, strong Linux, networking, databases, Kubernetes, SRE/DevOps toolset knowledge, experience with ClickHouse/Spark/Presto/Doris/Hadoop, coding in Python/Shell/Java/Go, strong problem-solving and communication.
ClickHouse, Spark, Presto, Doris, Hadoop, Kubernetes, Python, Shell, Java, Go
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Redwood City, California, United States
HybridFull Time
Luma AI
Luma AI: Develops multimodal AI for video generation and creative production.
5+ YOE5+ years SRE or infrastructure experience, deep Linux and low-level performance debugging, Terraform, Airflow, Ray, AWS or OCI, high-performance networking (InfiniBand/RDMA/RoCE), security/compliance familiarity.
Linux, Python, Go, Bash, Terraform, Airflow, Ray, AWS, OCI, InfiniBand, RDMA, RoCE, DCGM, ROCm, Kubernetes, NVIDIA, AMD
2mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Santa Clara, California, United States
$152k-$245k/yr OnsiteFull Time
Palo Alto Networks
Palo Alto NetworksNASDAQ: PANW: Provides enterprise-grade network, cloud, and endpoint security software.
BS or MS in CS or related field; expertise in configuration management (Ansible, Terraform, Kubernetes); Python and/or Go; Kubernetes with autoscaling; production engineering/DevOps/SRE experience; public cloud (GCP/AWS); Linux networking; CI/CD with GitLab/GitHub; distributed systems; strong communication; ownership and monitoring as code.
Kubernetes, Docker, GCP, AWS, Ansible, Terraform, Vault, GitLab, Spinnaker, Pub/Sub, Bigtable, Memorystore, BigQuery, RabbitMQ, Kafka, MySQL, Python, Go, Shell scripting, Golang
3w
Save
Mark Applied
Hide
Site Reliability Engineer, Compute Platform
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Experience with Linux, networking, databases, Kubernetes, ClickHouse/Hadoop/Doris/Spark/Presto, scripting or programming (Python, Shell, Java, Go), incident management, and capacity planning.
ClickHouse, Spark, Presto, Doris, Hadoop, Kubernetes, Linux, Python, Shell, Java, Go
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
San Jose, California, United States
$119k-$170k/yr HybridFull Time
Zscaler
ZscalerNASDAQ: ZS: Provides cloud-native cybersecurity solutions through zero trust architecture.
5+ YOE5+ years Linux/UNIX admin, Kubernetes/Docker, automation (Ansible), network/security fundamentals, and strong security practices.
Docker, Kubernetes, Ansible, Python, Golang, BASH, Openstack, CEPH, HashiCorp Vault, nftables
1mo
Save
Mark Applied
Hide
Distinguished Technologist Mechanical Engineer (Network Infrastructure)
Sunnyvale, California, United States
$163k-$348k/yr OnsiteFull Time
Hewlett Packard Enterprise
Hewlett Packard EnterpriseNYSE: HPE: Provides edge-to-cloud IT infrastructure and platform services.
15+ YOEBS in Mechanical Engineering, 15+ years electromechanical product development experience; expertise in chassis, packaging, manufacturability, reliability, SolidWorks/PLM, FEA, GD&T, thermal and cooling technologies; strong leadership and troubleshooting skills.
SolidWorks, EPDM, FEA, ECAD, IDF
1mo
Save
Mark Applied
Hide
Distinguished Technologist Mechanical Engineer (Network Infrastructure)
Sunnyvale, California, United States
$163k-$348k/yr OnsiteFull Time
Hewlett Packard Enterprise
Hewlett Packard EnterpriseNYSE: HPE: Providing global edge-to-cloud infrastructure and IT solutions for businesses.
15+ YOEBS in Mechanical Engineering, 15+ years electromechanical product development experience; expertise in chassis, packaging, manufacturability, thermal and reliability; SolidWorks and PLM experience; tolerance/GD&T, FEA, ECAD-MCAD, and supplier engagement experience.
SolidWorks, EPDM, PLM, FEA, ECAD-MCAD, IDF, GD&T
1w
Save
Mark Applied
Hide
Cloud Engineer
Palo Alto, California, United States
$125k-$200k/yr OnsiteFull Time
Sage Care
Sage Care: AI-powered care navigation and scheduling platform for health systems.
4+ YOE4+ years DevOps/SRE experience with GCP, Kubernetes (GKE), Terraform, CI/CD, and Bazel; strong networking, IAM, cloud security, and production reliability skills.
GCP, Terraform, Kubernetes, GKE, Bazel, CI/CD
3mo
Save
Mark Applied
Hide
Lead Systems Engineer - Traffic Management
Durham or Miami or Palo Alto or Washington
$15k-$23k/yr HybridFull Time
Nubank
NubankNYSE: NU: Digital financial platform offering banking, credit, and investment services.
Experience operating large-scale cloud distributed systems, Kubernetes, service mesh (Istio/Linkerd/Envoy), AWS networking, IaC with Pulumi or Terraform, strong reliability/observability and AI-assisted workflows.
Kubernetes, Istio, Linkerd, Envoy, Finagle, AWS, ALB, NLB, VPC, Route 53, IAM, Pulumi, Terraform
4d
Save
Mark Applied
Hide
Staff Backend Software Engineer
Sunnyvale, California, United States
$210k-$314k/yr OnsiteFull Time
Carbon
Carbon: Manufacturer of industrial 3D printers and advanced polymer materials.
8+ YOE8+ years building and operating production backend services; deep Linux and networking fundamentals; experience with OS provisioning, reliability, and mentoring; bachelor's in CS/Engineering or equivalent.
Linux, systemd, Bazel, MQTT, Python, Go, Node, TypeScript, C++
1mo
Save
Mark Applied
Hide
Senior Kubernetes DevOps Engineer - K0rdent AI Apps/Core Services
San Jose, California, United States
HybridFull Time
Mirantis
Mirantis: Develops software for managing cloud and AI infrastructure.
Kubernetes operator expertise; experience with Linux, virtualization, networking, and storage; system architecture, scalability, and reliability; debugging production systems; microservices and enterprise IT experience; excellent English communication.
Kubernetes, k0rdent, k0s, k0smotron, Kubeflow, Kserve, vLLM, NVIDIA AI Enterprise, KubeVirt, Harbor, ArgoCD, Grafana, OpenStack, Linux