386 network reliability engineer jobs at 225 companies in United States

2mo
Save
Mark Applied
Hide
Network Reliability Engineer
Austin or Atlanta or Denver or Seattle or Washington
HybridFull Time
Cloudflare
CloudflareNYSE: NET: Cloud-based security, performance, and reliability services for Internet applications.
3+ YOE3+ years network/site reliability engineering experience; BA/BS in CS or equivalent; experience with Saltstack, Ansible, Chef, NX-OS/JUNOS/EOS/Cumulus/Sonic; Linux administration; iproute2, Traffic Control, Devlink; software development in Go and Python; AI/LLM tooling experience.
Saltstack, Ansible, Chef, NX-OS, JUNOS, EOS, Cumulus, Sonic, LLM, iproute2, Traffic Control, Devlink, Go, Python, AirFlow, Temporal, FRR, Bird, GoBGP, C, C++, rust, Linux, Linux kernel, Prometheus, Grafana, Thanos, Clickhouse, Kubernetes, Docker, Consul
1w
Save
Mark Applied
Hide
Network Reliability Engineer, Infrastructure Services
California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designing and manufacturing consumer electronics, software, and digital services.
Extensive software, systems, or infrastructure engineering experience; expertise designing highly available distributed systems, systems programming, networking, SDN, and cross-functional technical leadership.
JSON, Protocol Buffers (ProtoBuf), REST, RPC, XML, Kubernetes (K8s), OpenStack, Infrastructure as Code
2mo
Save
Mark Applied
Hide
Staff Network Site Reliability Engineer
Amsterdam or United Kingdom or Israel or United States or Europe or North America
$180k-$224k/yr RemoteFull Time
Nebius Group
Nebius GroupNASDAQ: NBIS: Building a full-stack AI cloud infrastructure platform.
Strong Linux and networking fundamentals, SRE/incident response experience, ability to write automation (Go/Python), familiarity with IaC, CI/CD, container platforms, observability and reliability engineering practices.
Infrastructure as Code (IaC), CI/CD, Go, Python, eBPF, XDP, DPDK, perf, ftrace
2mo
Save
Mark Applied
Hide
Senior Network & Site Reliability Engineer
San Francisco, California, United States
$210k-$240k/yr OnsiteFull Time
Alembic Technologies
Alembic Technologies: San Francis private Causal AI platform helping enterprise marketers measure which actions drive growth and business outcomes.
8+ YOE8+ years in network or infrastructure engineering (5+ years datacenter ops); strong network security and architecture skills; hands-on with BGP, QoS, MPLS, IPsec, EVPN/VXLAN, ECMP; IaC (Ansible, Terraform, Nornir); NetBox/Infoblox; Kubernetes networking; Linux; monitoring stacks; Python/Bash.
NVIDIA DGX SuperPOD, Grace Blackwell, BGP, VPNs, WAN, QoS, MPLS, IPsec, EVPN, VXLAN, ECMP, Ansible, Terraform, Nornir, NetBox, Infoblox, Kubernetes, Prometheus, Grafana, Datadog, ELK, OpenTelemetry, Python, Bash, Cumulus Linux, InfiniBand, Spectrum-X, BlueField, Spark, Airflow, Kafka, NFS, LustreFS, iSCSI, Linux
1w
Save
Mark Applied
Hide
Senior Network Reliability Engineer - DGX Cloud
Santa Clara or United States
$136k-$265k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOE5+ years in network operations; deep TCP/IP and routing knowledge; troubleshooting, incident management, cloud networking, automation, and excellent communication skills; bachelor's degree or equivalent experience.
TCP/IP, BGP, OSPF, MPLS, IS-IS, VxLAN, EVPN, QoS, GRE, IPsec, DNS, MACsec, AWS, Azure, GCP, OCI, Arista, Fortinet, Juniper, Mellanox, Cumulus OS, Infiniband, Unix, Linux, Python, Shell, Netbox, Nautobot, Prometheus, Grafana, Panoptes
2w
Save
Mark Applied
Hide
10351 - Network Reliability Engineer
Irvine, California, United States
$115k-$125k/yr OnsiteFull Time
Hyundai AutoEver America
Hyundai AutoEver AmericaKorea Exchange (KOSPI): 307950: Public South Korean mobility-software provider serving Hyundai Motor Group with vehicle software, maps, and OTA services.
Requires enterprise network engineering, Cisco routing and switching, Palo Alto firewalls, Cisco ISE, TCP/IP, BGP, OSPF, LAN/WAN, monitoring, telemetry, Linux, Python, Ansible, and automation experience.
Cisco, Cisco ISE, Palo Alto, Splunk, SolarWinds, Grafana, Prometheus, SNMP, syslog, NetFlow/IPFIX, Python, Ansible, REST APIs, Git, Linux, AWS, Azure, Google Cloud, OpenStack, Terraform, OpenTelemetry, ITSM
2w
Save
Mark Applied
Hide
Sr Network Reliability Engineer
Los Angeles or Florida or Denver or Reston or El Segundo
$198k-$277k/yr OnsiteFull Time
Blue Origin
Blue Origin: Developing reusable space vehicles and infrastructure.
8+ YOEBachelor’s degree or equivalent experience and 8+ years designing scalable network infrastructure. Requires networking, automation, security, multi-vendor configuration, communication skills, and U.S. work authorization.
TCP/IP, IPv4, IPv6, MPLS, BGP, OSPF, IS-IS, DHCP, DNS, NTP, Python, Go, Rust, C++, Juniper, JNCIS-ENT, JNCIS-DevOps, JNCIP, Palo Alto Networks, NIST SP 800-series, NIST SP 800-57, FIPS 140-2, FIPS 140-3, VPN, VLAN
1mo
Save
Mark Applied
Hide
Senior Network Reliability Engineer, Incident Management
Mountain View or Espoo or Bengaluru
HybridFull Time
Skylo Technologies
Skylo Technologies: Private telecommunications providing satellite connectivity for smartphones, vehicles, and IoT devices where cellular networks are unavailable.
5+ YOE5+ years telecom or network operations experience, incident/outage management expertise, observability and Kubernetes literacy, ticketing and on-call tool proficiency, strong written and verbal communication.
Grafana, Prometheus, Loki, kubectl, Jira, ServiceNow, PagerDuty, Python, Bash
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (Network)
San Francisco or Golden
$157k-$239k/yr OnsiteFull Time
Loft Orbital
Loft Orbital: Space infrastructure helping governments, companies, and research institutions deploy and operate missions in low Earth orbit.
4+ YOE4–5 years network engineering experience, hands-on SDN and public-cloud networking (ideally GCP), Kubernetes/Docker familiarity, IaC (Terraform) and GitOps, SRE mindset (SLOs, observability), degree or equivalent experience.
GCP, k8s, Docker, Terraform, Grafana, ArgoCD, FluxCD, Cockpit, GitOps
2mo
Save
Mark Applied
Hide
Senior Network Reliability Engineer - DGX Cloud
Santa Clara or United States
$136k-$265k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOE5+ years network operations experience; deep knowledge of TCP/IP, BGP, OSPF, MPLS, EVPN/VxLAN and related protocols; experience in CSPs (AWS, Microsoft Azure, GCP, OCI); scripting/automation and vendor familiarity (Arista, Juniper, Fortinet).
TCP/IP, BGP, OSPF, MPLS, IS-IS, VxLAN, EVPN, QoS, GRE, IPsec, DNS, MACsec, AWS, Microsoft Azure, GCP, OCI, Arista, Fortinet, Juniper, Mellanox, Cumulus OS, NetBox, Nautobot, Prometheus, Grafana, Panoptes, Python, Shell
2w
Save
Mark Applied
Hide
Sr Lead Network Reliability Engineer
United States
$149k-$208k/yr RemoteFull Time
Coupa Software
Coupa Software: AI-driven business spend management platform for global enterprises.
10+ YOEComputer Science degree or equivalent and 10+ years' experience; expertise in cloud networking, firewalls, Kubernetes, DNS, BGP, Linux, infrastructure as code, troubleshooting, and programming.
AWS, Azure, GCP, Palo Alto Network Firewall, Fortinet, CheckPoint, Kubernetes, TCP/IP, DNS, BGP, TLS, Golang, Python, Java, Ruby, Terraform, CloudFormation, Chef, Ansible, Spacelift, Linux, VPN, VPC, CNI, NetworkPolicy, Microsoft?, AI
1w
Save
Mark Applied
Hide
Site Reliability Engineer Intern (Data Infra) - 2027 Fall
San Jose, California, United States
OnsiteInternship
ByteDance
ByteDance: Global technology specializing in AI-powered content platforms.
Currently pursuing a bachelor's degree in computer science or related technical discipline; programming experience in C, C++, Java, Python, Go, or Rust; knowledge of Unix/Linux internals, networking, and distributed systems.
C, C++, Java, Python, Go, Rust, Unix, Linux, Kubernetes, Redis, MySQL, Flink, Nginx, Docker, OpenStack, Hadoop, Spark
2w
Save
Mark Applied
Hide
Senior Network Reliability Engineer
Zionsville, Indiana, United States
$135k-$190k/yr RemoteFull Time
Group 1001
Group 1001: Private insurance and financial-services collective providing life insurance, annuities, online investing, and technology services to individuals.
Deep TCP/IP, BGP, OSPF, VPN, and SD-WAN knowledge; production Terraform, Ansible, and Python experience; security-platform experience; and monitoring and observability expertise.
Terraform, Pulumi, Ansible, Python, Cloudflare, Zscaler ZIA/ZPA, ZTNA, Transit Gateway, Cloud WAN, VPCs, Route53, IPAM, Kubernetes, Istio, Linkerd, Consul Connect, Cilium, Hubble, OPA/Rego, Sentinel, Grafana, Elastic, Open Telemetry, Datadog, Prometheus, Pixie, AWS, AFT, Control Tower, BGP, OSPF, VPN, SD-WAN, TCP/IP
1w
Save
Mark Applied
Hide
Plant Engineer and Network Reliability Manager
East Bernstadt, Kentucky, United States
OnsiteFull Time
Sazerac Company
Sazerac Company: Global producer and marketer of award-winning spirits.
10+ YOEBachelor’s degree in engineering or related discipline and 10+ years of relevant engineering, reliability, maintenance, manufacturing, or operations experience. Requires project management, cross-functional leadership, budget, timeline, and AutoCAD skills.
AutoCAD
3mo
Save
Mark Applied
Hide
Senior Network Reliability Engineer - DGX Cloud
Santa Clara or United States
$136k-$265k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOE5+ years in network operations; strong TCP/IP, BGP, OSPF, MPLS, EVPN; experience with AWS/Azure/GCP; hands-on with automation; Bachelor’s in CS or related field.
TCP/IP, BGP, OSPF, MPLS, EVPN, VxLAN, GRE, IPsec, DNS, MACsec, Arista, Fortinet, Juniper, Mellanox, Cumulus OS, Infiniband, Netbox, Nautobot, Prometheus, Grafana, Python, Shell, AWS, Azure, GCP, OCI
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: Private supply-chain AI software serving corporations, government agencies, and banks with risk and compliance technology.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
2mo
Save
Mark Applied
Hide
ENGINEER AUTOMATION RELIABILITY (Miami, FL, US, 33182)
Miami, Florida, United States
FieldFull Time
Cemex
CemexNYSE: CX: Global building materials providing construction solutions.
12+ YOEBachelors in electrical/electronic/computer/networking engineering and minimum 12 years of industrial reliability/maintenance/automation/networking/cybersecurity experience; strong electrical, PLC, HMI, CMMS, and reliability methodology skills.
CMMS, SAP, AutoCAD, Microsoft PowerPoint, Microsoft Visio, Microsoft Excel, Microsoft Word, Rockwell PLCs, Modicon PLCs, HMIs, Process Control Network, CEMS/DAHS, PIMs, Networking, Cybersecurity, Six Sigma
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States
$150k-$200k/yr RemoteFull Time
Runpod
Runpod: AI cloud computing platform providing on-demand GPUs and serverless compute to developers, researchers, and AI companies.
5+ YOE5+ years SRE or production engineering experience; strong Linux, networking, container, distributed systems, SLI/SLO, incident response, and scripting skills.
Prometheus, Grafana, Python, Go, Bash, Linux, Slack
3mo
Save
Mark Applied
Hide
Senior Systems Reliability Engineer
California or Oregon
$150k-$225k/yr RemoteFull Time
IEX Group
IEX Group: Private U.S. exchange operator and financial technology serving broker-dealers, investors, and capital-markets participants.
Automation experience with Ansible or similar tools; Linux, Python, Bash, Git; experience with large distributed systems; networking and data center knowledge; able to troubleshoot across hardware, software, and network.
Ansible, Linux, Python, Bash, Git, Arista, Cisco, Corvil, Solarflare, Mellanox, TCP/IP, Networking, Packet Analysis
3mo
Save
Mark Applied
Hide
Sr. Site Reliability Engineer- Product Reliability Engineering
Austin, Texas, United States
$111k-$172k/yr HybridFull Time
Visa
VisaNYSE: V: Global leader in digital payments and transaction technology.
2+ YOE2+ years experience with Bachelor’s or 5+ years experience; programming in Python/Java/.NET/C#/PowerShell/Bash; Linux/Unix; networking; automation; AI frameworks.
Python, Java, .NET, C#, PowerShell, Bash, Linux, Networking, AI frameworks