Forge GlobalNYSE: FRGE: Financial technology operating a private-market marketplace and data, custody, and investment solutions for companies and investors.
10+ YOE5+ MgmtRequires 5+ years leading SRE, DevOps, cloud operations, or reliability functions; 10+ years in engineering or operations; bachelor's degree or equivalent; cloud infrastructure, distributed systems, observability, CI/CD, and automation experience.
RedditNYSE: RDDT: Social news aggregation, web content rating, and discussion platform.
8+ YOE8+ years in site reliability or infrastructure engineering, distributed systems, cloud-native architecture, observability, automation, incident management, performance optimization, and backend software engineering.
OktaNASDAQ: OKTA: Identity management and access control software provider.
3+ Mgmt3+ years technical leadership experience; experience with cloud-native architectures, Kubernetes, Terraform, CI/CD, observability platforms; strong software development and automation background; US Person status required.
Amazon Web Services (AWS), Kubernetes, Terraform, Grafana, Splunk, APM, CI/CD
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
12+ YOE5+ MgmtBachelor's, master's, or doctoral degree in a related field or equivalent experience; 12+ years in SRE or IT service management and 5+ years leading global IT operations or service management teams.
Mountain View or Mountain View or California or United States
$222k-$301k/yrOnsiteFull Time
IntuitNASDAQ: INTU: A global financial technology platform powering prosperity.
8+ YOE3+ Mgmt8+ years in systems, SRE, or infrastructure engineering; 3+ years managing engineering teams; AWS at scale; distributed systems, Kubernetes, IaC, observability, incident management, and AI Ops experience; bachelor's degree required.
Runloop AI: Runloop AI provides AI infrastructure, secure code sandboxes, and evaluation tools for developers building software-engineering agents.
5+ YOE5+ years software engineering experience with 3+ years in SRE/DevOps, strong Python or Go skills, containerization, cloud infra, monitoring, networking, Linux administration, on‑call and incident management.
AbbottNYSE: ABT: Global healthcare technology focused on life-changing medical innovations.
Ensure reliability, scalability, and performance of a medical-device remote monitoring platform; expertise in cloud (Azure), Kubernetes, observability, automation, and incident management; bachelor's in a technical discipline.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK/EFK, Datadog, Linux
Site Reliability Engineering Manager, Vehicle Software
Sunnyvale, California, United States
$276k-$294k/yrHybridFull Time
Wayve: British autonomous-driving software licensing vehicle-agnostic AI Driver technology to automakers and fleet owners.
8+ YOE3+ MgmtRequires 8+ years building production software systems, 3+ years people leadership, SRE and reliability practices, architecture experience, and production coding in C++, Rust, Python, or Go.
Site Reliability Engineering Manager, Vehicle Software
Sunnyvale, California, United States
$276k-$294k/yrHybridFull Time
Wayve: British autonomous-driving software licensing vehicle-agnostic AI Driver technology to automakers and fleet owners.
8+ YOE3+ MgmtRequires 8+ years building production software systems, 3+ years of people leadership, SRE and reliability expertise, architecture experience, and hands-on coding in C++, Rust, Python, or Go.
Site Reliability Engineering (SRE) Manager, Apple Maps
Cupertino, California, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designing and manufacturing consumer electronics, software, and digital services.
Build, manage, and deliver highly available, automated infrastructure for Apple Maps at global scale; focus on reliability, scalability, and operational excellence.
Cerebras SystemsNasdaq Global Select Market: CBRS: Designs processors and systems for AI training and inference.
15+ YOE15+ years in SRE/infrastructure/platform engineering with large-scale fleets; experience in capacity management, orchestration, observability, SLOs/SLIs, incident response, and cross-team architecture.
AbbottNYSE: ABT: Global healthcare technology focused on life-changing medical innovations.
Senior SRE with strong distributed systems, cloud (Azure), Kubernetes, observability, automation, incident management, and cross-functional communication skills for a medical device remote monitoring platform.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK, EFK, Datadog, Linux
Scottsdale or San Francisco or Chicago or New York City or Phoenix or Chicago or San Francisco
$118k-$183k/yrHybridFull Time
Early Warning Services: U.S. bank-owned fintech and consumer reporting agency providing identity, fraud-risk, and real-time payment solutions to financial institutions.
5+ YOEBachelor's degree in business, computer science, or related field; 5+ years of technical experience; incident management, Linux, scripting, observability, Git, security protocols, and enterprise production experience.
Graphon AI: Graphon AI is a private enterprise-AI software building multimodal relational memory for organizations and AI agents.
Proficient in Bash and Python; experience with infrastructure-as-code, Docker, CI/CD, multi-cloud deployments, networking and identity access; comfortable managing production environments and using AI tools.
Bash, Python, Infrastructure-as-code, Docker, CI/CD, AI tools
BitdeerThe Nasdaq Stock Market LLC: BTDR: Public Singaporean Bitcoin mining and AI cloud infrastructure serving enterprises with computing, datacenters, and mining solutions.
5+ YOERequires 5+ years of Kubernetes operations, 2+ years managing GPU workloads, Terraform, Helm, GitOps, SRE practices, monitoring, Go or Python, and multi-tenant platform experience.