Kerrigan Robotics: AI-powered orchestration platform for factory automation systems.
5+ YOE5+ years SRE/DevOps experience; deep Kubernetes knowledge; expertise with Pulumi/Terraform, Helm; proficiency in Go, Linux networking, observability (Prometheus, Grafana), and CI/CD (GitHub Actions).
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
10+ YOE3+ Mgmt10+ years software engineering experience with SRE/platform focus, 3+ years managing SRE teams, proficiency in cloud (AWS/Azure/GCP), deep SRE principles, incident management, bachelor's in CS or equivalent, able to work 2+ days/week in Sunnyvale.
Pune or India or Milpitas or Seattle or Princeton or Cape Town or London or Zurich or Singapore or Mexico City
OnsiteFull Time
ZensarNational Stock Exchange of India: ZENSARTECH: Global technology firm providing digital transformation and infrastructure services.
15+ YOE15+ years IT experience in product owner/manager or technical leadership; strong knowledge of IT operations, monitoring/observability, ITSM (ServiceNow/Remedy), automation (StackStorm), cloud/hybrid, DevOps/SRE, Agile/SAFe, and AIOps initiatives.
8+ YOEBachelor's degree or equivalent,8 years software/systems engineering experience,5 years SRE experience,5 years software design experience,EMR not mentioned; strong troubleshooting and stakeholder management skills.
Site Reliability Engineering (SRE) Manager, Apple Maps
Cupertino, California, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Build, manage, and deliver highly available, automated infrastructure for Apple Maps at global scale; focus on reliability, scalability, and operational excellence.
WalmartNYSE: WMT: Multinational retail operating discount stores and supermarkets.
5+ YOEPlan and execute medium- to high-complexity technical programs, manage dependencies and risks, translate product needs into technical designs, apply SRE and cloud best practices, and align stakeholders across engineering and product teams.
6+ YOEBS/MS in Engineering or CS (or equivalent),6+ years product management experience,experience with cloud operations,SRE,distributed systems,release processes,and using AI to augment product work.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
8+ YOE8+ years in incident management/SRE/infrastructure operations; experience with large-scale distributed infrastructure, data center operations, GPU clusters, networking, cloud platforms; incident frameworks (ITIL/SRE); strong leadership, communication, and stakeholder management.
PagerDuty, ServiceNow, Jira, Datadog, Prometheus, Grafana, Incident command system (ICS)
Glean: AI platform for enterprise search and automated workplace agents
8+ YOE8+ years TPM/infrastructure or SRE experience with 3+ years leading infra/platform programs; BS/MS in CS/Engineering or related; strong cloud, ML/LLM, reliability, and cross-functional leadership skills.
Microsoft Teams, Zoom, ServiceNow, Zendesk, GitHub, AWS, GCP, Azure, LLM
Senior Technical Program Manager, DGX Cloud Software Products and Services
Santa Clara, California, United States
$168k-$322k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
8+ YOEMS EE or CS or equivalent; 8+ years in program management of large-scale software or infrastructure; strong analytical and cross-functional leadership; experience with cloud/infrastructure and SRE; proficient with Jira/Aha!/Confluence and Git.
Kody: An agentic commerce platform providing integrated in-person payment solutions.
Deep AWS and GitHub experience, strong monitoring/logging and scripting skills, incident management ownership, and absolute fluency in Mandarin and English.
Bellevue or Livingston or New York City or Sunnyvale or San Francisco
$182k-$242k/yrOnsiteFull Time
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
7+ YOE3+ Mgmt3+ years engineering management and 7+ years technical experience in cloud operations/SRE; strong knowledge of cloud platforms, K8S, observability, incident management, and systems programming (Go); experience defining SLAs/SLOs and running on-call.
Engineering Manager - Cloud Infrastructure and Devops
San Jose or Scottsdale or Chicago or Austin
$177k-$262k/yrHybridFull Time
PayPalNASDAQ: PYPL: Global digital payments platform for consumers and merchants.
7+ YOE2+ Mgmt7+ years software engineering experience, 2+ years engineering management leading teams of 4+, strong background in cloud infrastructure/DevOps/SRE, experience with AWS, Kubernetes, Terraform, and hiring/stakeholder management.
ThoughtSpot: AI-powered analytics platform for enterprise business intelligence.
5+ YOE5–8 years in TAM/SRE/technical support for enterprise SaaS; strong cloud (AWS/GCP/Azure), SQL, data warehousing, incident management, customer-facing communication and technical advisory skills.
GEICO: Provides vehicle and property insurance services to consumers.
6+ YOE5+ Mgmt8+ years with customer communications platforms, 6+ years coding, 5+ years managing technical teams; strong cloud, SRE, and software architecture skills; BS in IT or equivalent.
Amazon Connect, AWS, IAM, VPC, CloudWatch, Python, Java, Go, .NET, JavaScript, TypeScript, XML, JSON, REST, Lambda, API Gateway, S3, DynamoDB, RDS, SNS, SQS, EventBridge, Kinesis, Lex, Polly, AWS Management Console, Azure Portal, GCP, X-Ray, Jira, Azure DevOps, Microsoft Project, Microsoft Visio, Microsoft Excel, Microsoft PowerPoint, Microsoft Outlook, Zoom, Slack
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
5+ YOE3+ Mgmt5+ years SRE/DevOps experience,3+ years managing technical teams,BS/MS or equivalent,expertise with Kubernetes,Terraform,cloud providers,CI/CD and incident response,programming and scripting skills.