Plaud: Develops AI-powered voice recorders and automated transcription software.
5+ YOE5+ years SRE/Infra experience, strong cloud (AWS/GCP/Azure/OCI), Kubernetes, on-call incident management, and proficiency in Go, Python, or Java.
Pune or India or Milpitas or Seattle or Princeton or Cape Town or London or Zurich or Singapore or Mexico City
OnsiteFull Time
ZensarNational Stock Exchange of India: ZENSARTECH: Global technology firm providing digital transformation and infrastructure services.
15+ YOE15+ years IT experience in product owner/manager or technical leadership; strong knowledge of IT operations, monitoring/observability, ITSM (ServiceNow/Remedy), automation (StackStorm), cloud/hybrid, DevOps/SRE, Agile/SAFe, and AIOps initiatives.
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bengaluru or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Washington or Dallas
$200k-$275k/yrOnsiteFull Time
SingleStore: Provides a distributed SQL database for real-time analytics.
Product leadership experience with distributed databases, query processing, AI agents or MCP servers, observability systems, SRE workflows, complex technical initiatives, and cross-functional people leadership.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
7+ YOE7+ years product management experience including 3+ years on technical infrastructure or platform products; experience defining observability, SLIs/SLOs/SLA, partnering with SRE and fleet engineering; strong writing, influence, and execution skills.
Seattle or United States or Europe or Africa or Europe
$215k/yrOnsiteFull Time
Edge Delta: Software platform for processing and analyzing observability data.
Master's degree in software or computer engineering required. Leads product strategy and execution through engineering collaboration, enterprise customer engagement, and product performance analysis.
Senior Lead Software Engineer - Agentic AI SRE Platform
Seattle, Washington, United States
$171k-$260k/yrOnsiteFull Time
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years software engineering experience delivering production agentic AI and distributed systems; strong Python, cloud (AWS), Terraform, AI agent frameworks; SRE and reliability experience preferred.
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bangalore or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Dallas
$200k-$275k/yrOnsiteFull Time
SingleStore: Real-time distributed SQL database for transactions and analytics.
Requires deep distributed database, query processing, performance optimization, observability, and AI product expertise, plus multi-team leadership, customer engagement, execution, communication, and mentorship skills.
IPinfo: Provides IP geolocation and network intelligence data via API.
8+ YOE5+ Mgmt8+ years in infrastructure/SRE/platform roles with hands-on scaling, 5+ years managing infra teams, DDoS mitigation and incident response experience, deep cloud (AWS/GCP/Azure), networking, Terraform/Ansible, Linux.
Terraform, Ansible, AWS, GCP, Azure, BGP, DNS, CDN, Linux
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
7+ YOE2+ Mgmt7+ years in software or infrastructure engineering with 2+ years leadership; SRE fundamentals, incident management, observability, change management; strong automation and people development skills.
Infrastructure Engineer - IT Merchandising - On-Premise Systems (SRE)
Issaquah, Washington, United States
$85k-$225k/yrOnsiteFull Time
Costco WholesaleNasdaq: COST: Operates a global chain of membership-only big-box warehouse clubs.
3+ YOE3+ years managing Windows and Linux servers, experience with Puppet, JBoss, SQL/Oracle, scripting (PowerShell, Python, JavaScript), monitoring and enterprise operations.
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
5+ YOE3+ Mgmt5+ years SRE/DevOps experience,3+ years managing technical teams,BS/MS or equivalent,expertise with Kubernetes,Terraform,cloud providers,CI/CD and incident response,programming and scripting skills.
Major Incident Manager, Incident Management -TikTok USDS
Seattle or Bellevue
$130k-$246k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
4+ YOE4+ years incident management/production support/SRE experience in large-scale SaaS or cloud environments, cloud and monitoring proficiency, scripting/automation skills, ITIL knowledge, strong communication and incident leadership.
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
7+ YOE2+ Mgmt7+ years related experience, Bachelor's degree (CS/Engineering/IT) plus leadership experience; expertise in AWS/Azure, CI/CD, GitLab, IaC (Terraform/CloudFormation/Bicep), SRE practices, cloud security, and observability.
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOE5+ years SRE/DevOps experience with Kubernetes and Linux, proficiency in Bash/Python, experience managing infrastructure and automation, and ability to obtain Top Secret/SCI or DOE Q clearance.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
10+ YOE3+ Mgmt10+ years in Production Engineering/SRE/Infrastructure; 3-5 years in people management; strong Kubernetes, Terraform, PostgreSQL experience; AI/GPU workloads; multi-region/multi-cloud scaling; incident leadership.
Kubernetes, Terraform, PostgreSQL, AI workflows, GPU workloads