Pune or India or Milpitas or Seattle or Princeton or Cape Town or London or Zurich or Singapore or Mexico City
OnsiteFull Time
ZensarNational Stock Exchange of India: ZENSARTECH: Global technology firm providing digital transformation and infrastructure services.
15+ YOE15+ years IT experience in product owner/manager or technical leadership; strong knowledge of IT operations, monitoring/observability, ITSM (ServiceNow/Remedy), automation (StackStorm), cloud/hybrid, DevOps/SRE, Agile/SAFe, and AIOps initiatives.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
8+ YOE8+ years in incident management/SRE/infrastructure operations; experience with large-scale distributed infrastructure, data center operations, GPU clusters, networking, cloud platforms; incident frameworks (ITIL/SRE); strong leadership, communication, and stakeholder management.
PagerDuty, ServiceNow, Jira, Datadog, Prometheus, Grafana, Incident command system (ICS)
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
Experience with system design, scalable cloud infrastructure, configuration management, programming in Python or Go, distributed systems, automation, documentation, and cross-functional collaboration.
AWS, GCP, Azure, Chef, Ansible, Terraform, GitHub Actions, Python, Go
Kody: An agentic commerce platform providing integrated in-person payment solutions.
Deep AWS and GitHub experience, strong monitoring/logging and scripting skills, incident management ownership, and absolute fluency in Mandarin and English.
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
5+ YOE3+ Mgmt5+ years SRE/DevOps experience,3+ years managing technical teams,BS/MS or equivalent,expertise with Kubernetes,Terraform,cloud providers,CI/CD and incident response,programming and scripting skills.
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
12+ YOE4+ MgmtBachelor's degree plus 12 years' relevant experience, 4+ years managing technical teams, hands-on security/software engineering and DevSecOps experience, ability to translate security/compliance into automated engineering workflows and partner across engineering, security, compliance, and SRE.
Bloom EnergyNYSE: BE: Manufactures solid oxide fuel cell systems for onsite power.
10+ YOE3+ MgmtBachelor’s degree and 10+ years in cloud, infrastructure, platform engineering, or DevOps, including 3+ years in senior technical leadership. Requires AWS, networking, IaC, CI/CD, Kubernetes, observability, security, and SRE expertise.
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
15+ YOE5+ Mgmt15+ years leading IT infrastructure and operations with 5+ years in senior management; experience managing large global teams, IAM, SRE/automation, AI/ML in support environments, XaaS and on‑prem architectures, and compliance frameworks (NIST, SOX, SOC).
ByteDance: Developing AI-driven content platforms and mobile applications.
Bachelor's in related field and strong experience with large-scale Linux host management, core data-center services (DNS, NTP, DHCP, NAT, APT, Kerberos), DevOps tooling, SRE practices, and troubleshooting.
Slalom: Provides business and technology consulting and software engineering services.
Senior leader with deep technology delivery and leadership experience in product engineering, cloud modernization, AI-accelerated engineering, SRE/operations, and go-to-market capability development.
Tech Lead Cloud Site Reliability Engineer - DCS Cloud
San Jose, California, United States
OnsiteFull Time
ByteDance: Developing AI-driven content platforms and mobile applications.
5+ YOE5+ years SRE/DevOps/Linux operations experience, bachelor’s in CS or related field, proficiency in Go/Python/C++, strong troubleshooting, monitoring, and reliability practices.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
10+ YOE10+ years in software/platform engineering or SRE; 5+ years Kubernetes at scale; strong Go and Python; deep Kubernetes internals; GPU orchestration; multi-tenant infrastructure; distributed systems; observability at scale; Linux networking; IaC and GitOps.
Tech Lead Site Reliability Engineer, TikTok Generalized Arch USTO
San Jose, California, United States
$245k-$450k/yrOnsiteFull Time
TikTok: Global short-form video hosting and social media platform.
5+ YOEBachelor's in CS or related, strong CS foundation, Linux and storage/network knowledge, proficiency in Python/Go/Java/PHP/C/C++, strong problem solving and communication; 5+ years SRE/cloud experience preferred.
Director, Enterprise IT Infrastructure (CA, US, 95110)
San Jose, California, United States
$191k-$280k/yrOnsiteFull Time
QuantumScapeNYSE: QS: Develops next-generation solid-state batteries for electric vehicles.
15+ YOE5+ Mgmt15+ years IT experience with 5+ years leading enterprise infrastructure; Bachelor's in CS/IT/Engineering; hands-on expertise with GCP/Azure, Kubernetes (GKE), networking, identity, endpoints, SRE, automation, and OT/IT integration; strong leadership and cross-functional communication.
San Francisco or Oakland or San Jose or California
$240k-$290k/yrOnsiteFull Time
Slalom: A global business and technology consulting firm.
Senior go-to-market leader with deep product engineering, cloud modernization, AI-accelerated engineering, and large-scale delivery experience; ability to lead teams, drive revenue, support billable delivery, and engage clients and partners (AWS, Microsoft, Google). Residency in San Francisco/East Bay/Silicon Valley required.
California or United States or San Jose or Palo Alto
$191k-$280k/yrOnsiteFull Time
QuantumScape: Develops solid-state lithium-metal batteries for electric vehicles.
15+ YOE5+ MgmtBachelor's in computer science, IT, engineering, or related field; 15+ years of IT experience; 5+ years leading enterprise infrastructure or IT operations teams; expertise in networking, identity, endpoints, hybrid infrastructure, service delivery, and people leadership.