80 site operations engineer jobs at 59 companies in Menlo Park, CA
2w
Save
Mark Applied
Hide
2w
Staff Site Reliability Engineer-Production Operations
Palo Alto, California, United States
$186k-$233k/yrHybridFull Time
Rivian and Volkswagen Group Technologies: A joint venture creating software-defined vehicle technology and connected services for electric vehicles.
Senior SRE with incident command experience, strong systems engineering for distributed systems, hands-on coding (Python or Go), observability expertise (Datadog or comparable), and experience with blameless post-incident practices.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
5+ YOE5+ years systems and software engineering experience for large-scale internet services; expertise in SRE principles, containers, observability, incident management, Python and Go, and applying AI/ML to operations.
EarnIn: Provides immediate access to earned wages through a mobile app.
7+ YOE7+ years in SRE or related field; experience applying AI/LLMs to operations; strong SLO/SLI and incident management; software engineering in Python or Go; observability and IaC proficiency; AI-assisted development tools; fintech/regulated environment experience.
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
Proven experience in site reliability, DevOps, or infrastructure operations; proficiency with Python, Go, or Bash; knowledge of distributed systems, monitoring/observability, incident management, and automation; experience supporting production systems and on-call rotations.
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
2+ YOE1+ MgmtBachelor's degree in a technical field; 2+ years lab engineering/technical experience; strong lab equipment knowledge; junior management experience; ability to travel within the Bay Area.
Retool: Software platform for building custom internal business applications.
Experience operating production infrastructure (AWS), Kubernetes, Terraform, Postgres; programming in Go/Python/TypeScript/Java/Ruby; building observability and automation for customer-facing SaaS systems.
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
2+ YOEBachelor's in a technical field or equivalent experience, 2+ years experience, lab equipment and engineering background, junior management experience, ability to assemble and move equipment, debug PCBs, run tests on Windows/Linux, and occasionally travel within the Bay Area.
Microsoft Visio, Microsoft SharePoint, Jira, Windows, Linux, oscilloscopes
CircleNYSE: CRCL: Digital currency issuer and blockchain financial infrastructure provider.
6+ YOE6+ years SRE/DevOps experience preferred; strong Kubernetes, IaC (Terraform/Pulumi), cloud networking, Go or Python, CI/CD, observability, and distributed systems experience; blockchain node operation experience is a plus.
Senior Site Reliability Engineer Platform Private Cloud Engineer
San Jose, California, United States
$94k-$130k/yrOnsiteFull Time
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
7+ YOERequires 7+ years designing and operating enterprise or cloud environments, private cloud and Kubernetes expertise, scripting, IaC tools, Unix/Linux knowledge, and a CS or engineering degree.
Nectar Social: AI platform for social commerce and community management.
5+ YOE5+ years operating production systems; cloud (AWS); infrastructure as code; programming; startup environment; reliability-focused with cost awareness.
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yrHybridFull Time
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
8+ YOE8+ years operations experience in cloud and Linux, strong networking and web server knowledge, proficiency with Terraform/Chef and scripting (Bash, Python, Go), experience with automation tools and on-call duty.
TikTok: Global short-form video hosting and social media platform.
1+ YOEMaster's degree and 1 year of related experience, or bachelor's degree and 3 years; requires 1 year supporting critical systems, monitoring, troubleshooting, data operations, error analysis, and runbook creation.
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOEBachelor's degree in CS or related,2+ years in Linux operations/SRE/DevOps,programming in Go/Python/C++,cloud and reliability practices experience,strong troubleshooting and communication skills.
Member of Technical Staff, Site Reliablity Engineer
San Francisco, California, United States
$200k-$270k/yrHybridFull Time
Vapi: Build and deploy AI-powered voice agents via flexible APIs.
Experience running incident command and postmortems, operating SLOs/error budgets, capacity planning and load testing, Kubernetes production ops, KEDA autoscaling, and shipping services in Go or TypeScript.
CBRENYSE: CBRE: Provides global commercial real estate services and investment management.
5+ YOEBachelor's degree preferred or equivalent experience; 5+ years in engineering operations, facilities management, or building systems; leadership, budgeting, reporting, vendor management, and Microsoft Office expertise required.
Microsoft Office, Microsoft Excel, Microsoft Word, Microsoft Outlook, Microsoft PowerPoint
Site Reliability Engineer - USDS (Multiple Positions)
San Jose, California, United States
$188k-$259k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
1+ YOERequires degree in CS/Engineering/Information Systems/Data Science/Mathematics plus relevant experience; experience with Linux administration, monitoring, troubleshooting, SDLC and cloud-native operations.
San Francisco or New York or Denver or Austin or Calgary or Toronto
$118k-$135k/yrHybridFull Time
IntersectNASDAQ: GOOGL: Develops-located data centers and renewable energy infrastructure.
1+ YOEB.S. in Electrical Engineering; 1–3 years field experience; interpret electrical diagrams; analyze data from SCADA/CMMS; strong communication; safety-focused.
San Francisco or San Jose or New York City or Milpitas or Mountain View or Holmdel or Goleta or Redwood City or Fremont or Sunnyvale or Brooklyn or Palo Alto
$168k-$245k/yrHybridFull Time
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
4+ YOERequires 7+ years with a bachelor's, 4+ with a master's, or 1 with a PhD; 4+ years in SRE or related engineering, 3+ years operating Kubernetes, Helm, CI/CD, cloud platforms, and Python or Go.