73 site operations engineer jobs at 55 companies in Fairview, CA
2w
Save
Mark Applied
Hide
2w
Staff Site Reliability Engineer-Production Operations
Palo Alto, California, United States
$186k-$233k/yrHybridFull Time
Rivian and Volkswagen Group Technologies: A joint venture creating software-defined vehicle technology and connected services for electric vehicles.
Senior SRE with incident command experience, strong systems engineering for distributed systems, hands-on coding (Python or Go), observability expertise (Datadog or comparable), and experience with blameless post-incident practices.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
5+ YOE5+ years systems and software engineering experience for large-scale internet services; expertise in SRE principles, containers, observability, incident management, Python and Go, and applying AI/ML to operations.
EarnIn: Provides immediate access to earned wages through a mobile app.
7+ YOE7+ years in SRE or related field; experience applying AI/LLMs to operations; strong SLO/SLI and incident management; software engineering in Python or Go; observability and IaC proficiency; AI-assisted development tools; fintech/regulated environment experience.
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
2+ YOE1+ MgmtBachelor's degree in a technical field; 2+ years lab engineering/technical experience; strong lab equipment knowledge; junior management experience; ability to travel within the Bay Area.
Retool: Software platform for building custom internal business applications.
Experience operating production infrastructure (AWS), Kubernetes, Terraform, Postgres; programming in Go/Python/TypeScript/Java/Ruby; building observability and automation for customer-facing SaaS systems.
CircleNYSE: CRCL: Digital currency issuer and blockchain financial infrastructure provider.
6+ YOE6+ years SRE/DevOps experience preferred; strong Kubernetes, IaC (Terraform/Pulumi), cloud networking, Go or Python, CI/CD, observability, and distributed systems experience; blockchain node operation experience is a plus.
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
2+ YOEBachelor's in a technical field or equivalent experience, 2+ years experience, lab equipment and engineering background, junior management experience, ability to assemble and move equipment, debug PCBs, run tests on Windows/Linux, and occasionally travel within the Bay Area.
Microsoft Visio, Microsoft SharePoint, Jira, Windows, Linux, oscilloscopes
Senior Site Reliability Engineer Platform Private Cloud Engineer
San Jose, California, United States
$94k-$130k/yrOnsiteFull Time
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
7+ YOERequires 7+ years designing and operating enterprise or cloud environments, private cloud and Kubernetes expertise, scripting, IaC tools, Unix/Linux knowledge, and a CS or engineering degree.
Nectar Social: AI platform for social commerce and community management.
5+ YOE5+ years operating production systems; cloud (AWS); infrastructure as code; programming; startup environment; reliability-focused with cost awareness.
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yrHybridFull Time
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
8+ YOE8+ years operations experience in cloud and Linux, strong networking and web server knowledge, proficiency with Terraform/Chef and scripting (Bash, Python, Go), experience with automation tools and on-call duty.
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOEBachelor's degree in CS or related,2+ years in Linux operations/SRE/DevOps,programming in Go/Python/C++,cloud and reliability practices experience,strong troubleshooting and communication skills.
Member of Technical Staff, Site Reliablity Engineer
San Francisco, California, United States
$200k-$270k/yrHybridFull Time
Vapi: Build and deploy AI-powered voice agents via flexible APIs.
Experience running incident command and postmortems, operating SLOs/error budgets, capacity planning and load testing, Kubernetes production ops, KEDA autoscaling, and shipping services in Go or TypeScript.
Site Reliability Engineer - Global SRE, Monetization Technology
San Jose, California, United States
$156k-$317k/yrOnsiteFull Time
TikTok: Global short-form video hosting and social media platform.
3+ YOEBachelor's or equivalent, 3+ years programming experience (C,C++,Java,Python,Perl,Go), Unix/Linux and IP networking expertise, production operations and automation experience, strong communication and ownership.
C, C++, Java, Python, Perl, Go, Unix, Linux, IP networking
Site Reliability Engineer - USDS (Multiple Positions)
San Jose, California, United States
$188k-$259k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
1+ YOERequires degree in CS/Engineering/Information Systems/Data Science/Mathematics plus relevant experience; experience with Linux administration, monitoring, troubleshooting, SDLC and cloud-native operations.
San Francisco or New York or Denver or Austin or Calgary or Toronto
$118k-$135k/yrHybridFull Time
IntersectNASDAQ: GOOGL: Develops-located data centers and renewable energy infrastructure.
1+ YOEB.S. in Electrical Engineering; 1–3 years field experience; interpret electrical diagrams; analyze data from SCADA/CMMS; strong communication; safety-focused.
San Francisco or San Jose or New York City or Milpitas or Mountain View or Holmdel or Goleta or Redwood City or Fremont or Sunnyvale or Brooklyn or Palo Alto
$187k-$268k/yrHybridFull Time
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
6+ YOERequires 8+ years with a bachelor's, 6+ with a master's, or 3+ with a PhD; 6+ years in SRE or infrastructure engineering, 5+ years operating Kubernetes, cloud, CI/CD, and Python or Go.
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Design and build secure end-to-end solutions, develop server-side systems, APIs and tooling to operate large-scale services while upholding privacy and high performance.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOERequires 5+ years operating Linux systems in production or HPC environments, large-scale storage experience, incident response, monitoring, Kubernetes, CI/CD, Python or Go, and Terraform or Ansible.