165 software reliability engineer jobs at 110 companies in Cotati, CA

2mo
Save
Mark Applied
Hide
Software Reliability Engineer
Mountain View or San Francisco
$175k-$215k/yr HybridFull Time
Waymo
Waymo: Autonomous driving technology for ride-hailing and logistics.
2+ YOE2+ years in C++, Java, or Python; interest in distributed and production systems; BS degree or equivalent experience; 3+ years preferred.
C++, Java, Python
1mo
Save
Mark Applied
Hide
Senior Reliability Engineer
Alameda, California, United States
$90k-$180k/yr OnsiteFull Time
Abbott
AbbottNYSE: ABT: Manufactures medical devices, diagnostics, and nutritional health products.
4+ YOEAssociate's degree (engineering preferred), minimum 4 years in FDA/ISO-regulated environments, experience in failure analysis, hardware/software integration, embedded systems, and strong project management skills.
1mo
Save
Mark Applied
Hide
Senior Reliability Engineer
Alameda, California, United States
$90k-$180k/yr OnsiteFull Time
Abbott
AbbottNYSE: ABT: Provides medical devices, diagnostics, and science-based nutritional products.
4+ YOEAssociate degree in engineering, 4+ years in FDA/ISO-regulated environment, experience in product failure analysis, hardware/software integration, embedded systems and electrical design, strong problem-solving and project management skills.
1mo
Save
Mark Applied
Hide
Software Engineer, Reliability Platforms
San Francisco or Sunnyvale or New York City
$160k-$235k/yr OnsiteFull Time
DoorDash
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
5+ YOE5+ years in infrastructure/platform/backend engineering; fluent in Go or similar; AWS, containerization, and IaC experience (Terraform or Pulumi); SRE concepts (SLOs, error budgets); platform engineering mindset and familiarity with AI tools.
Go, AWS, Terraform, Pulumi
4w
Save
Mark Applied
Hide
Staff Software Engineer, Reliability Engineering & SDLC Governance
Carlsbad or Germantown or San Jose or San Francisco or New York City
$165k-$261k/yr OnsiteFull Time
Viasat
ViasatNASDAQ: VSAT: Provides global satellite broadband and secure networking communication services.
8+ YOE8+ years software engineering experience with SDLC governance, SRE/DevOps knowledge, observability, incident analysis, and cross-team influence to improve reliability and customer experience.
Prometheus, Grafana, OpenTelemetry, Datadog, AS9115
1mo
Save
Mark Applied
Hide
Staff Quality & Reliability Engineer
San Francisco or New York City
HybridFull Time
Beast Industries
Beast Industries: Produces digital media and consumer goods for MrBeast brands.
Expert in software quality engineering and site reliability for consumer-scale distributed systems; owns test strategy, SLOs/error budgets, CI/CD test gates, observability, incident response, and reliability tooling.
CI/CD
3w
Save
Mark Applied
Hide
Site Reliability Engineer
San Francisco, California, United States
HybridFull Time
Runloop
Runloop: Provides infrastructure and secure sandboxes for AI agents.
5+ YOE5+ years software engineering experience with 3+ years in SRE/DevOps, strong Python or Go skills, containerization, cloud infra, monitoring, networking, Linux administration, on‑call and incident management.
AWS, GCP, Azure, Grafana, Prometheus, Datadog, Python, Go, Docker, Kubernetes, Terraform, Pulumi, Sentry, RUM, CI/CD
4d
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
San Francisco, California, United States
$350k-$475k/yr OnsiteFull Time
Thinking Machines
Thinking Machines: Building AI systems to extend human will and judgment.
Experience in distributed systems/cloud/site reliability, software automation for reliability, incident response and postmortems, strong communication and coordination skills.
Tinker, Kubernetes, LoRA, CI/CD
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Platform for personalized customer engagement and cross-channel messaging.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
1mo
Save
Mark Applied
Hide
Software Engineer, Infrastructure & Reliability
San Francisco, California, United States
HybridFull Time
CrewAI
CrewAI: Platform for orchestrating collaborative multi-agent AI systems.
Experience building and operating production SaaS infrastructure: cloud, containers, CI/CD, observability, secrets, databases, and automation using Python/Ruby/Go/Bash.
AWS, Docker, CI/CD, GitHub Actions, ECS, ECR, Kubernetes, Helm, PostgreSQL, Redis, Celery, FastAPI, Rails, Sentry, OpenTelemetry, Python, Ruby, Go, Bash, Terraform
5d
Save
Mark Applied
Hide
Staff Software Engineer, Quality & Reliability
Los Angeles or San Francisco or Toronto or Raleigh or United States
$172k-$229k/yr HybridFull Time
BuildOps
BuildOps: SaaS platform for managing commercial contracting businesses.
Extensive experience solving cross-cutting reliability and quality problems, leading multi-team initiatives, systems thinking, cloud (AWS) experience, strong programming in TypeScript or Java, observability and CI/CD familiarity, and strong communication.
TypeScript, Java, AWS, CI/CD
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
San Francisco, California, United States
$149k-$224k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
5+ YOE5+ years systems and software engineering experience for large-scale internet services; expertise in SRE principles, containers, observability, incident management, Python and Go, and applying AI/ML to operations.
Temporal, Airflow, Argo Workflows, Docker, Kubernetes, DNS, HTTP, Grafana, Prometheus, ELK, Splunk, Datadog, Python, Go, Linux, Claude Code, GitHub Copilot, Codex, Cursor, AWS, GCP, MCP
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Pleasanton or Austin or San Francisco or United States
OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOESenior SRE with strong infrastructure, automation, and programming experience (Terraform, Chef, Ansible, Python, Java, Bash). Minimum multi-year experience in software engineering or equivalent; participates in on-call and incident response.
Terraform, Chef, Ansible, Python, Java, Bash, Kubernetes, Helm, Jenkins, Grafana, Prometheus, OCI - DevOps, Oracle Cloud Guard, Oracle Observability and Management
1mo
Save
Mark Applied
Hide
Senior Software Engineer - Observability and Reliability
San Francisco, California, United States
$170k-$240k/yr OnsiteFull Time
Sigma Computing
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
5+ YOERequires strong CS fundamentals, 5+ years building and maintaining software, experience with Go, Open Telemetry, Kubernetes, cloud platforms (GCP/AWS/Azure) and on-call/incident management.
Go, Open Telemetry, Kubernetes, GCP, AWS, Azure, SQL, Python
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
2mo
Save
Mark Applied
Hide
Founding Engineer - Site Reliability
San Francisco or United States
$185k-$285k/yr RemoteFull Time
uRun
uRun: Infrastructure cloud for interactive, stateful AI inference.
7+ YOE7+ years in site reliability or infrastructure engineering; strong SLOs, incident response, and observability; Kubernetes and cloud (AWS); software engineering fundamentals; first SRE at a company.
Kubernetes, AWS, Prometheus, Grafana, Datadog, Automation, VPC, GPU compute
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Lincoln or San Francisco
$125k-$165k/yr RemoteFull Time
TELCOR
TELCOR: Provides healthcare software for laboratory and point-of-care operations.
2+ YOEExperience with distributed systems, Redis, queuing systems, Kubernetes (2+ yrs), AWS, Terraform (2+ yrs), observability, and production operations.
Redis, Kubernetes, AWS, Terraform
1w
Save
Mark Applied
Hide
Sr. Software Engineer
San Francisco, California, United States
$166k-$200k/yr HybridFull Time
Pilot
Pilot: Software-powered bookkeeping, tax, and CFO services for businesses.
5+ YOE5+ years software engineering experience, production Python, strong engineering fundamentals, ownership of end-to-end systems, production reliability and observability, strong communication and mentoring skills.
Python, JavaScript, TypeScript, Vue.js, Terraform, AWS, Postgres
2mo
Save
Mark Applied
Hide
Senior/Staff Software Engineer, Core Infrastructure
San Francisco, California, United States
$160k-$210k/yr RemoteFull Time
Zip
Zip: AI-powered intake-to-procure platform for enterprise spend management
6+ YOE6+ years software engineering in infrastructure; BS or higher in CS or related; Kubernetes/EKS, multi-region, observability; experience in a small company; quick learner.
Kubernetes, EKS, Networking, Observability, Reliability, Performance Engineering
3w
Save
Mark Applied
Hide
Senior Software Engineer
San Francisco, California, United States
$210k-$240k/yr OnsiteFull Time
Fazeshift
Fazeshift: AI-powered automation platform for enterprise accounts receivable workflows.
5+ YOE5+ years professional software engineering experience, strong TypeScript and full-stack skills, API and reliability instincts, startup experience preferred, mentoring and ownership mindset.
TypeScript, React