31 devops reliability engineer jobs at 26 companies in Cotati, CA

2mo
Save
Mark Applied
Hide
DevOps / Site Reliability Engineer
San Francisco, California, United States
OnsiteFull Time
Reactor
Reactor: Building infrastructure for real-time generative world models.
Production Kubernetes, CI/CD, GitOps, IaC; Go/Python/Bash tooling; on-call; incident response; observability and SLOs; multi-cloud.
Kubernetes, CI/CD, GitOps, Helm, Kustomize, Terraform, Go, Python, Bash, Docker, Observability, SRE tooling
1mo
Save
Mark Applied
Hide
Site Reliability/Devops Engineer
San Francisco, California, United States
$100k-$200k/yr OnsiteFull Time
Graphon
Graphon: Developing graph-native AI models for multimodal data reasoning.
Proficient in Bash and Python; experience with infrastructure-as-code, Docker, CI/CD, multi-cloud deployments, networking and identity access; comfortable managing production environments and using AI tools.
Bash, Python, Infrastructure-as-code, Docker, CI/CD, AI tools
2mo
Save
Mark Applied
Hide
DevOps Engineer
San Francisco, California, United States
$120k-$200k/yr OnsiteFull Time
Droyd
Droyd: Builds autonomous robotic systems and in-house hardware, enabling robots to operate reliably in real production environments.
Experienced DevOps/infrastructure engineer with strong Linux/Ubuntu skills, Jetson/JetPack and cross-compilation (x86_64→aarch64) experience, Docker Compose/systemd, OTA and fleet management, and reliability-focused tooling.
Linux, Ubuntu, Jetson, JetPack, Docker Compose, systemd, JWKS, QUIC, Balena, Balena Etcher
1mo
Save
Mark Applied
Hide
DevOps Engineer
San Francisco, California, United States
RemoteFull Time
CareerSwift
CareerSwift: AI-powered recruitment and job search automation platform
4+ YOE4+ years DevOps experience; strong knowledge of Docker, Kubernetes, Terraform, CI/CD, AWS/GCP, and scripting with Python/Bash; experience with monitoring and improving deployment reliability, scalability, and security.
CI/CD, AWS, GCP, Docker, Kubernetes, Terraform, Python, Bash
1mo
Save
Mark Applied
Hide
Technology, DevOps/Site Reliability Engineer
New York or San Francisco
$160k-$200k/yr HybridFull Time
BTIG
BTIG: Provides global institutional trading and investment banking services.
2+ YOE2+ years IT support experience, strong customer-facing skills, Windows 10/11 and Microsoft 365 proficiency, experience with SCCM/Endpoint Manager and ticketing systems (ServiceNow); willing to obtain MS900.
ServiceNow, Windows 10, Windows 11, Microsoft Office 365, Microsoft OneDrive, System Center Configuration Manager, Endpoint Manager, Zoom, Bloomberg, Thomson Reuters, ICE, Fidessa, Redi+, Global Relay, Cisco, Active Directory, VPNWIFI
3w
Save
Mark Applied
Hide
Site Reliability Engineer
San Francisco, California, United States
HybridFull Time
Runloop
Runloop: Provides infrastructure and secure sandboxes for AI agents.
5+ YOE5+ years software engineering experience with 3+ years in SRE/DevOps, strong Python or Go skills, containerization, cloud infra, monitoring, networking, Linux administration, on‑call and incident management.
AWS, GCP, Azure, Grafana, Prometheus, Datadog, Python, Go, Docker, Kubernetes, Terraform, Pulumi, Sentry, RUM, CI/CD
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Platform for personalized customer engagement and cross-channel messaging.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
3w
Save
Mark Applied
Hide
Lead DevOps Engineer
Sydney or Melbourne or San Francisco or London or United States or Australia
HybridFull Time
Deputy
Deputy: Workforce management software for scheduling and managing hourly employees.
6+ YOE2+ Mgmt6+ years DevSecOps experience, 2+ years lead/staff experience, expertise in platform reliability, infrastructure security, Kubernetes, cloud DR, SLO/SLI and incident response; Bachelor’s in CS or equivalent.
Kubernetes, AWS, GitHub, GitLab, Jenkins, Snyk, SAST, DAST, CI/CD, OAuth2, OpenID
5d
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Bellevue or San Francisco
$147k-$226k/yr OnsiteFull Time
Okta
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
5+ YOE5+ years SRE/DevOps experience; expert AWS multi-account governance; Terraform and Python automation; Kubernetes and observability experience; strong networking, Linux, security and documentation skills.
AWS, AWS Orgs, IAM, Identity Center, StackSets, Terraform, Python, GitLab, GitHub Actions, Kubernetes, Splunk, CloudWatch, Grafana, BGP, IPsec, VPCs, TGWs, VPC endpoints, Linux
5d
Save
Mark Applied
Hide
Staff Site Reliability Engineer
San Francisco, California, United States
$195k-$258k/yr RemoteFull Time
Circle
CircleNYSE: CRCL: Digital currency issuer and blockchain financial infrastructure provider.
6+ YOE6+ years SRE/DevOps experience preferred; strong Kubernetes, IaC (Terraform/Pulumi), cloud networking, Go or Python, CI/CD, observability, and distributed systems experience; blockchain node operation experience is a plus.
Kubernetes, Helm, Terraform, Pulumi, Go, Python, CI/CD, RBAC, VPCs, DNS, SQL
1mo
Save
Mark Applied
Hide
Customer Reliability Engineer - Infrastructure
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yr RemoteFull Time
Astronomer
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
Apache Airflow, AWS, Azure, CI/CD, GCP, Infrastructure as Code (IaC), Kubernetes, Linux, Python
3w
Save
Mark Applied
Hide
Site Reliability Engineer II
Scottsdale or San Francisco or Chicago or New York City
$86k-$126k/yr HybridFull Time
Early Warning Services
Early Warning Services: Operates payment and risk solutions for the financial industry.
2+ YOEBachelor's or equivalent, minimum 2 years DevOps/Dev/SRE experience, Linux/Unix experience, infrastructure automation (Chef/Ansible/Puppet, Terraform), containerization (Docker,Kubernetes), cloud (AWS/GCP/Azure), on-call rotation.
Linux, Unix, Chef, Ansible, Puppet, Terraform, Docker, Kubernetes, AWS, GCP, Azure, Java, Ruby, Python, JavaScript, Go
1mo
Save
Mark Applied
Hide
Senior Database Reliability Engineer
San Francisco, California, United States
$203k-$254k/yr OnsiteFull Time
Crunchyroll
Crunchyroll: Operates a global streaming platform for anime and manga.
8+ YOEBachelor's degree in CS/IT, 8+ years in database operations/SRE or related, hands-on IaC (Terraform/CloudFormation/Pulumi), AWS and CI/CD experience, 24x7 production support expertise, monitoring/observability tooling knowledge.
Terraform, CloudFormation, Pulumi, AWS, RDS/Aurora, DynamoDB, Datadog, CloudWatch, DevOps Guru, Database Performance Insights, CI/CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - Hiring Sprint
San Francisco, California, United States
$196k-$255k/yr HybridFull Time
Airbyte
Airbyte: Open-source data integration platform for automated data movement.
7+ YOE7+ years in infrastructure/platform engineering/SRE/DevOps; hands-on Kubernetes, Helm, Terraform; observability with Prometheus/Grafana/Datadog; CI/CD ownership; ability to read backend code; fluency with LLMs and agentic tools.
Kubernetes, Helm, Terraform, AWS, GCP, Prometheus, Grafana, Datadog, CI/CD, Java, Python, Airbyte, CDKs, LLMs
1w
Save
Mark Applied
Hide
Sr. Site Reliability Engineer, tvScientific
San Francisco or United States
$140k-$288k/yr HybridFull Time
Pinterest
PinterestNYSE: PINS: Visual discovery engine for finding inspiration and creative ideas.
4+ YOE4+ years SRE/DevOps experience with production AWS and Kubernetes; hands-on with ArgoCD, Terraform/Terragrunt, Helm, GitHub Actions, Bash or Python; strong troubleshooting and incident response skills.
AWS, Kubernetes, EKS, ArgoCD, GitOps, Terraform, Terragrunt, Helm, GitHub Actions, Bash, Python, Linux, IAM
2mo
Save
Mark Applied
Hide
Senior Staff Site Reliability Engineer
San Francisco or New York City or Chicago
$245k-$270k/yr HybridFull Time
Ironclad
Ironclad: AI platform for digital contract lifecycle management
8+ YOE8+ years DevOps/SRE; 5+ years coding; Kubernetes and GCP expertise; build resilient infra; GitOps with Terraform/Pulumi, CircleCI, ArgoCD; AI tooling experience; strong communication; cross-functional collaboration.
Kubernetes, Google Cloud Platform, Terraform, Pulumi, CircleCI, ArgoCD, Claude Code, Cursor, Zed
2d
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Charlotte or Chandler or San Francisco or Columbus
$119k-$224k/yr HybridFull Time
Wells Fargo
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
5+ YOERequires 5+ years in systems engineering or architecture, 5+ years SRE, cloud observability experience, hosting platforms, and familiarity with DevOps, Agile, and IT service management.
Elasticsearch, Kibana, Kafka, Airflow, Logstash, Grafana, Elastic APM, Jaeger, Zipkin, AWS, OCP, Kubernetes, PKS, Azure, VMware, Unix, Linux, Windows, Jenkins, Maven, Gradle, Groovy, Artifactory, GIT, Harness IO, Spinnaker, Terraform, UDeploy, AIOPS, ServiceNow, Remedy, Big Panda, Netcool
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Robotics & Cloud Infrastructure
Brooklyn or New York City or Richmond or Europe
$164k-$220k/yr RemoteFull Time
Bedrock Ocean Exploration
Bedrock Ocean Exploration: Maps the ocean floor using autonomous underwater robotic vehicles.
5+ YOE5+ years SRE/DevOps experience with on-call ownership; strong automation using Python/Go/Bash; Terraform and AWS hands-on; containerization (Docker, Kubernetes); observability (Prometheus, Grafana); Linux and networking expertise; East Coast location and US work authorization required.
Python, Go, Bash, Terraform, AWS, Docker, Kubernetes, Prometheus, Grafana, ROS 2, ROS, Jetson, Linux, IAM
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Spend
San Francisco, California, United States
$160k-$250k/yr HybridFull Time
Airwallex
Airwallex: Global financial platform for business payments and money management.
6+ YOE6+ years in SRE/DevOps/infrastructure-focused role; Bachelor's in Computer Science or Software Engineering; AWS/GCP, Kubernetes; lead SRE strategy; production systems with high availability and compliance.
AWS, GCP, Kubernetes, Observability, Incident Response
4w
Save
Mark Applied
Hide
Staff Software Engineer, Reliability Engineering & SDLC Governance
Carlsbad or Germantown or San Jose or San Francisco or New York City
$165k-$261k/yr OnsiteFull Time
Viasat
ViasatNASDAQ: VSAT: Provides global satellite broadband and secure networking communication services.
8+ YOE8+ years software engineering experience with SDLC governance, SRE/DevOps knowledge, observability, incident analysis, and cross-team influence to improve reliability and customer experience.
Prometheus, Grafana, OpenTelemetry, Datadog, AS9115