392 observability engineer jobs at 230 companies in Cotati, CA
1mo
Save
Mark Applied
Hide
1mo
Platform Engineer, APIs & Observability
San Francisco or New York City
$180k-$200k/yrHybridFull Time
StackAI: Build and deploy custom AI agents without writing code.
4+ YOE4+ years building backend services and public APIs; strong REST/OpenAPI skills, observability and distributed tracing, Python and FastAPI, familiarity with TypeScript/Node.js, and analytics-driven metrics and reporting.
Railway: Infrastructure platform for automated application deployment and cloud hosting.
Experience building distributed systems, observability tooling, backend services in Golang/Rust, GRPC; familiarity with Terraform and Ansible; strong communication and ownership.
Observability Lead - Cloud SRE & Network Reliability (193698)
Fremont or San Francisco or Oakland
$114k-$253k/yrHybridFull Time
Lam ResearchNASDAQ: LRCX: Manufacturing equipment used to fabricate advanced semiconductor microchips.
12+ YOE6+ MgmtBS/MS/PhD or equivalent, 12+ years in infrastructure/SRE/DevOps/network engineering, 6+ years leading SRE/observability teams; multi-cloud networking, DR/BCP, observability platforms, IaC, automation, Python/Go experience.
PinterestNYSE: PINS: Visual discovery engine for finding inspiration and creative ideas.
7+ YOE7+ years in distributed systems and data engineering; expert in Java, Python, Go, or Scala; strong observability (metrics/logs/traces) with OpenTelemetry/Prometheus/Grafana; experience building scalable observability platforms; cloud-native (Kubernetes); product mindset and collaboration skills.
Bengaluru or San Francisco or Boston or New York City or Austin or Tokyo or London
HybridFull Time
Postman: Platform for building, testing, and managing software APIs.
10+ YOE10+ years software engineering experience with distributed systems, observability expertise, production operations, strong programming in Go/Java/Python/Node.js, and experience with monitoring/logging/tracing tools.
SalesforceNYSE: CRM: Provides cloud-based customer relationship management and enterprise software.
Experience building and maintaining high-volume log pipelines and distributed observability services; strong communication, mentoring, testing, debugging, and security knowledge; familiarity with observability tooling.
Databricks: A unified platform for data analytics and artificial intelligence.
12+ YOE12+ years in distributed systems, observability or governance; strong CS fundamentals; cross-functional communication; BS in CS (MS/PhD a plus).
Principal Software Engineer, Full Stack (Observability)
San Francisco, California, United States
$197k-$314k/yrHybridFull Time
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
10+ YOE10+ years of software engineering experience and 3+ years in technical leadership or a principal role; Java, React or LWC, distributed systems, observability, streaming, and AI-assisted development expertise required.
Claude Code, Java, React, Lightning Web Components (LWC), OpenTelemetry, Kafka, HBase, ClickHouse, Grafana, PromQL, AI coding assistants
LaunchDarkly: Provides a feature management platform for software development teams.
5+ YOE5+ years of professional software engineering; TypeScript/React frontend and Go backend; experience with integrations, data pipelines or APIs; IaC tooling; RBAC and GitOps; strong communication.
Senior Software Engineer - Observability and Reliability
New York City or San Francisco or London or Sydney
$170k-$240k/yrOnsiteFull Time
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
5+ YOE5+ years building high-quality software, strong CS fundamentals, experience building observability tools, proficiency with Go, OpenTelemetry, Kubernetes, participation in on-call rotations, cloud service administration (GCP/AWS/Azure) preferred.
HUD: Platform for reinforcement learning environments and AI agent evaluations.
Production infra and backend experience owning uptime, performance, deployment safety, and cost; strong AWS, Kubernetes/EKS, Terraform, CI/CD, observability, and backend engineering skills.
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Staff-level experience building production infrastructure software, strong distributed systems and telemetry pipeline background, networking and high-performance network knowledge, experience processing high-volume operational data.
Principal Engineer - Isovalent Secure Workload Observability (remote)
Milpitas or San Jose or Sunnyvale or Mountain View or San Francisco or Palo Alto
$231k-$332k/yrHybridFull Time
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
10+ YOE10+ years building software (Go, C++, Java), 5+ years technical leadership, degree in computer science/engineering or related field, experience with scalable distributed systems, Kubernetes, cloud APIs, algorithms and performance engineering.
Triumph: Skill-based mobile gaming platform for real money tournaments.
Experience operating and scaling large production systems; deep Postgres knowledge; CI/CD; observability tooling; able to lead a function independently.
Felt Technologies: Provides embedded telehealth and provider network infrastructure for applications.
Proficiency with NextJS, React, Postgres; experience with serverless environments, observability, and healthtech/PII/PHI best practices; mobile SDK experience (Swift/Kotlin) is a plus. Must work US timezones.
Mondrio: AI-native platform for agentic pricing and revenue management.
8+ YOE8+ years engineering experience with production LLMs, built evals/observability, shipped LLM features, strong communication, product-engineer instincts, US work authorization required.
Clera: AI talent agent matching professionals with high-growth startup roles
8+ YOE8+ years engineering experience with production LLM systems, building evals and observability, experience with LLM agents and data-residency/SOC2/GDPR constraints, strong communication and product-engineer instincts.
Loft Orbital: Deploy and operate satellite missions for organizations and governments.
5+ YOERequires 5+ years in space systems engineering on Earth observation missions, expertise in imagery calibration, image quality metrics and optical payloads, Python coding experience, and strong communication and organizational skills.