405 observability engineer jobs at 236 companies in Sonoma, CA
1mo
Save
Mark Applied
Hide
1mo
Platform Engineer, APIs & Observability
San Francisco or New York City
$180k-$200k/yrHybridFull Time
StackAI: Build and deploy custom AI agents without writing code.
4+ YOE4+ years building backend services and public APIs; strong REST/OpenAPI skills, observability and distributed tracing, Python and FastAPI, familiarity with TypeScript/Node.js, and analytics-driven metrics and reporting.
Railway: Infrastructure platform for automated application deployment and cloud hosting.
Experience building distributed systems, observability tooling, backend services in Golang/Rust, GRPC; familiarity with Terraform and Ansible; strong communication and ownership.
San Francisco or Toronto or New York City or Montreal
$165k-$330k/yrHybridFull Time
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Deep observability experience, scalable telemetry pipeline knowledge, Prometheus, Grafana, ClickHouse, or OpenTelemetry experience, and proficiency in Python, Rust, or Go.
Prometheus, Grafana, ClickHouse, OpenTelemetry, Python, Rust, Go
PinterestNYSE: PINS: Visual discovery engine for finding inspiration and creative ideas.
7+ YOE7+ years in distributed systems and data engineering; expert in Java, Python, Go, or Scala; strong observability (metrics/logs/traces) with OpenTelemetry/Prometheus/Grafana; experience building scalable observability platforms; cloud-native (Kubernetes); product mindset and collaboration skills.
Bengaluru or San Francisco or Boston or New York City or Austin or Tokyo or London
HybridFull Time
Postman: Platform for building, testing, and managing software APIs.
10+ YOE10+ years software engineering experience with distributed systems, observability expertise, production operations, strong programming in Go/Java/Python/Node.js, and experience with monitoring/logging/tracing tools.
SalesforceNYSE: CRM: Provides cloud-based customer relationship management and enterprise software.
Experience building and maintaining high-volume log pipelines and distributed observability services; strong communication, mentoring, testing, debugging, and security knowledge; familiarity with observability tooling.
Databricks: A unified platform for data analytics and artificial intelligence.
12+ YOE12+ years in distributed systems, observability or governance; strong CS fundamentals; cross-functional communication; BS in CS (MS/PhD a plus).
Principal Software Engineer, Full Stack (Observability)
San Francisco, California, United States
$197k-$314k/yrHybridFull Time
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
10+ YOE10+ years of software engineering experience and 3+ years in technical leadership or a principal role; Java, React or LWC, distributed systems, observability, streaming, and AI-assisted development expertise required.
Claude Code, Java, React, Lightning Web Components (LWC), OpenTelemetry, Kafka, HBase, ClickHouse, Grafana, PromQL, AI coding assistants
LaunchDarkly: Provides a feature management platform for software development teams.
5+ YOE5+ years of professional software engineering; TypeScript/React frontend and Go backend; experience with integrations, data pipelines or APIs; IaC tooling; RBAC and GitOps; strong communication.
Senior Software Engineer - Observability and Reliability
New York City or San Francisco or London or Sydney
$170k-$240k/yrOnsiteFull Time
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
5+ YOE5+ years building high-quality software, strong CS fundamentals, experience building observability tools, proficiency with Go, OpenTelemetry, Kubernetes, participation in on-call rotations, cloud service administration (GCP/AWS/Azure) preferred.
HUD: Platform for reinforcement learning environments and AI agent evaluations.
Production infra and backend experience owning uptime, performance, deployment safety, and cost; strong AWS, Kubernetes/EKS, Terraform, CI/CD, observability, and backend engineering skills.
Triumph: Skill-based mobile gaming platform for real money tournaments.
Experience operating and scaling large production systems; deep Postgres knowledge; CI/CD; observability tooling; able to lead a function independently.
Principal Engineer - Isovalent Secure Workload Observability (remote)
Milpitas or San Jose or Sunnyvale or Mountain View or San Francisco or Palo Alto
$231k-$332k/yrHybridFull Time
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
10+ YOE10+ years building software (Go, C++, Java), 5+ years technical leadership, degree in computer science/engineering or related field, experience with scalable distributed systems, Kubernetes, cloud APIs, algorithms and performance engineering.
Felt Technologies: Provides embedded telehealth and provider network infrastructure for applications.
Proficiency with NextJS, React, Postgres; experience with serverless environments, observability, and healthtech/PII/PHI best practices; mobile SDK experience (Swift/Kotlin) is a plus. Must work US timezones.
Mondrio: AI-native platform for agentic pricing and revenue management.
8+ YOE8+ years engineering experience with production LLMs, built evals/observability, shipped LLM features, strong communication, product-engineer instincts, US work authorization required.
Clera: AI talent agent matching professionals with high-growth startup roles
8+ YOE8+ years engineering experience with production LLM systems, building evals and observability, experience with LLM agents and data-residency/SOC2/GDPR constraints, strong communication and product-engineer instincts.
Loft Orbital: Deploy and operate satellite missions for organizations and governments.
5+ YOERequires 5+ years in space systems engineering on Earth observation missions, expertise in imagery calibration, image quality metrics and optical payloads, Python coding experience, and strong communication and organizational skills.