492 observability engineer jobs at 299 companies in American Canyon, CA
2mo
Save
Mark Applied
Hide
2mo
Platform Engineer, APIs & Observability
San Francisco or New York City
$180k-$200k/yrHybridFull Time
StackAI: Enterprise AI software helping businesses build and deploy no-code agents for automated workflows.
4+ YOE4+ years building backend services and public APIs; strong REST/OpenAPI skills, observability and distributed tracing, Python and FastAPI, familiarity with TypeScript/Node.js, and analytics-driven metrics and reporting.
Railway: Private cloud infrastructure platform that helps developers deploy and manage applications.
Experience building distributed systems, observability tooling, backend services in Golang/Rust, GRPC; familiarity with Terraform and Ansible; strong communication and ownership.
San Francisco or Toronto or New York City or Montreal
$165k-$330k/yrHybridFull Time
Baseten: Private AI infrastructure provider that trains, deploys, and serves artificial-intelligence models for businesses.
Deep observability experience, scalable telemetry pipeline knowledge, Prometheus, Grafana, ClickHouse, or OpenTelemetry experience, and proficiency in Python, Rust, or Go.
Prometheus, Grafana, ClickHouse, OpenTelemetry, Python, Rust, Go
PinterestNYSE: PINS: A visual discovery engine for finding inspiration.
7+ YOE7+ years in distributed systems and data engineering; expert in Java, Python, Go, or Scala; strong observability (metrics/logs/traces) with OpenTelemetry/Prometheus/Grafana; experience building scalable observability platforms; cloud-native (Kubernetes); product mindset and collaboration skills.
Bengaluru or San Francisco or Boston or New York City or Austin or Tokyo or London
HybridFull Time
Postman: Private software providing an API development and lifecycle-management platform for developers and organizations.
10+ YOE10+ years software engineering experience with distributed systems, observability expertise, production operations, strong programming in Go/Java/Python/Node.js, and experience with monitoring/logging/tracing tools.
Databricks: Data and AI software providing a unified platform.
12+ YOE12+ years in distributed systems, observability or governance; strong CS fundamentals; cross-functional communication; BS in CS (MS/PhD a plus).
Principal Software Engineer, Full Stack (Observability)
San Francisco, California, United States
$197k-$314k/yrHybridFull Time
SalesforceNYSE: CRM: The #1 AI CRM driving customer success together.
10+ YOE10+ years of software engineering experience and 3+ years in technical leadership or a principal role; Java, React or LWC, distributed systems, observability, streaming, and AI-assisted development expertise required.
Claude Code, Java, React, Lightning Web Components (LWC), OpenTelemetry, Kafka, HBase, ClickHouse, Grafana, PromQL, AI coding assistants
Blackhawk Network: Private American fintech providing businesses and consumers gift cards, rewards, incentives, and digital payment solutions.
6+ YOE6+ years in platform engineering/SRE/DevOps/architecture; experience designing large-scale observability for cloud/AWS; expertise with New Relic, Splunk, Datadog, Coralogix; OpenTelemetry and observability pipelines; strong communication.
LaunchDarkly: B2B feature-management software helping engineering teams safely manage code releases and AI agents in production.
5+ YOE5+ years of professional software engineering; TypeScript/React frontend and Go backend; experience with integrations, data pipelines or APIs; IaC tooling; RBAC and GitOps; strong communication.
Senior Software Engineer - Observability and Reliability
New York City or San Francisco or London or Sydney
$170k-$240k/yrOnsiteFull Time
Sigma Computing: Private cloud analytics platform helping business and technical teams analyze live warehouse data and build AI applications.
5+ YOE5+ years building high-quality software, strong CS fundamentals, experience building observability tools, proficiency with Go, OpenTelemetry, Kubernetes, participation in on-call rotations, cloud service administration (GCP/AWS/Azure) preferred.
HUD: Private AI platform building reinforcement-learning environments, evaluating models, and delivering post-training data to frontier AI labs.
Production infra and backend experience owning uptime, performance, deployment safety, and cost; strong AWS, Kubernetes/EKS, Terraform, CI/CD, observability, and backend engineering skills.
RobloxNYSE: RBLX: Global platform for user-created immersive digital experiences.
3+ Mgmt3+ years engineering management experience; strong background in building and operating large-scale distributed systems and data infrastructure; experience with observability or AI/ML infrastructure preferred; excellent communication and cross-functional collaboration skills.
Experience operating and scaling large production systems; deep Postgres knowledge; CI/CD; observability tooling; able to lead a function independently.
Felt Technologies: Felt is an AI-native enterprise GIS platform for teams building maps, dashboards, and apps from spatial data.
Proficiency with NextJS, React, Postgres; experience with serverless environments, observability, and healthtech/PII/PHI best practices; mobile SDK experience (Swift/Kotlin) is a plus. Must work US timezones.
Mondrio: AI-powered B2B pricing software that helps companies continuously optimize pricing and guide deals with expert support.
8+ YOE8+ years engineering experience with production LLMs, built evals/observability, shipped LLM features, strong communication, product-engineer instincts, US work authorization required.
Clera: AI-powered talent agent matching candidates to startup roles.
8+ YOE8+ years engineering experience with production LLM systems, building evals and observability, experience with LLM agents and data-residency/SOC2/GDPR constraints, strong communication and product-engineer instincts.