1,685 observability jobs at 772 companies in California
1mo
Save
Mark Applied
Hide
1mo
Senior Observability Engineer
Woodland Hills, California, United States
$120k-$130k/yrOnsiteFull Time
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
6+ YOE6+ years experience in observability/telemetry architecture, alerting optimization, dynamic baselining, SLO governance, and telemetry data pipelines.
Atlanta or Jersey City or New York City or Los Angeles or New Jersey or United States or Canada or United Kingdom or Ireland or Portugal or Romania or Australia or Puerto Rico
$149k-$186k/yrHybridFull Time
FanDuelNYSE: FLUT: Offers online sports betting and daily fantasy sports services.
Hands-on observability/SRE experience; expertise with Datadog, monitoring, alerting, incident management, SLOs/SLIs; Kubernetes, AWS, Terraform; proficiency in Go/Java/Python/TypeScript; strong analytical and collaboration skills.
Rednote: A lifestyle-focused social media and e-commerce discovery platform.
3+ YOEBachelor's degree or higher, 3+ years of computer science experience, Java or Go proficiency, distributed systems and concurrency knowledge, observability tools experience, and fluent English and Chinese.
Warsaw or Bellevue or Ireland or Livingston or London or New York or Sunnyvale
OnsiteFull Time
CoreWeaveNasdaq: CRWV: Cloud computing infrastructure specialized for large-scale AI workloads.
Experience building and operating network observability: Prometheus/Grafana/Alertmanager, gNMI/SNMP, Python/Go/Bash, Kubernetes, Linux networking; telemetry collectors/exporters, on-call support, and mentoring junior engineers.
Prometheus, Grafana, Alertmanager, gNMI, SNMP, Python, Golang, Go, Bash, Kubernetes, Ansible, Jinja2, TensorFlow, scikit-learn, OpenTelemetry, Jaeger, Zipkin, Arista EOS, NVIDIA Cumulus Linux, Nokia SR OS, SR Linux, Linux
CoupangNYSE: CPNG: Provides online retail, grocery delivery, and video streaming services.
8+ YOEBachelor's in CS/EE/Math,8+ years building large-scale distributed systems,deep observability experience (metrics,logs,tracing),SLO/KPI definition,cloud and container familiarity,programming in Go/Java/Python/Ruby.
Cirrascale: Provides specialized GPU-based cloud infrastructure for AI workloads.
5+ YOEBachelors in CS/CE or equivalent; 5+ years observability and distributed systems experience; 1+ year HPE OpsRamp; strong Bash and Python; experience with OpenTelemetry, Prometheus, Grafana, Datadog, ELK, ThousandEyes; cloud skills (AWS/GCP/OpenStack/Proxmox/k8s).
San Francisco or Toronto or New York City or Montreal
$165k-$330k/yrHybridFull Time
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Deep observability experience, scalable telemetry pipeline knowledge, Prometheus, Grafana, ClickHouse, or OpenTelemetry experience, and proficiency in Python, Rust, or Go.
Prometheus, Grafana, ClickHouse, OpenTelemetry, Python, Rust, Go
Clockwork Systems: Software-driven network fabrics for GPU cluster performance optimization.
Lead architecture and development of a high-performance network observability platform; strong distributed systems, Linux networking, and observability tooling; mentor engineers.
PinterestNYSE: PINS: Visual discovery engine for finding inspiration and creative ideas.
7+ YOE7+ years in distributed systems and data engineering; expert in Java, Python, Go, or Scala; strong observability (metrics/logs/traces) with OpenTelemetry/Prometheus/Grafana; experience building scalable observability platforms; cloud-native (Kubernetes); product mindset and collaboration skills.
Databricks: A unified platform for data analytics and artificial intelligence.
12+ YOE12+ years in distributed systems, observability or governance; strong CS fundamentals; cross-functional communication; BS in CS (MS/PhD a plus).
Sr. Partner Dev Mgr, DevOps and Observability, AMER, Observability
Austin or Seattle or Santa Monica or New York or East Palo Alto or Chicago or United States or San Francisco
$148k-$220k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
6+ YOE6+ years GTM/business development/sales experience, 5+ years negotiating business agreements, experience with Observability/DevOps/Data/Analytics/SaaS, strong strategic and co-selling skills.
Railway: Infrastructure platform for automated application deployment and cloud hosting.
Experience building distributed systems, observability tooling, backend services in Golang/Rust, GRPC; familiarity with Terraform and Ansible; strong communication and ownership.
Bengaluru or San Francisco or Boston or New York City or Austin or Tokyo or London
HybridFull Time
Postman: Platform for building, testing, and managing software APIs.
10+ YOE10+ years software engineering experience with distributed systems, observability expertise, production operations, strong programming in Go/Java/Python/Node.js, and experience with monitoring/logging/tracing tools.
RobinhoodNASDAQ: HOOD: Provides a commission-free platform for investing and financial services.
8+ YOE8+ years software engineering experience, deep Kubernetes and public cloud (AWS) expertise, strong coding in Go/Python, experience owning observability/control-plane and telemetry pipelines.