32 platform reliability engineer jobs at 21 companies in Kent, WA

1mo
Save
Mark Applied
Hide
Senior Platform Reliability Engineer
San Francisco or New York City or Seattle
$182k-$250k/yr HybridFull Time
Grow Therapy
Grow Therapy: Platform connecting mental health providers with patients and insurance.
6+ YOE6+ years operating production systems; hands-on AWS, Kubernetes (EKS), Terraform; experience defining SLOs/SLAs and observability (DataDog); strong communication and systems-thinking skills; PostgreSQL experience a plus.
AWS, Kubernetes, EKS, Terraform, DataDog, PostgreSQL, Gem
4w
Save
Mark Applied
Hide
Site Reliability Engineer - Compute Platform
Seattle, Washington, United States
$148k-$368k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Bachelor's degree in CS/Engineering, strong Linux, networking, database knowledge, Kubernetes and SRE tool experience, coding in Python/Shell/Java/Go, and strong problem-solving skills.
ClickHouse, Spark, Presto, Doris, Hadoop, Kubernetes, Linux, Python, Shell, Java, Go
1mo
Save
Mark Applied
Hide
Sr. Kubernetes Platform Site Reliability Engineer (Starlink)
Redmond, Washington, United States
$165k-$230k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOEBachelor's in CS/IT/engineering + 5+ years SRE/DevOps (or 7+ years experience), Linux, Terraform/Ansible, Kubernetes/OCI containers, scripting (Bash/Python), development in Python/C++/Go, strong networking and CI/CD knowledge.
Linux, Terraform, Ansible, Kubernetes, OCI containers, Bash, Python, C++, Go, Bazel, Makefiles, TCP/IP
2w
Save
Mark Applied
Hide
Service Reliability Engineer (SRE)
Seattle, Washington, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Design and build secure end-to-end server-side solutions and APIs for large-scale media and service platforms; collaborate across teams to deliver reliable services globally.
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Platform Responsibility - USDS
Seattle, Washington, United States
$178k-$342k/yr OnsiteFull Time
TikTok USDS Joint Venture
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
5+ YOE5+ years SRE/DevOps experience, BS in CS or related, strong Unix/Linux and distributed systems knowledge, experience with observability stacks and incident response, familiarity integrating AI/LLM into workflows.
Prometheus, Grafana, DataDog, Kubernetes, Kafka
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer - Core Cloud Platform
San Francisco or San Jose or Bellevue
$240k-$356k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
7+ YOE7+ years SRE or production infrastructure experience, deep Kubernetes and Terraform knowledge, experience with observability and SLOs, proficiency in Go or Python, on-call and incident leadership experience.
Kubernetes, Terraform, Argo CD, Flux, Helm, Kustomize, OpenTelemetry, Prometheus, Grafana, Datadog, Go, Python, etcd, GitOps
1w
Save
Mark Applied
Hide
Sr. Site Reliability Engineer I
Seattle, Washington, United States
$134k-$215k/yr HybridFull Time
Axon
AxonNASDAQ: AXON: Develops weapons and software for law enforcement and safety.
5+ YOE5+ years platform engineering experience with Kubernetes and cloud (AWS/Azure), programming (Python/Go/C#/Java), CI/CD, IaC (Terraform/Pulumi), observability, and strong debugging and documentation skills.
Kubernetes, AKS, EKS, Python, Go, C#, Java, CI/CD, APM, Terraform, Pulumi
3w
Save
Mark Applied
Hide
Staff Engineer, Deployment Platform
Seattle, Washington, United States
HybridFull Time
Stripe
Stripe: Provides online payment processing and financial infrastructure for businesses.
10+ YOE10+ years building and shipping production infrastructure systems; deep distributed-systems and deployment orchestration expertise; Kubernetes, container, reliability, incident response, and large cross-team project leadership experience.
Kubernetes, Kafka, Terraform, AWS, Azure
1d
Save
Mark Applied
Hide
Lead Software Engineer - AI Platform Reliability
Seattle, Washington, United States
$152k-$215k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years software engineering experience, strong Python coding, system design, observability and incident response, experience with AI/ML platforms preferred, mentor engineers and participate in on-call rotations.
Python
4w
Save
Mark Applied
Hide
Tech Lead Cloud Site Reliability Engineer - DCS Cloud
Seattle, Washington, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
5+ YOEBachelor's in CS or related,5+ years SRE/Linux/DevOps experience,proficient in Go/Python/C++,familiar with public cloud platforms,monitoring,incident response,and strong troubleshooting and communication skills.
Linux, Go, Python, C++, OCI, AWS, Azure, GCP, KVM/QEMU, Docker, Kubernetes, containerd, cgroups, namespaces, CUDA, MIG
1mo
Save
Mark Applied
Hide
Staff Software Engineer, Storage Platform
Bellevue or Menlo Park or Toronto
$230k-$270k/yr HybridFull Time
Robinhood
RobinhoodNASDAQ: HOOD: Provides a commission-free platform for investing and financial services.
5+ YOEDeep expertise in PostgreSQL/Aurora, distributed systems (sharding, replication, transactions), proficiency in Go or Rust, experience with Kubernetes and AWS services, and strong reliability/performance engineering skills.
PostgreSQL, Aurora PostgreSQL, Go, Rust, Kubernetes, RDS, DynamoDB
3w
Save
Mark Applied
Hide
Senior Infrastructure Engineer
Seattle, Washington, United States
$160k-$220k/yr OnsiteFull Time
Aurelian
Aurelian: AI-powered voice automation for 9-1-1 non-emergency calls.
4+ YOE4+ years in infrastructure/platform/backend engineering, experience with reliability and scale, comfortable across backend and cloud, experience building analytics/observability/developer tooling.
ClickHouse
3mo
Save
Mark Applied
Hide
Staff Software Engineer, Kubernetes Platform
San Francisco or New York City or Seattle
$320k-$405k/yr HybridFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
Significant software engineering experience with distributed systems; deep Kubernetes expertise; proficiency in Go, Python, Rust, or C++; strong debugging and reliability focus; excellent communication.
Kubernetes, etcd, apiserver, controller-runtime, Go, Python, Rust, C++, Linux, eBPF
2w
Save
Mark Applied
Hide
Staff Software Engineer, Test Platform
San Francisco or Seattle
$154k-$264k/yr OnsiteFull Time
SoFi
SoFiNasdaq: SOFI: Mobile-first platform for banking, lending, and investment services.
8+ YOEBachelor's or Master's in CS or related,8+ years software development,cloud (AWS),containers (Docker,Kubernetes),distributed systems,Java/Kotlin/Python/Go,automated testing and reliability expertise.
AWS, Docker, Kubernetes, Istio, Envoy, Java, Kotlin, Python, Go, Locust, Artillery, Cypress, Gremlin, AWS FIS, Datadog, Elastic, Splunk, Argo, GitLab CI/CD
1mo
Save
Mark Applied
Hide
Staff Software Engineer, Infrastructure
Canada or United States or Seattle or Paris or New York City
$238k-$382k/yr RemoteFull Time
Docker
Docker: Provides a platform for building, sharing, and running containerized applications.
8+ YOE8+ years backend/infrastructure engineering; strong Go; experience with Kubernetes/EKS, cloud platforms, networking, and reliability; Bachelor's or equivalent; experience leading cross-team technical initiatives and strong written/verbal communication.
Go, Terraform, Argo CD, EKS, Envoy Gateway, Grafana Cloud, OpenTelemetry, Prometheus, Grafana, GitHub Actions, Kubernetes, Linux, GitOps, Covey Scout
1mo
Save
Mark Applied
Hide
Principal AI Systems Engineer — C++ / Applied AI
San Jose or San Francisco or Seattle or New York City or Chicago or California
$190k-$361k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years professional software engineering with deep C++ expertise, production integration of AI/LLMs, cross‑platform systems design, reliability and observability, architecture leadership, and strong communication skills.
C++, LLMs, GPT, Claude, Gemini, JSON-RPC, gRPC, WebSockets, REST, TLS, CI, Windows, macOS, Linux
3mo
Save
Mark Applied
Hide
SWE - Backend Infrastructure Engineer
San Francisco or Bellevue or New York
$175k-$280k/yr OnsiteFull Time
Sesame
Sesame: Designing wearable computers with lifelike voice-driven AI agents.
3+ YOEStrong systems thinker with reliability engineering experience; 3+ years in infrastructure, platform, or ML systems; Kubernetes production experience; strong communication.
Kubernetes, Terraform, CloudFormation, Pulumi, TorchServe, Seldon, KServe, Ray Serve, PyTorch, APIs, Database design
2mo
Save
Mark Applied
Hide
Staff Software Engineer - Parameters
Bellevue, Washington, United States
$236k-$339k/yr OnsiteFull Time
Snowflake
SnowflakeNYSE: SNOW: Cloud-based platform for data storage, processing, and analytics.
8+ YOE8+ years in software engineering with distributed systems; cloud platforms (AWS/Azure/GCP); strong Java/Scala/C++/Python skills; leadership and collaboration; reliable, scalable infra experience.
Java, Scala, C++, Python, AWS, Azure, GCP
2w
Save
Mark Applied
Hide
Senior Systems Software Engineer
Seattle or United States
$102k-$210k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
3+ YOEDesign and develop systems software for hyperscale server/storage/GPU platforms, cross-layer debugging with firmware/hardware, automate provisioning, and maintain reliability; 3+ years experience expected.
CI/CD, API
1mo
Save
Mark Applied
Hide
Sr. Director of Engineering, Managed PostgreSQL and AI-Native Database Platform
Seattle, Washington, United States
$262k-$327k/yr HybridFull Time
DigitalOcean
DigitalOceanNew York Stock Exchange: DOCN: Simplifies cloud infrastructure for developers, startups, and SMBs.
15+ YOE6+ Mgmt15+ years engineering experience, 6+ years managing managers, deep distributed systems and database expertise, proven production reliability operations, strategic roadmap and executive communication skills.
pgvector, pgvectorscale, RDS/Aurora, AlloyDB, Cloud SQL, Azure Database for PostgreSQL, Neon, Crunchy, EDB, Yugabyte, Citus, Timescale, Supabase

Explore Jobs

Expand Your Job Search