32 cloud reliability engineer jobs at 23 companies in West Hollywood, CA

1mo
Save
Mark Applied
Hide
Staff Cloud Reliability Engineer
Irvine or Los Angeles
$180k-$200k/yr OnsiteFull Time
Viant Technology
Viant TechnologyNASDAQ: DSP: Provides an AI-powered programmatic advertising platform for marketers.
8+ YOE8+ years in DevOps/SRE, 3+ years Linux, cloud (AWS/Google), serverless (AWS Lambda/Google Cloud Functions), Docker/Kubernetes, Terraform, CI/CD (GitHub Actions), Python or Go, SQL/BigQuery; participate in on-call rotation.
Linux, AWS, Google, AWS Lambda, Google Cloud Functions, Docker, Kubernetes, Terraform, GitHub Actions, Python, GoLang, SQL, Google BigQuery
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Los Angeles, California, United States
$140k-$199k/yr HybridFull Time
Green Dot
Green DotNYSE: GDOT: Provides mobile banking and payment solutions to consumers and businesses.
7+ YOE7+ years in release/reliability engineering, cloud platform experience (AWS/Azure/GCP), automated deployment and observability proficiency, scripting with PowerShell/Bash/Python, excellent troubleshooting and communication skills.
AWS, Azure, GCP, PowerShell, Bash, Python
1mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Orlando or Glendale
$176k-$235k/yr HybridFull Time
The Walt Disney Company
The Walt Disney CompanyNYSE: DIS: Produces media content and operates global theme parks.
10+ YOE10+ years experience; expertise in observability, multi-cloud (AWS/Azure/GCP), CI/CD, Terraform/Cloud Formation/Ansible/Chef, containers/Kubernetes, Git, and leadership/mentoring skills.
Gitlab, AWS CodeBuild, CodeDeploy, CodePipeline, Azure DevOps, Terraform, Cloud Formation, Ansible, Chef, Harness, Kubernetes, AWS, Azure, GCP, Datadog, New Relic, Dynatrace, Git
1mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Orlando or Glendale
$176k-$247k/yr HybridFull Time
The Walt Disney Company
The Walt Disney CompanyNYSE: DIS: Produces movies, operates theme parks, and provides streaming services.
10+ YOE10+ years experience; expertise in multi-cloud (AWS, Azure, GCP), observability, CI/CD, infrastructure as code (Terraform/CloudFormation), Linux systems administration; bachelor's degree or equivalent; strong leadership and communication.
AWS, Azure, GCP, Git, Gitlab, AWS CodeBuild, CodeDeploy, CodePipeline, Azure DevOps, Terraform, Cloud Formation, Ansible, Chef, Harness, Kubernetes, Datadog, New Relic, Dynatrace
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Los Angeles, California, United States
$140k-$180k/yr OnsiteFull Time
K2 Space
K2 Space: Develops high-power satellite platforms for heavy-lift launch vehicles.
5+ YOE5+ years SRE/DevOps experience or BS in CS/IT/STEM, deep cloud (AWS/GCP/Azure), IaC, Kubernetes, Linux, programming (Go/Python), strong security and reliability experience.
Infrastructure-as-Code (IaC), AWS, GCP, Azure, Kubernetes, Terraform, Ansible, Go, Python, Linux
4w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Glendale, California, United States
$138k-$221k/yr RemoteFull Time
ServiceTitan
ServiceTitanNASDAQ: TTAN: Cloud-based business management software for home service contractors.
8+ YOE8+ years experience building scalable distributed systems, expertise in cloud infrastructure, Kubernetes, CI/CD, logging/metrics, and strong programming skills.
.NET, C#, Visual Basic, PowerShell, Java, Jenkins, Team City, Kubernetes, Kafka, Event Hubs, SQS, Snowflake, Databricks Delta, API gateways, Azure, AWS, Git, Elasticsearch-Logstash-Kibana, DataDog, Grafana
4w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Sunnyvale or Sylmar
$90k-$180k/yr OnsiteFull Time
Abbott
AbbottNYSE: ABT: Provides medical devices, diagnostics, and science-based nutritional products.
Senior SRE with strong distributed systems, cloud (Azure), Kubernetes, observability, automation, incident management, and cross-functional communication skills for a medical device remote monitoring platform.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK, EFK, Datadog, Linux
2w
Save
Mark Applied
Hide
Sr. Site Reliability Engineer (Starlink)
Hawthorne or Palo Alto or Redmond
$165k-$270k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOEBachelor's in CS/engineering/math with 5 years software experience or 7+ years SRE/DevOps experience; Linux experience required; Kubernetes, Kafka, cloud-native tooling, and programming in Python/Go/Java/C#/Scala preferred.
Linux, Kubernetes, Istio, Apache Kafka, Apache Spark, HBase, HDFS, Apache Flink, Python, C#, Java, Scala, Go
1mo
Save
Mark Applied
Hide
Associate Site Reliability Engineer
Santa Monica, California, United States
$30-$56/hr OnsiteFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Develops computer software, consumer electronics, and video games.
1+ YOE1–3 years SRE/DevOps or cloud experience; familiarity with Linux, HTTP, DNS, containers, Kubernetes, Git, and scripting (Bash/Python); exposure to monitoring, logs, metrics, and incident management; strong troubleshooting and communication skills.
Linux, HTTP, DNS, Kubernetes, Git, Bash, Python
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, TS Clearance
Costa Mesa or Irvine or Washington
$191k-$287k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology building autonomous military hardware and software.
6+ YOEActive U.S. Top Secret clearance, 6+ years engineering experience, deep Kubernetes and cloud (AWS/Azure) expertise, experience with Terraform and infrastructure-as-code, proficiency in Go/Python/Rust/C++, and strong system troubleshooting skills.
Kubernetes, EKS, Docker, Helm, ArgoCD, Terraform, Python, AWS, Azure, Go, Rust, C++, KubeVirt, qemu, Linux, CI/CD, Lattice OS
3mo
Save
Mark Applied
Hide
Site Reliability Engineer II
Los Angeles, California, United States
$130k-$145k/yr OnsiteFull Time
AXS
AXS: Provides digital ticketing and marketing solutions for live events.
4+ YOE4-6 years in site reliability or DevOps; BA/BS preferred but not required; cloud operations, infrastructure as code, containers/orchestration, CI/CD; programming/scripting ability to automate tasks.
Cloud, Containers, Orchestration, Infrastructure as Code, CI/CD, Python, Bash, Go
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (FedRAMP / Security) - CA
Los Angeles or Israel
$170k-$350k/yr RemoteFull Time
Coralogix
Coralogix: AI-powered observability and security data platform.
5+ YOE5+ years SRE/DevOps experience, strong Kubernetes and cloud (AWS) skills, experience with monitoring and observability tools, infrastructure as code (Terraform/Crossplane), networking knowledge, FedRAMP/security experience preferred, Golang experience advantageous.
Kubernetes, Kops, AWS, Kafka, Prometheus, Thanos, Coralogix, Git, Argo CD, Istio, Grafana, Terraform, Crossplane, Golang, Apache Kafka
1d
Save
Mark Applied
Hide
10351 - Network Reliability Engineer
Irvine, California, United States
$115k-$125k/yr OnsiteFull Time
Hyundai AutoEver America
Hyundai AutoEver AmericaKorea Exchange: 307950: Provides automotive software and IT services for mobility systems.
Requires enterprise network engineering, Cisco routing and switching, Palo Alto firewalls, Cisco ISE, TCP/IP, BGP, OSPF, LAN/WAN, monitoring, telemetry, Linux, Python, Ansible, and automation experience.
Cisco, Cisco ISE, Palo Alto, Splunk, SolarWinds, Grafana, Prometheus, SNMP, syslog, NetFlow/IPFIX, Python, Ansible, REST APIs, Git, Linux, AWS, Azure, Google Cloud, OpenStack, Terraform, OpenTelemetry, ITSM
1w
Save
Mark Applied
Hide
Site Reliability Engineer III
Irvine, California, United States
$133k-$185k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
3+ YOE3+ years applied SRE experience, proficiency with SRE principles, one programming language (Python, Java/Spring Boot, .Net), observability, CI/CD, container orchestration, and cloud infrastructure.
Python, Java, Spring Boot, .Net, EKS, EC2, ALB, NLB, Route 53, Terraform, Kubernetes, TLS, mTLS, CloudWatch
2w
Save
Mark Applied
Hide
Staff AI Platform & Reliability Engineer
Santa Ana, California, United States
$172k-$220k/yr OnsiteFull Time
eJam
eJam: Builds and scales direct-to-consumer e-commerce brands.
Senior Python/GCP engineer to design and operate AI generation services, provider integrations, billing/usage metering, tenant security and platform reliability.
Python 3.12, FastAPI, Pydantic, asyncio, Cloud Run, Pub/Sub, Cloud Tasks, GCS, Firestore, Cloud SQL/Postgres, SQLAlchemy, Alembic, Terraform, IAM/OIDC, Secret Manager, CI/CD, NestJS, TypeScript, Claude Code, Codex, Cursor
1w
Save
Mark Applied
Hide
Staff Software Engineer, Quality & Reliability
Los Angeles or San Francisco or Toronto or Raleigh or United States
$172k-$229k/yr HybridFull Time
BuildOps
BuildOps: SaaS platform for managing commercial contracting businesses.
Extensive experience solving cross-cutting reliability and quality problems, leading multi-team initiatives, systems thinking, cloud (AWS) experience, strong programming in TypeScript or Java, observability and CI/CD familiarity, and strong communication.
TypeScript, Java, AWS, CI/CD
1mo
Save
Mark Applied
Hide
Associate Software Engineer, Reliability
Irvine, California, United States
$31-$56/hr OnsiteFull Time
Blizzard Entertainment
Blizzard EntertainmentNASDAQ: MSFT: Develops and publishes interactive video games and services.
2+ YOE2+ years relevant experience with Linux, DevOps practices, Python and/or C#, cloud APIs, OS/networking knowledge, strong diagnostic and communication skills, and willingness to share on-call duties.
Python, C#, Go, C++, Linux
2d
Save
Mark Applied
Hide
Software Development Engineer – Performance & Reliability
Lake Forest, California, United States
$92k-$154k/yr HybridFull Time
AVEVA
AVEVA: Industrial software for engineering and operational performance management.
6+ YOERequires 6+ years in software development in test, TypeScript/JavaScript and k6 expertise, distributed architecture testing, CI/CD pipelines, API testing, cloud platforms, OAuth 2.0/OIDC, and identity management.
TypeScript, JavaScript, k6, Azure DevOps, GitHub Actions, REST, HTTP/2, gRPC, Azure, OAuth 2.0, OIDC, Grafana, Application Insights, Open Telemetry, LLM
1d
Save
Mark Applied
Hide
Manager - Production Operations & Site Reliability Engineering
Lake Forest, California, United States
$140k-$182k/yr OnsiteFull Time
Alcon
AlconNYSE: ALC: Manufactures ophthalmic surgical equipment and vision care products.
5+ YOEBachelor’s degree or equivalent experience, 5 years of relevant experience, English fluency, and expertise in production operations, SRE, cloud platforms, AWS, Kubernetes, automation, observability, and regulated healthcare systems.
AWS, EKS, EC2, RDS, S3, ElastiCache, AWS MQ, Route53, Kubernetes, Istio, Datadog, CloudWatch, Docker, APM, CI/CD, Infrastructure as Code, HL7, FHIR, DICOM
3mo
Save
Mark Applied
Hide
Senior Platform Engineer
Santa Monica or Lower Manhattan or San Francisco or Los Angeles
$150k-$200k/yr HybridFull Time
Pivotal Health
Pivotal Health: AI platform automating healthcare insurance claim disputes for providers.
5+ YOE5+ years in platform, infrastructure, or software engineering; strong Python; cloud-native systems (GCP); Terraform; CI/CD; containers; event-driven architectures; security and reliability.
Python, Terraform, GitHub Actions, Kubernetes, Docker, Kafka, Pub/Sub, Kinesis, Google Cloud Platform