33 platform reliability engineer jobs at 22 companies in Los Angeles, CA

2w
Save
Mark Applied
Hide
Staff AI Platform & Reliability Engineer
Santa Ana, California, United States
$172k-$220k/yr OnsiteFull Time
eJam
eJam: Builds and scales direct-to-consumer e-commerce brands.
Senior Python/GCP engineer to design and operate AI generation services, provider integrations, billing/usage metering, tenant security and platform reliability.
Python 3.12, FastAPI, Pydantic, asyncio, Cloud Run, Pub/Sub, Cloud Tasks, GCS, Firestore, Cloud SQL/Postgres, SQLAlchemy, Alembic, Terraform, IAM/OIDC, Secret Manager, CI/CD, NestJS, TypeScript, Claude Code, Codex, Cursor
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Kubernetes Platform (Starshield)
Hawthorne or Redmond
$125k-$175k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
1+ YOEBachelor's in CS/IT/engineering + 1+ year SRE/DevOps experience (or 3+ years experience), Linux, Terraform/Ansible, Kubernetes and OCI containers, Bash/Python scripting, development in Python/C++/Go, willingness to obtain Top Secret clearance.
Terraform, Ansible, OCI containers, Kubernetes, Bash, Python, C++, Go, Bazel, Makefiles, TCP/IP, Linux
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Los Angeles, California, United States
$140k-$199k/yr HybridFull Time
Green Dot
Green DotNYSE: GDOT: Provides mobile banking and payment solutions to consumers and businesses.
7+ YOE7+ years in release/reliability engineering, cloud platform experience (AWS/Azure/GCP), automated deployment and observability proficiency, scripting with PowerShell/Bash/Python, excellent troubleshooting and communication skills.
AWS, Azure, GCP, PowerShell, Bash, Python
2mo
Save
Mark Applied
Hide
Site Reliability Engineer, Client Platform
Los Angeles, California, United States
$164k-$270k/yr OnsiteFull Time
Hadrian
Hadrian: Building autonomous factories for aerospace and defense manufacturing.
Experience building scalable automation, strong scripting (Python, Bash, PowerShell), IaC (Ansible, Terraform), MDM platform administration, patch/vulnerability remediation, endpoint hardening, and translating CMMC/compliance into code-managed baselines. ITAR eligibility required.
Python, Bash, PowerShell, Ansible, Terraform, Chef, Salt, Puppet, Pulumi, Fleet DM, Microsoft Intune, Workspace ONE, Jamf, osquery
4w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Sunnyvale or Sylmar
$90k-$180k/yr OnsiteFull Time
Abbott
AbbottNYSE: ABT: Provides medical devices, diagnostics, and science-based nutritional products.
Senior SRE with strong distributed systems, cloud (Azure), Kubernetes, observability, automation, incident management, and cross-functional communication skills for a medical device remote monitoring platform.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK, EFK, Datadog, Linux
3mo
Save
Mark Applied
Hide
Senior Platform Engineer
Santa Monica or Lower Manhattan or San Francisco or Los Angeles
$150k-$200k/yr HybridFull Time
Pivotal Health
Pivotal Health: AI platform automating healthcare insurance claim disputes for providers.
5+ YOE5+ years in platform, infrastructure, or software engineering; strong Python; cloud-native systems (GCP); Terraform; CI/CD; containers; event-driven architectures; security and reliability.
Python, Terraform, GitHub Actions, Kubernetes, Docker, Kafka, Pub/Sub, Kinesis, Google Cloud Platform
1d
Save
Mark Applied
Hide
Data Platform Engineer - AI Platform
United States or North America or San Francisco or Los Angeles or New York City or Washington or London or Singapore
RemoteFull Time
TRM Labs
TRM Labs: An AI-powered intelligence technology helping agencies investigate crime and disrupt illicit activity.
U.S. citizenship, distributed OLAP or serving-layer operations experience, query tuning, data pipeline reliability, incident response, AI tool fluency, independent infrastructure ownership, and on-call readiness.
StarRocks, Claude, Trino, ClickHouse, Cursor, Slack, Otter.ai, Fireflies, Fathom, Cluey
2d
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Costa Mesa, California, United States
$191k-$253k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology building autonomous military hardware and software.
10+ YOE10+ years in SRE, production, or infrastructure engineering; expertise in distributed systems, Kubernetes, cloud platforms, networking, storage, observability, deployments, incident response, and systems programming.
Kubernetes, AWS, GCP, Azure, Go, Python, Rust, Datadog, Grafana, Prometheus, OpenTelemetry
1mo
Save
Mark Applied
Hide
Manager, Software Engineering (Reliability Platform)
United States or California or Washington or New York or New Jersey or Connecticut or Los Angeles or San Francisco
$204k-$290k/yr RemoteFull Time
Affirm
AffirmNasdaq: AFRM: Financial platform providing installment loans for consumer purchases.
7+ YOE2+ Mgmt7+ years backend/full-stack engineering experience with 2+ years engineering leadership; SRE/production engineering experience; observability and platform-building experience; strong programming (Python, Kotlin, Java); Bachelor\u0002s degree or equivalent experience.
Python, Kotlin, Java
1mo
Save
Mark Applied
Hide
Principal Engineer, Digital Workplace Technology Engineering
New York or New Jersey or Los Angeles or United States
$180k-$210k/yr RemoteFull Time
NBCUniversal
NBCUniversalNASDAQ: CMCSA: Produces and distributes entertainment content and theme park experiences.
12+ YOE12+ years in enterprise/platform engineering; deep expertise in Microsoft 365, endpoint management (Intune, JAMF, SCCM), Azure, enterprise AI (Copilot, Power Platform), architecture, SRE principles, and large-scale platform transformations.
Microsoft 365, Microsoft Teams, Microsoft SharePoint, Microsoft OneDrive, Microsoft Exchange Online, Microsoft Entra ID, Microsoft Purview, Microsoft Defender, Microsoft Intune, JAMF, Microsoft SCCM, Microsoft Copilot, Copilot Studio, Microsoft Power Platform, Microsoft Power Apps, Microsoft Power Automate, Microsoft Azure, MLOps, DevOps
3mo
Save
Mark Applied
Hide
Infrastructure Platform Manager
Beverly Hills, California, United States
$120k-$160k/yr OnsiteFull Time
WME Group
WME Group: Global holding for talent, media, and entertainment representation.
8+ YOE2+ Mgmt8+ years in infrastructure/DevOps/ SRE or platform engineering; 2+ years leading engineers; strong Terraform/OpenTofu, CI/CD (GitHub Actions/Spacelift); DevSecOps practices; platform governance and reliability.
Terraform/OpenTofu, GitHub Actions, Spacelift, policy-as-code, secrets management, IaC scanning, SAST/DAST
2d
Save
Mark Applied
Hide
Release Engineer
Mexico City or Los Angeles or United States or Dominican Republic or Mexico or Ukraine
$1059k-$1324k/yr RemoteFull Time
SimplePractice
SimplePractice: Practice management software for health and wellness professionals.
3+ YOERequires 3+ years in Release Engineering, DevOps, Platform Engineering, Site Reliability Engineering, or related work; CI/CD production SaaS experience, Git, Linux, scripting, troubleshooting, and collaboration skills.
SemaphoreCI, Ruby, Ruby on Rails, Git, AWS, Linux, Terraform, Docker, Kubernetes, Bash, Python, Datadog, CloudWatch
2d
Save
Mark Applied
Hide
Software Development Engineer – Performance & Reliability
Lake Forest, California, United States
$92k-$154k/yr HybridFull Time
AVEVA
AVEVA: Industrial software for engineering and operational performance management.
6+ YOERequires 6+ years in software development in test, TypeScript/JavaScript and k6 expertise, distributed architecture testing, CI/CD pipelines, API testing, cloud platforms, OAuth 2.0/OIDC, and identity management.
TypeScript, JavaScript, k6, Azure DevOps, GitHub Actions, REST, HTTP/2, gRPC, Azure, OAuth 2.0, OIDC, Grafana, Application Insights, Open Telemetry, LLM
1d
Save
Mark Applied
Hide
Manager - Production Operations & Site Reliability Engineering
Lake Forest, California, United States
$140k-$182k/yr OnsiteFull Time
Alcon
AlconNYSE: ALC: Manufactures ophthalmic surgical equipment and vision care products.
5+ YOEBachelor’s degree or equivalent experience, 5 years of relevant experience, English fluency, and expertise in production operations, SRE, cloud platforms, AWS, Kubernetes, automation, observability, and regulated healthcare systems.
AWS, EKS, EC2, RDS, S3, ElastiCache, AWS MQ, Route53, Kubernetes, Istio, Datadog, CloudWatch, Docker, APM, CI/CD, Infrastructure as Code, HL7, FHIR, DICOM
2mo
Save
Mark Applied
Hide
Staff Software Engineer, Storage Platform
Long Beach, California, United States
$181k-$249k/yr OnsiteFull Time
Relativity Space
Relativity Space: Designing and manufacturing 3D-printed rockets and launch vehicles.
5+ YOE5+ years kernel/driver development experience (PCI/PCIe, block storage), strong Linux internals and storage systems (ZFS/OpenZFS, NVMe, NFS), experience with fault injection and reliability modeling, hardware lab prototyping.
Linux, OpenZFS, ZFS, NFS, NVMe, PCI, PCIe, Yocto, Buildroot, ftrace, perf, bpftrace, kdump, vmcore, serial console, logic analyzer
1w
Save
Mark Applied
Hide
VP, Data Platform Engineering
Los Angeles, California, United States
$330k-$360k/yr HybridFull Time
AXS
AXS: Provides digital ticketing and marketing solutions for live events.
13+ YOE7+ MgmtRequires 13+ years in high-growth technology and 7+ years leading data engineering or platform teams, with expertise in scalable data platforms, cloud systems, governance, reliability, and global team leadership.
Snowflake, Amazon Redshift, BigQuery, Airflow, dbt, Kafka, Amazon Kinesis, AWS, Oracle, DynamoDB, Scrum, Kanban, CI/CD, SRE
1w
Save
Mark Applied
Hide
Lead Machine Learning Operations Engineer
Burbank, California, United States
$157k-$235k/yr OnsiteFull Time
Paramount
ParamountNASDAQ: PSKY: Global media producing films, television, and streaming content.
5+ YOERequires 5+ years in ML engineering, MLOps, ML platforms, applied ML, data platforms, or reliability engineering; production ML operations, cross-team technical leadership, SQL, and ML observability experience.
SQL
4d
Save
Mark Applied
Hide
SRE
Hyderabad or Pasay City or Los Angeles or World Technology Center
OnsiteFull Time
VXI Global Solutions
VXI Global Solutions: Global provider of customer care and business process outsourcing.
Requires observability experience with Prometheus, Grafana, OpenTelemetry, Google Cloud tools, and SolarWinds; telemetry analysis, incident troubleshooting, Python or Bash scripting, and SRE or platform engineering experience preferred.
Prometheus, Grafana, OpenTelemetry, SolarWinds, Google Cloud Platform, Cloud Monitoring, Logging, Trace, Python, Bash, Infrastructure-as-Code, Datadog, New Relic
3w
Save
Mark Applied
Hide
Staff Data Engineer
Los Angeles or Santa Ana or New Jersey or Texas or Florida or Shanghai or Hong Kong or Japan or Canada or Mexico or Germany or France or United States
$165k-$250k/yr RemoteFull Time
Collectors
Collectors: Providing authentication and grading services for high-value collectibles.
8+ YOE8+ years building data platforms with BigQuery/dbt and modern ingestion tools, expert SQL and Python, cloud experience (GCP/AWS), AI-assisted development (Claude Code), and strong reliability and governance practice.
Hevo, Estuary, Google BigQuery, dbt, Claude Code, Python, SQL, GCP, AWS, CI/CD
1w
Save
Mark Applied
Hide
Senior Director, Software Engineering – Customer Content Platform
Eagan or Frisco or New York City or Toronto or San Francisco or Los Angeles or Irvine or McLean or Washington
$159k-$295k/yr HybridFull Time
Thomson Reuters
Thomson ReutersNASDAQ: TRI: Provides professional software, data, and news services globally.
12+ YOE5+ Mgmt12+ years of software engineering experience, including 5+ years leading engineering managers and distributed teams, with expertise in enterprise-scale data platforms, distributed systems, compliance, migrations, and reliability.
Electronic Medical Records (EMR), Microsoft Excel, iManage, NetDocuments, Microsoft SharePoint, HighQ, Box, Dropbox, Google Drive, MCP

Explore Jobs

Expand Your Job Search