32 platform reliability engineer jobs at 20 companies in Cudahy, CA
3w
Save
Mark Applied
Hide
3w
Staff AI Platform & Reliability Engineer
Santa Ana, California, United States
$172k-$220k/yrOnsiteFull Time
eJam: Builds and scales direct-to-consumer e-commerce brands.
Senior Python/GCP engineer to design and operate AI generation services, provider integrations, billing/usage metering, tenant security and platform reliability.
Site Reliability Engineer, Kubernetes Platform (Starshield)
Hawthorne or Redmond
$125k-$175k/yrOnsiteFull Time
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
1+ YOEBachelor's in CS/IT/engineering + 1+ year SRE/DevOps experience (or 3+ years experience), Linux, Terraform/Ansible, Kubernetes and OCI containers, Bash/Python scripting, development in Python/C++/Go, willingness to obtain Top Secret clearance.
Green DotNYSE: GDOT: Provides mobile banking and payment solutions to consumers and businesses.
7+ YOE7+ years in release/reliability engineering, cloud platform experience (AWS/Azure/GCP), automated deployment and observability proficiency, scripting with PowerShell/Bash/Python, excellent troubleshooting and communication skills.
Brooklyn or New York City or Los Angeles or Santa Monica or United States
$230k-$260k/yrHybridFull Time
Pivotal Health: AI platform automating healthcare insurance claim disputes for providers.
8+ YOE8+ years in SRE, infrastructure, platform engineering, or large-scale production systems; expertise in cloud infrastructure, distributed systems, networking, containers, orchestration, infrastructure as code, observability, automation, and incident management.
AbbottNYSE: ABT: Provides medical devices, diagnostics, and science-based nutritional products.
Senior SRE with strong distributed systems, cloud (Azure), Kubernetes, observability, automation, incident management, and cross-functional communication skills for a medical device remote monitoring platform.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK, EFK, Datadog, Linux
United States or North America or San Francisco or Los Angeles or New York City or Washington or London or Singapore
RemoteFull Time
TRM Labs: An AI-powered intelligence technology helping agencies investigate crime and disrupt illicit activity.
U.S. citizenship, distributed OLAP or serving-layer operations experience, query tuning, data pipeline reliability, incident response, AI tool fluency, independent infrastructure ownership, and on-call readiness.
Anduril Industries: Defense technology building autonomous military hardware and software.
10+ YOE10+ years in SRE, production, or infrastructure engineering; expertise in distributed systems, Kubernetes, cloud platforms, networking, storage, observability, deployments, incident response, and systems programming.
Principal Engineer, Digital Workplace Technology Engineering
New York or New Jersey or Los Angeles or United States
$180k-$210k/yrRemoteFull Time
NBCUniversalNASDAQ: CMCSA: Produces and distributes entertainment content and theme park experiences.
12+ YOE12+ years in enterprise/platform engineering; deep expertise in Microsoft 365, endpoint management (Intune, JAMF, SCCM), Azure, enterprise AI (Copilot, Power Platform), architecture, SRE principles, and large-scale platform transformations.
Microsoft 365, Microsoft Teams, Microsoft SharePoint, Microsoft OneDrive, Microsoft Exchange Online, Microsoft Entra ID, Microsoft Purview, Microsoft Defender, Microsoft Intune, JAMF, Microsoft SCCM, Microsoft Copilot, Copilot Studio, Microsoft Power Platform, Microsoft Power Apps, Microsoft Power Automate, Microsoft Azure, MLOps, DevOps
WME Group: Global holding for talent, media, and entertainment representation.
8+ YOE2+ Mgmt8+ years in infrastructure/DevOps/ SRE or platform engineering; 2+ years leading engineers; strong Terraform/OpenTofu, CI/CD (GitHub Actions/Spacelift); DevSecOps practices; platform governance and reliability.
Software Development Engineer – Performance & Reliability
Lake Forest, California, United States
$92k-$154k/yrHybridFull Time
AVEVA: Industrial software for engineering and operational performance management.
6+ YOERequires 6+ years in software development in test, TypeScript/JavaScript and k6 expertise, distributed architecture testing, CI/CD pipelines, API testing, cloud platforms, OAuth 2.0/OIDC, and identity management.
Manager - Production Operations & Site Reliability Engineering
Lake Forest, California, United States
$140k-$182k/yrOnsiteFull Time
AlconNYSE: ALC: Manufactures ophthalmic surgical equipment and vision care products.
5+ YOEBachelor’s degree or equivalent experience, 5 years of relevant experience, English fluency, and expertise in production operations, SRE, cloud platforms, AWS, Kubernetes, automation, observability, and regulated healthcare systems.
Relativity Space: Designing and manufacturing 3D-printed rockets and launch vehicles.
5+ YOE5+ years kernel/driver development experience (PCI/PCIe, block storage), strong Linux internals and storage systems (ZFS/OpenZFS, NVMe, NFS), experience with fault injection and reliability modeling, hardware lab prototyping.
AXS: Provides digital ticketing and marketing solutions for live events.
13+ YOE7+ MgmtRequires 13+ years in high-growth technology and 7+ years leading data engineering or platform teams, with expertise in scalable data platforms, cloud systems, governance, reliability, and global team leadership.
Hyderabad or Pasay City or Los Angeles or World Technology Center
OnsiteFull Time
VXI Global Solutions: Global provider of customer care and business process outsourcing.
Requires observability experience with Prometheus, Grafana, OpenTelemetry, Google Cloud tools, and SolarWinds; telemetry analysis, incident troubleshooting, Python or Bash scripting, and SRE or platform engineering experience preferred.
Prometheus, Grafana, OpenTelemetry, SolarWinds, Google Cloud Platform, Cloud Monitoring, Logging, Trace, Python, Bash, Infrastructure-as-Code, Datadog, New Relic
Varsity Technologies: Managed IT services and AI solutions for social-impact organizations.
3+ YOERequires 3–5 years of IT support experience, advanced Microsoft 365 and macOS support, networking, AI platform, RMM/PSA, client communication, driver's license, and reliable transportation.
Microsoft 365, Exchange, Microsoft SharePoint, Microsoft Teams, Entra ID, Active Directory (AAD), macOS, MDM, Apple Business Manager, iOS, Apple Cloud, DNS, DHCP, VPN, Microsoft Copilot, Google Gemini, ChatGPT, Claude, RMM, PSA, CompTIA Network+, CompTIA Security+, Microsoft 365 Administrator (MS-102)
Los Angeles or Santa Ana or New Jersey or Texas or Florida or Shanghai or Hong Kong or Japan or Canada or Mexico or Germany or France or United States
$165k-$250k/yrRemoteFull Time
Collectors: Providing authentication and grading services for high-value collectibles.
8+ YOE8+ years building data platforms with BigQuery/dbt and modern ingestion tools, expert SQL and Python, cloud experience (GCP/AWS), AI-assisted development (Claude Code), and strong reliability and governance practice.
Hevo, Estuary, Google BigQuery, dbt, Claude Code, Python, SQL, GCP, AWS, CI/CD
NetflixNASDAQ: NFLX: Provider of global streaming entertainment and video content.
3+ Mgmt3+ years managing distributed engineering teams, platform mindset, experience with reliable 24x7 services, strong communication, hiring/coaching, and ability to drive adoption of data platforms and governance.
graph-based data modeling, knowledge graphs, RDF, ontologies
Columbus or Boston or New York City or Chicago or Austin or Los Angeles
$150k-$190k/yrRemoteFull Time
Loop Returns: Software platform automating e-commerce returns and post-purchase experiences.
3+ YOE3+ years engineering management experience, experience with platform reliability/system health, AI agent adoption, technical depth in high-risk domains, ability to manage and grow engineers.