90 infrastructure reliability engineer jobs at 77 companies in North Babylon, NY
1mo
Save
Mark Applied
Hide
1mo
Customer Reliability Engineer - Infrastructure
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yrRemoteFull Time
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
Senior Site Reliability and Infrastructure Engineer
New York City, New York, United States
$160k-$220k/yrHybridFull Time
Treeswift: Provides AI-augmented vegetation management and asset monitoring for utilities.
7+ YOE7+ years experience in observability, SRE, infrastructure or DevOps; hands-on Terraform, Kubernetes, Linux, CI/CD; experience with Airflow-style pipelines and cloud services; strong debugging and communication skills.
Austin or New York City or San Francisco or Seattle
$203k-$232k/yrOnsiteFull Time
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience in reliability engineering for infrastructure or complex hardware, building availability/RAM models, leading cross-discipline FMEAs, and mining field failure data.
Site Reliability Engineer, Tech Infrastructure - USDS
New York, New York, United States
$137k-$259k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOE3+ years SRE/systems engineering experience, bachelor\u0002s in CS or related, proficiency in Python/Go/Java/Shell, Linux, cloud and distributed systems, monitoring and incident management.
Bayview Asset Management: Investment management firm specializing in credit and mortgage assets.
5+ YOE2+ Mgmt5+ years building and operating production infrastructure; strong cloud and DevOps experience; hands-on leadership; expertise in IaC and platform reliability.
Senior Site Reliability Engineer, Robotics & Cloud Infrastructure
Brooklyn or New York City or Richmond or Europe
$164k-$220k/yrRemoteFull Time
Bedrock Ocean Exploration: Maps the ocean floor using autonomous underwater robotic vehicles.
5+ YOE5+ years SRE/DevOps experience with on-call ownership; strong automation using Python/Go/Bash; Terraform and AWS hands-on; containerization (Docker, Kubernetes); observability (Prometheus, Grafana); Linux and networking expertise; East Coast location and US work authorization required.
Canada or United States or Seattle or Paris or New York City
$238k-$382k/yrRemoteFull Time
Docker: Provides a platform for building, sharing, and running containerized applications.
8+ YOE8+ years backend/infrastructure engineering; strong Go; experience with Kubernetes/EKS, cloud platforms, networking, and reliability; Bachelor's or equivalent; experience leading cross-team technical initiatives and strong written/verbal communication.
Narmi: Provides digital banking and account opening software for financial institutions.
6+ YOE6+ Mgmt6+ years in infrastructure, DevOps, or SRE roles, including engineering leadership; AWS, networking, security, Python, Bash, Linux, system architecture, debugging, and automation skills required.
Claryo: AI-powered spatial software for optimizing warehouse operations
3+ YOE3+ years SRE/infrastructure experience, strong Linux and networking fundamentals, experience with Kubernetes, cloud platforms, observability tooling, and debugging distributed systems in production.
Retool: Software platform for building custom internal business applications.
Experience operating production infrastructure (AWS), Kubernetes, Terraform, Postgres; programming in Go/Python/TypeScript/Java/Ruby; building observability and automation for customer-facing SaaS systems.
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
Requires a computer science bachelor's degree or equivalent experience, Python, production ML inference, AWS cloud infrastructure, Kubernetes, vulnerability management, distributed-systems debugging, and on-call participation.
Cox Enterprises: Providing global communications, automotive services, and media solutions.
5+ YOE5+ years in software/platform/infrastructure engineering, strong coding (Python/Go/Java), AWS and Terraform experience, SRE and observability knowledge, system design and incident response skills.
Senior Site Reliability Engineer, Messaging Services
Secaucus or New York or New Jersey or United States
$140k-$150k/yrRemoteFull Time
NBA: Operates professional basketball leagues and manages global media rights.
10+ YOEBachelor's in computer science or related, 10+ years in enterprise messaging/infrastructure/reliability, deep Exchange Online/M365 and email security experience, PowerShell and Microsoft Graph automation, Slack/Teams support, incident response and executive support.
Microsoft Exchange Online (M365), Microsoft Outlook, SMTP, Proofpoint, PowerShell, Microsoft Graph, Slack, Microsoft Teams, DMARC, DKIM, SPF, ARC
Optimal Market Technologies: Operates a broker-dealer platform for wholesale options execution.
Hands-on Linux and network administration, strong scripting/automation, Infrastructure-as-Code experience, production support and incident response ownership, staff management/mentoring, familiarity with C++, Python, SQL, Azure, and real-time trading systems.
C++, Python, SQL, Linux, CentOS7, RHEL9, Microsoft Azure, Microsoft Azure Virtual Desktop (AVD), Claude Code, FIX Protocol, PostgreSQL, AERON
Zurich or United Kingdom or Germany or San Francisco or New York
HybridFull Time
Namespace: Cloud infrastructure platform for faster software builds and tests.
Experience engineering large-scale infrastructure; strong network architecture and performance tuning; hands-on hardware deployment; proficient with orchestration and automation tools; track record of reliable systems and efficiency gains.
Sesame: Designing wearable computers with lifelike voice-driven AI agents.
3+ YOEStrong systems thinker with reliability engineering experience; 3+ years in infrastructure, platform, or ML systems; Kubernetes production experience; strong communication.