130 site reliability manager jobs at 59 companies in Darien, CT

1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
3mo
Save
Mark Applied
Hide
Site Reliability Engineer - Enterprise Technology
New York City, New York, United States
$200k-$250k/yr OnsiteFull Time
Hudson River Trading
Hudson River Trading: Proprietary quantitative trading firm specializing in automated market making.
5+ YOERequires 5+ years in site reliability or related disciplines, Python proficiency, container infrastructure experience, CI/CD, IaC, configuration management, and cloud platform experience. Technical leadership is preferred.
Linux, Kubernetes, Python, Jenkins, GitHub Actions, ArgoCD, Terraform, SaltStack, Chef, Puppet, Ansible, AWS, Azure, GCP
4d
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
HybridFull Time
Chariot
Chariot: Payment infrastructure for charitable giving and Donor Advised Funds.
4+ YOERequires 4+ years software development, 2+ years backend application development, bachelor's degree preferred, and proficiency with Go, Node, Docker, Terraform, Kubernetes, Postgres, REST APIs, gRPC, and AWS.
Go, Node, Docker, Terraform, Kubernetes, Postgres, REST APIs, gRPC, AWS
3w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Plano or Jersey City
$152k-$215k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied SRE experience, proficiency in reliability/scalability/security, experience with CI/CD, containers, observability, programming in Python/Java/.Net, and using enterprise AI for SRE workflows.
Python, Java, Spring Boot, .Net, CI/CD
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York, New York, United States
$178k-$258k/yr RemoteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
Requires a computer science bachelor's degree or equivalent experience, Python, production ML inference, AWS cloud infrastructure, Kubernetes, vulnerability management, distributed-systems debugging, and on-call participation.
Python, PHP, Node.js, Ruby, SageMaker, OpenAI, Bedrock, EC2 Auto Scaling Groups, Kubernetes, AWS, Azure, GCP, AMI, LangGraph, LLM gateway, MCP, Aurora PostgreSQL, Memcached, Terraform, Terragrunt, Atlantis, Chef, Ansible, SSM, Docker, bash, Jenkins, Argo CD, New Relic, Splunk, Grafana, Prometheus, Fastly, Datadome, WAF, vLLM, LangSmith, Lambda
3d
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Brooklyn or New York City or Los Angeles or Santa Monica or United States
$230k-$260k/yr HybridFull Time
Pivotal Health
Pivotal Health: AI platform automating healthcare insurance claim disputes for providers.
8+ YOE8+ years in SRE, infrastructure, platform engineering, or large-scale production systems; expertise in cloud infrastructure, distributed systems, networking, containers, orchestration, infrastructure as code, observability, automation, and incident management.
AI, HIPAA, PHI, SOC 2, 401(k)
1mo
Save
Mark Applied
Hide
Staff Engineer, Site Reliability
New York City or Los Angeles or United States or Canada
RemoteFull Time
Babylist
Babylist: Operates an e-commerce platform and universal registry for baby products.
Hands-on Terraform expertise, proven AWS experience (EKS, RDS, networking, CDNs), production Kubernetes operation, CI/CD design, observability and alerting, on-call/incident management, cross-functional developer support, familiarity with AI tooling.
Ruby on Rails, AWS, Sidekiq, MySQL, Redis, Terraform, EKS, RDS, Kubernetes, CircleCI, GitHub Actions, Datadog, Sentry, PagerDuty, Cronitor, Claude, ChatGPT
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Observability
Chicago or New York City
$160k-$200k/yr HybridFull Time
Ripple
Ripple: Provides blockchain solutions for global payments and liquidity.
7+ YOE7+ years SRE/Platform experience focused on observability, New Relic, Terraform, PowerShell, Azure/AWS, incident management (Incident.IO/PagerDuty/OpsGenie), and coaching engineering teams.
New Relic, NRQL, Terraform, PowerShell, Azure, AWS, Azure DevOps, Octopus Deploy, Incident.IO, PagerDuty, OpsGenie, Slack, Python, Bash, Jira, SQL Server
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer (SRE)
Chicago or New York City
$175k-$220k/yr HybridFull Time
Optimal Market Technologies
Optimal Market Technologies: Operates a broker-dealer platform for wholesale options execution.
Hands-on Linux and network administration, strong scripting/automation, Infrastructure-as-Code experience, production support and incident response ownership, staff management/mentoring, familiarity with C++, Python, SQL, Azure, and real-time trading systems.
C++, Python, SQL, Linux, CentOS7, RHEL9, Microsoft Azure, Microsoft Azure Virtual Desktop (AVD), Claude Code, FIX Protocol, PostgreSQL, AERON
1w
Save
Mark Applied
Hide
Sr. Site Reliability Engineer - Paze
Scottsdale or Chicago or San Francisco or New York City or Phoenix or California or Illinois
$106k-$156k/yr HybridFull Time
Early Warning Services
Early Warning Services: Operates payment and risk solutions for the financial industry.
3+ YOEBachelor's degree in business, computer science, or related field; 3+ years of related technical or software development experience; Linux administration, Git, scripting, observability, incident management, and enterprise-scale experience required.
Linux, Git, Java, Ruby, Python, JavaScript, Go, AWS, Docker, Kubernetes, Swarm, CI/CD, TCP/UDP/IP
1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Playout
Stamford, Connecticut, United States
$145k-$175k/yr HybridFull Time
NBCUniversal
NBCUniversalNASDAQ: CMCSA: Produces and distributes entertainment content and theme park experiences.
8+ YOEBachelor's degree or equivalent experience, 8 years of engineering experience in broadcast playout, Linux administration, cloud and networking expertise, monitoring, containerization, and 24/7 on-call availability.
Linux, Splunk, Grafana, ServiceNow, Docker, Kubernetes, AWS, Snell, Harris, Imagine, Amagi, Evertz, GrassValley, Harmonic, CoralBay, Veset, TS, HEVC, H.264, HLS, CMAF, SCTE-35, SCTE-224, ESAM, SRT, RIST, Slack
3d
Save
Mark Applied
Hide
Site Reliability Engineer- Team Lead
New York City or Europe or United States or Asia-Pacific
$170k-$190k/yr HybridFull Time
Pico
Pico: Provides managed infrastructure and data services to financial markets.
Bachelor's degree or higher in engineering or related discipline, operational team leadership experience, Linux performance expertise, networking knowledge, financial technology experience, and programming or scripting skills.
Linux, Python, C, C++, Java
5d
Save
Mark Applied
Hide
Site Reliability Team Leader
Tel Aviv-Yafo or Israel or New York City or London or Edinburgh or Brazil or Estonia or Ukraine
OnsiteFull Time
Optimove
Optimove: AI-powered platform for customer marketing and retention.
5+ YOE5+ years in SRE, platform, DevOps, or infrastructure engineering; Kubernetes, GCP/AWS, programming, automation, CI/CD, observability, Linux, networking, distributed systems, and strong communication skills.
Kubernetes, GCP, AWS, Python, Go, Bash, CI/CD, Datadog, Prometheus, Grafana, Linux, Terraform, Ansible, Kafka, Pub/Sub, Redis, OpenTelemetry, Canary, Blue/Green, Feature Flags
3mo
Save
Mark Applied
Hide
Site Reliability Engineer, Storage - Enterprise Technology
New York City, New York, United States
$200k-$250k/yr OnsiteFull Time
Hudson River Trading
Hudson River Trading: A quantitative firm using technology to trade global financial markets.
5+ YOE5+ years SRE experience; Linux, Kubernetes, Python; NetApp storage; CI/CD tools; IaC and configuration management.
Linux, Kubernetes, Python, NetApp, NFS, SMB, Jenkins, GitHub Actions, ArgoCD, Terraform, SaltStack, Chef, Puppet, Ansible
2mo
Save
Mark Applied
Hide
Site Reliability Engineer, Tech Infrastructure - USDS
New York, New York, United States
$137k-$259k/yr OnsiteFull Time
TikTok USDS Joint Venture
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOE3+ years SRE/systems engineering experience, bachelor\u0002s in CS or related, proficiency in Python/Go/Java/Shell, Linux, cloud and distributed systems, monitoring and incident management.
Python, Go, Java, Shell, Linux, Docker, Kubernetes, Prometheus, Grafana
3w
Save
Mark Applied
Hide
Staff , Site Reliability Engineer - Cloud Platform
New York City, New York, United States
$190k-$210k/yr HybridFull Time
Butterfly Network
Butterfly NetworkNYSE: BFLY: Handheld whole-body ultrasound scanners powered by semiconductor technology.
8+ YOE8+ years managing production systems; deep hands-on AWS and Kubernetes experience; strong programming/scripting and automation skills; observability platform ownership; incident response leadership and mentoring ability.
Kubernetes, EKS, AWS, NewRelic, Datadog, DICOM, HL7, FHIR, PACS, VNA, EMR
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
2w
Save
Mark Applied
Hide
Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid)
New York City or Austin or Sunnyvale or Redmond
$140k-$215k/yr HybridFull Time
CrowdStrike
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
10+ YOE10+ years building distributed systems, 5+ years developing SaaS microservices, expert programming skills, distributed-systems expertise, architectural leadership, and a Computer Science degree or equivalent experience.
Go, Java, Scala, Kotlin, Python, Node.js, Kubernetes, AWS, Cassandra, Kafka, Elasticsearch, OpenSearch, Google Cloud Platform (GCP), Oracle Cloud Infrastructure (OCI), GitHub, Stack Overflow
1w
Save
Mark Applied
Hide
Site Reliability Engineer, Global Banking & Markets, Vice President
New York City, New York, United States
$150k-$250k/yr OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities, and investment management firm.
8+ YOE8+ years of software or reliability engineering experience; proficiency in a major programming language; cloud, distributed systems, SRE, automation, observability, incident response, and risk management expertise.
Java, GCP, AWS, Kubernetes, Docker, Claude Code, GitHub Copilot Agent Mode, Devin, Gemini Code Assist, Apache Kafka, Spring Boot, gRPC, Protocol Buffers, Apache Camel, Spring Integration, Terraform, Helm, Prometheus, Grafana, OpenTelemetry, SQL, NoSQL, Vert.x, Netty, Microsoft?
1mo
Save
Mark Applied
Hide
Technology, DevOps/Site Reliability Engineer
New York or San Francisco
$160k-$200k/yr HybridFull Time
BTIG
BTIG: Provides global institutional trading and investment banking services.
2+ YOE2+ years IT support experience, strong customer-facing skills, Windows 10/11 and Microsoft 365 proficiency, experience with SCCM/Endpoint Manager and ticketing systems (ServiceNow); willing to obtain MS900.
ServiceNow, Windows 10, Windows 11, Microsoft Office 365, Microsoft OneDrive, System Center Configuration Manager, Endpoint Manager, Zoom, Bloomberg, Thomson Reuters, ICE, Fidessa, Redi+, Global Relay, Cisco, Active Directory, VPNWIFI

Explore Jobs

Expand Your Job Search