130 site reliability manager jobs at 59 companies in Darien, CT
1mo
Save
Mark Applied
Hide
1mo
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
Hudson River Trading: Proprietary quantitative trading firm specializing in automated market making.
5+ YOERequires 5+ years in site reliability or related disciplines, Python proficiency, container infrastructure experience, CI/CD, IaC, configuration management, and cloud platform experience. Technical leadership is preferred.
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied SRE experience, proficiency in reliability/scalability/security, experience with CI/CD, containers, observability, programming in Python/Java/.Net, and using enterprise AI for SRE workflows.
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
Requires a computer science bachelor's degree or equivalent experience, Python, production ML inference, AWS cloud infrastructure, Kubernetes, vulnerability management, distributed-systems debugging, and on-call participation.
Brooklyn or New York City or Los Angeles or Santa Monica or United States
$230k-$260k/yrHybridFull Time
Pivotal Health: AI platform automating healthcare insurance claim disputes for providers.
8+ YOE8+ years in SRE, infrastructure, platform engineering, or large-scale production systems; expertise in cloud infrastructure, distributed systems, networking, containers, orchestration, infrastructure as code, observability, automation, and incident management.
Ripple: Provides blockchain solutions for global payments and liquidity.
7+ YOE7+ years SRE/Platform experience focused on observability, New Relic, Terraform, PowerShell, Azure/AWS, incident management (Incident.IO/PagerDuty/OpsGenie), and coaching engineering teams.
New Relic, NRQL, Terraform, PowerShell, Azure, AWS, Azure DevOps, Octopus Deploy, Incident.IO, PagerDuty, OpsGenie, Slack, Python, Bash, Jira, SQL Server
Optimal Market Technologies: Operates a broker-dealer platform for wholesale options execution.
Hands-on Linux and network administration, strong scripting/automation, Infrastructure-as-Code experience, production support and incident response ownership, staff management/mentoring, familiarity with C++, Python, SQL, Azure, and real-time trading systems.
C++, Python, SQL, Linux, CentOS7, RHEL9, Microsoft Azure, Microsoft Azure Virtual Desktop (AVD), Claude Code, FIX Protocol, PostgreSQL, AERON
Scottsdale or Chicago or San Francisco or New York City or Phoenix or California or Illinois
$106k-$156k/yrHybridFull Time
Early Warning Services: Operates payment and risk solutions for the financial industry.
3+ YOEBachelor's degree in business, computer science, or related field; 3+ years of related technical or software development experience; Linux administration, Git, scripting, observability, incident management, and enterprise-scale experience required.
NBCUniversalNASDAQ: CMCSA: Produces and distributes entertainment content and theme park experiences.
8+ YOEBachelor's degree or equivalent experience, 8 years of engineering experience in broadcast playout, Linux administration, cloud and networking expertise, monitoring, containerization, and 24/7 on-call availability.
New York City or Europe or United States or Asia-Pacific
$170k-$190k/yrHybridFull Time
Pico: Provides managed infrastructure and data services to financial markets.
Bachelor's degree or higher in engineering or related discipline, operational team leadership experience, Linux performance expertise, networking knowledge, financial technology experience, and programming or scripting skills.
Tel Aviv-Yafo or Israel or New York City or London or Edinburgh or Brazil or Estonia or Ukraine
OnsiteFull Time
Optimove: AI-powered platform for customer marketing and retention.
5+ YOE5+ years in SRE, platform, DevOps, or infrastructure engineering; Kubernetes, GCP/AWS, programming, automation, CI/CD, observability, Linux, networking, distributed systems, and strong communication skills.
Site Reliability Engineer, Tech Infrastructure - USDS
New York, New York, United States
$137k-$259k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOE3+ years SRE/systems engineering experience, bachelor\u0002s in CS or related, proficiency in Python/Go/Java/Shell, Linux, cloud and distributed systems, monitoring and incident management.
8+ YOE8+ years managing production systems; deep hands-on AWS and Kubernetes experience; strong programming/scripting and automation skills; observability platform ownership; incident response leadership and mentoring ability.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
10+ YOE10+ years building distributed systems, 5+ years developing SaaS microservices, expert programming skills, distributed-systems expertise, architectural leadership, and a Computer Science degree or equivalent experience.
Site Reliability Engineer, Global Banking & Markets, Vice President
New York City, New York, United States
$150k-$250k/yrOnsiteFull Time
Goldman SachsNYSE: GS: Global investment banking, securities, and investment management firm.
8+ YOE8+ years of software or reliability engineering experience; proficiency in a major programming language; cloud, distributed systems, SRE, automation, observability, incident response, and risk management expertise.
BTIG: Provides global institutional trading and investment banking services.
2+ YOE2+ years IT support experience, strong customer-facing skills, Windows 10/11 and Microsoft 365 proficiency, experience with SCCM/Endpoint Manager and ticketing systems (ServiceNow); willing to obtain MS900.
ServiceNow, Windows 10, Windows 11, Microsoft Office 365, Microsoft OneDrive, System Center Configuration Manager, Endpoint Manager, Zoom, Bloomberg, Thomson Reuters, ICE, Fidessa, Redi+, Global Relay, Cisco, Active Directory, VPNWIFI