20 service reliability engineer jobs at 17 companies in Cold Spring, NY
5d
Save
Mark Applied
Hide
5d
Service Reliability Engineer
London or Manchester or New York City
HybridFull Time
Fitch Group: Provides global credit ratings and financial market research services.
Deep SRE, DevOps, or platform engineering experience with AWS, Azure, Docker, Kubernetes, Linux, Windows, CI/CD, cloud security, networking, and Python, PowerShell, or Bash.
Site Reliability Engineer - Application Support (Director)
New York, New York, United States
$120k-$165k/yrOnsiteFull Time
Morgan StanleyNYSE: MS: Global financial services firm providing investment and wealth management.
5+ YOE3+ Mgmt5+ years supporting or developing enterprise applications, 3+ years leading teams, DevOps/SRE experience, observability tools, AWS services, scripting in Python, knowledge of Java and database engineering, on-call support experience.
Senior Site Reliability Engineer, Messaging Services
Secaucus or New York or New Jersey or United States
$140k-$150k/yrRemoteFull Time
NBA: Operates professional basketball leagues and manages global media rights.
10+ YOEBachelor's in computer science or related, 10+ years in enterprise messaging/infrastructure/reliability, deep Exchange Online/M365 and email security experience, PowerShell and Microsoft Graph automation, Slack/Teams support, incident response and executive support.
Microsoft Exchange Online (M365), Microsoft Outlook, SMTP, Proofpoint, PowerShell, Microsoft Graph, Slack, Microsoft Teams, DMARC, DKIM, SPF, ARC
Senior Site Reliability and Infrastructure Engineer
New York City, New York, United States
$160k-$220k/yrHybridFull Time
Treeswift: Provides AI-augmented vegetation management and asset monitoring for utilities.
7+ YOE7+ years experience in observability, SRE, infrastructure or DevOps; hands-on Terraform, Kubernetes, Linux, CI/CD; experience with Airflow-style pipelines and cloud services; strong debugging and communication skills.
Senior Software Engineer - Observability and Reliability
New York City or San Francisco or London or Sydney
$170k-$240k/yrOnsiteFull Time
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
5+ YOE5+ years building high-quality software, strong CS fundamentals, experience building observability tools, proficiency with Go, OpenTelemetry, Kubernetes, participation in on-call rotations, cloud service administration (GCP/AWS/Azure) preferred.
Executive Director – Site Reliability Engineering – WM Technology
New York, New York, United States
$195k-$215k/yrOnsiteFull Time
Morgan StanleyNYSE: MS: Provides global investment banking, wealth management, and advisory services.
12+ YOEBachelor's degree; 12-18+ years in technology production management in financial services; strong SRE leadership, incident management, monitoring/observability, and cloud/DevOps familiarity.
ID.me: Provides secure digital identity verification and authentication services.
4+ YOEBachelor's degree or equivalent, 4+ years backend software experience (Java/JVM preferred), strong SQL/PostgreSQL skills, REST/OpenAPI design experience, familiarity with AI-assisted development tooling, and production reliability practices.
Bastion: Infrastructure for enterprises to issue and manage regulated stablecoins.
Senior-level backend engineer experienced with Go and TypeScript/Node.js (Rust optional), building reliable services with gRPC/REST, Postgres/Redis/Kafka, Temporal, and AWS/Kubernetes/Terraform; strong CI/CD and observability skills.
Pivotal Health: AI platform automating healthcare insurance claim disputes for providers.
5+ YOE5+ years backend software engineering experience; strong system design, data modeling, API development, and service reliability skills; Python experience preferred; authorized to work in the U.S. without sponsorship.
Senior Lead Software Engineer-AI Foundation Services
Plano or Jersey City or Wilmington or McLean
$171k-$260k/yrOnsiteFull Time
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years software engineering experience building cloud-native AI/ML platform services with Kubernetes, CI/CD, and infrastructure-as-code; proficiency in Python/Java/Go; strong production reliability and secure-by-design practices.
Astrophysics Inc.: Designs and manufactures X-ray security inspection systems.
2+ YOEAssociate's or Bachelor's in Electronics/Computer Engineering, 2+ years field service/technical support experience, travel required, reliable vehicle and valid driver\u0002s license, strong troubleshooting and customer service skills, fluent English.
Senior Production Engineer – Listed Derivatives Trading
New York City or Jacksonville
$96k-$167k/yrOnsiteFull Time
CGINYSE: GIB: Provides information technology and business consulting services.
Production support expertise in incident management, root cause analysis, service reliability, monitoring, automation, and SLA compliance within trading or financial services; listed derivatives knowledge required.
Zocdoc: Online platform for searching and booking healthcare appointments.
2+ YOE2+ years experience with .NET and AWS; ability to integrate generative AI tools; experience building/scaling backend services; strong mentorship, communication, and a bias for correctness and reliability.
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Manage site reliability and ML operations for Apple's advertising platform; role emphasizes building and operating reliable ad services across Apple products.
Experience with elementary-age children, up-to-date First-Aid/CPR, strong communication, ability to motivate, creative and flexible, reliable and accountable.
BarclaysLondon Stock Exchange: BARC: Global bank providing retail, corporate, and investment financial services.
Requires extensive service management, reliability engineering, resilience, technology controls, cybersecurity, risk management, incident management, stakeholder leadership, and team development expertise.