130 systems reliability engineer jobs at 104 companies in Fairfield, CT
2mo
Save
Mark Applied
Hide
2mo
Systems Reliability Engineer
San Francisco or New York City
$150k-$170k/yrHybridFull Time
Claryo: AI-powered spatial software for optimizing warehouse operations
3+ YOE3+ years in SRE/infrastructure/distributed systems; strong Linux/networking; production experience; Kubernetes and cloud platforms; observability tools; multi-layer debugging; on-call readiness.
BAE Systems - Senior Reliability Engineer - C4ISR Systems
Greenlawn, New York, United States
$145k-$247k/yrHybridFull Time
BAE SystemsLondon Stock Exchange: BA: Designs and manufactures advanced defense and aerospace systems.
8+ YOEBachelor's in engineering and minimum 8 years reliability engineering experience; expertise in root cause analysis, FMEA, FRACAS, statistical modeling; proficiency with Microsoft Office; ability to obtain/maintain Secret clearance.
Microsoft Excel, Microsoft Word, Microsoft PowerPoint, Minitab, JMP, R, SAP, Teamcenter, Windchill, Python, MATLAB, FRACAS, SEM/EDS, X-ray, acoustic microscopy, thermal imaging
Two Sigma: Systematic investment management and quantitative trading firm.
1+ YOE1+ years reliability engineering experience (5+ preferred), BS/BA in Computer Science or technical discipline, proficiency in Python/Java/C++/Rust, experience with automation, UNIX/Linux and distributed systems preferred.
Beast Industries: Produces digital media and consumer goods for MrBeast brands.
Expert in software quality engineering and site reliability for consumer-scale distributed systems; owns test strategy, SLOs/error budgets, CI/CD test gates, observability, incident response, and reliability tooling.
New York or Los Angeles or Chicago or Houston or Tempe or Philadelphia or Dallas or North Miami Beach or Denver
$111k-$145k/yrHybridFull Time
JacobsNYSE: J: Global provider of professional engineering and technical services.
4+ YOEBachelor's in electrical engineering, 4+ years power-system and reliability analysis experience, familiarity with reliability indicators and NERC/FERC standards, strong analytical and communication skills.
Mistral AI: Developing frontier artificial intelligence models and enterprise AI solutions.
7+ YOE7+ years SRE/DevOps experience, Master’s in CS/Engineering or related, strong cloud and distributed systems skills, Kubernetes/CI-CD/infra-as-code proficiency, scripting experience, observability and on-call experience.
Legora: AI workspace for legal document research and drafting.
Extensive experience designing and operating large-scale production systems; lead reliability initiatives; strong distributed systems architecture; in-person NYC role.
Grow Therapy: Platform connecting mental health providers with patients and insurance.
6+ YOE6+ years operating production systems; hands-on AWS, Kubernetes (EKS), Terraform; experience defining SLOs/SLAs and observability (DataDog); strong communication and systems-thinking skills; PostgreSQL experience a plus.
Site Reliability Engineer, Tech Infrastructure - USDS
New York, New York, United States
$137k-$259k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOE3+ years SRE/systems engineering experience, bachelor\u0002s in CS or related, proficiency in Python/Go/Java/Shell, Linux, cloud and distributed systems, monitoring and incident management.
Retool: Software platform for building custom internal business applications.
Experience operating production infrastructure (AWS), Kubernetes, Terraform, Postgres; programming in Go/Python/TypeScript/Java/Ruby; building observability and automation for customer-facing SaaS systems.
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yrRemoteFull Time
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
CarrierNYSE: CARR: Manufactures HVAC, refrigeration, and fire safety systems.
3+ YOEBachelor's degree, 3+ years experience in refrigeration/HVAC product design and development, expertise in heat transfer, fan systems, thermal management, testing, and reliability analysis.
New York City or Boston or Miami or Pittsburgh or Raleigh
$127k-$249k/yrHybridFull Time
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years in software development and distributed systems; proficiency in Python or Go; experience with stateful storage systems; containerization (Kubernetes); cloud platforms (AWS, GCP, Azure); Linux networking; customer-focused and automation-minded.
Scottsdale or San Francisco or Chicago or New York
$66k-$82k/yrHybridFull Time
Early Warning Services: Operates payment and risk solutions for the financial industry.
2+ YOEBachelor's degree in CS/IS; 2-5 years in IT/DevOps; ITIL/ITSM knowledge; Agile; strong problem solving; ability to relate business needs to system capabilities.
ITIL, ITSM, Agile, Networking, Distributed Systems
Optimal Market Technologies: Operates a broker-dealer platform for wholesale options execution.
Hands-on Linux and network administration, strong scripting/automation, Infrastructure-as-Code experience, production support and incident response ownership, staff management/mentoring, familiarity with C++, Python, SQL, Azure, and real-time trading systems.
C++, Python, SQL, Linux, CentOS7, RHEL9, Microsoft Azure, Microsoft Azure Virtual Desktop (AVD), Claude Code, FIX Protocol, PostgreSQL, AERON
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.