138 systems reliability engineer jobs at 109 companies in Edison, NJ
2h
Save
Mark Applied
Hide
2h
Systems Reliability Engineer (SRE)
San Francisco or New York City
$150k-$170k/yrOnsiteFull Time
Claryo: AI-powered spatial software for optimizing warehouse operations
3+ YOE3+ years SRE/infrastructure experience, strong Linux and networking fundamentals, experience with Kubernetes, cloud platforms, observability tooling, and debugging distributed systems in production.
6+ YOEBachelor's in engineering and 6+ years experience (4 with MS); strong reliability, failure analysis, statistical skills; proficiency with Microsoft Office; ability to obtain security clearance.
Microsoft Excel, Microsoft Word, Microsoft PowerPoint, Minitab, JMP, Python, FRACAS
Two Sigma: Systematic investment management and quantitative trading firm.
1+ YOE1+ years reliability engineering experience (5+ preferred), BS/BA in Computer Science or technical discipline, proficiency in Python/Java/C++/Rust, experience with automation, UNIX/Linux and distributed systems preferred.
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
Beast Industries: Produces digital media and consumer goods for MrBeast brands.
Expert in software quality engineering and site reliability for consumer-scale distributed systems; owns test strategy, SLOs/error budgets, CI/CD test gates, observability, incident response, and reliability tooling.
New York or Los Angeles or Chicago or Houston or Tempe or Philadelphia or Dallas or North Miami Beach or Denver
$111k-$145k/yrHybridFull Time
JacobsNYSE: J: Global provider of professional engineering and technical services.
4+ YOEBachelor's in electrical engineering, 4+ years power-system and reliability analysis experience, familiarity with reliability indicators and NERC/FERC standards, strong analytical and communication skills.
Mistral AI: Developing frontier artificial intelligence models and enterprise AI solutions.
7+ YOE7+ years SRE/DevOps experience, Master’s in CS/Engineering or related, strong cloud and distributed systems skills, Kubernetes/CI-CD/infra-as-code proficiency, scripting experience, observability and on-call experience.
Legora: AI workspace for legal document research and drafting.
Extensive experience designing and operating large-scale production systems; lead reliability initiatives; strong distributed systems architecture; in-person NYC role.
Grow Therapy: Platform connecting mental health providers with patients and insurance.
6+ YOE6+ years operating production systems; hands-on AWS, Kubernetes (EKS), Terraform; experience defining SLOs/SLAs and observability (DataDog); strong communication and systems-thinking skills; PostgreSQL experience a plus.
Site Reliability Engineer, Tech Infrastructure - USDS
New York, New York, United States
$137k-$259k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOE3+ years SRE/systems engineering experience, bachelor\u0002s in CS or related, proficiency in Python/Go/Java/Shell, Linux, cloud and distributed systems, monitoring and incident management.
Retool: Software platform for building custom internal business applications.
Experience operating production infrastructure (AWS), Kubernetes, Terraform, Postgres; programming in Go/Python/TypeScript/Java/Ruby; building observability and automation for customer-facing SaaS systems.
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yrRemoteFull Time
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
CarrierNYSE: CARR: Manufactures HVAC, refrigeration, and fire safety systems.
3+ YOEBachelor's degree, 3+ years experience in refrigeration/HVAC product design and development, expertise in heat transfer, fan systems, thermal management, testing, and reliability analysis.
New York City or Boston or Miami or Pittsburgh or Raleigh
$127k-$249k/yrHybridFull Time
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years in software development and distributed systems; proficiency in Python or Go; experience with stateful storage systems; containerization (Kubernetes); cloud platforms (AWS, GCP, Azure); Linux networking; customer-focused and automation-minded.
MS in Reliability Engineering and reliability engineering experience; knowledge of failure modes and reliability testing (HALT, accelerated life), programming for test automation, database management, CRE preferred, and authorization to work in the U.S.