155 systems reliability engineer jobs at 126 companies in Marlboro, NJ
3w
Save
Mark Applied
Hide
3w
Systems Reliability Engineer (SRE)
San Francisco or New York City
$150k-$170k/yrOnsiteFull Time
Claryo: AI-powered spatial software for optimizing warehouse operations
3+ YOE3+ years SRE/infrastructure experience, strong Linux and networking fundamentals, experience with Kubernetes, cloud platforms, observability tooling, and debugging distributed systems in production.
6+ YOEBachelor's in engineering and 6+ years experience (4 with MS); strong reliability, failure analysis, statistical skills; proficiency with Microsoft Office; ability to obtain security clearance.
Microsoft Excel, Microsoft Word, Microsoft PowerPoint, Minitab, JMP, Python, FRACAS
Two Sigma: Systematic investment management and quantitative trading firm.
1+ YOE1+ years reliability engineering experience (5+ preferred), BS/BA in Computer Science or technical discipline, proficiency in Python/Java/C++/Rust, experience with automation, UNIX/Linux and distributed systems preferred.
ZT SystemsNASDAQ: SANM: Manufactures custom server and storage hardware for data centers.
2+ YOEB.S. in Electrical Engineering/Computer Science or related,2+ years relevant experience,knowledge of computer/hardware systems,Python/Unix scripting,statistics and reliability modeling,ability to lead cross-functional reliability efforts.
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
New York or Los Angeles or Chicago or Houston or Tempe or Philadelphia or Dallas or North Miami Beach or Denver
$111k-$145k/yrHybridFull Time
JacobsNYSE: J: Global provider of professional engineering and technical services.
4+ YOEBachelor's in electrical engineering, 4+ years power-system and reliability analysis experience, familiarity with reliability indicators and NERC/FERC standards, strong analytical and communication skills.
Mistral AI: Developing frontier artificial intelligence models and enterprise AI solutions.
7+ YOE7+ years SRE/DevOps experience, Master’s in CS/Engineering or related, strong cloud and distributed systems skills, Kubernetes/CI-CD/infra-as-code proficiency, scripting experience, observability and on-call experience.
Grow Therapy: Platform connecting mental health providers with patients and insurance.
6+ YOE6+ years operating production systems; hands-on AWS, Kubernetes (EKS), Terraform; experience defining SLOs/SLAs and observability (DataDog); strong communication and systems-thinking skills; PostgreSQL experience a plus.
Retool: Software platform for building custom internal business applications.
Experience operating production infrastructure (AWS), Kubernetes, Terraform, Postgres; programming in Go/Python/TypeScript/Java/Ruby; building observability and automation for customer-facing SaaS systems.
Site Reliability Engineer, Tech Infrastructure - USDS
New York, New York, United States
$137k-$259k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOE3+ years SRE/systems engineering experience, bachelor\u0002s in CS or related, proficiency in Python/Go/Java/Shell, Linux, cloud and distributed systems, monitoring and incident management.
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
Requires a computer science bachelor's degree or equivalent experience, Python, production ML inference, AWS cloud infrastructure, Kubernetes, vulnerability management, distributed-systems debugging, and on-call participation.
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yrRemoteFull Time
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
CarrierNYSE: CARR: Manufactures HVAC, refrigeration, and fire safety systems.
3+ YOEBachelor's degree, 3+ years experience in refrigeration/HVAC product design and development, expertise in heat transfer, fan systems, thermal management, testing, and reliability analysis.
Cox Enterprises: Providing global communications, automotive services, and media solutions.
5+ YOE5+ years in software/platform/infrastructure engineering, strong coding (Python/Go/Java), AWS and Terraform experience, SRE and observability knowledge, system design and incident response skills.
MS in Reliability Engineering and reliability engineering experience; knowledge of failure modes and reliability testing (HALT, accelerated life), programming for test automation, database management, CRE preferred, and authorization to work in the U.S.
Scottsdale or San Francisco or Chicago or New York
$66k-$82k/yrHybridFull Time
Early Warning Services: Operates payment and risk solutions for the financial industry.
2+ YOEBachelor's degree in CS/IS; 2-5 years in IT/DevOps; ITIL/ITSM knowledge; Agile; strong problem solving; ability to relate business needs to system capabilities.
ITIL, ITSM, Agile, Networking, Distributed Systems
Optimal Market Technologies: Operates a broker-dealer platform for wholesale options execution.
Hands-on Linux and network administration, strong scripting/automation, Infrastructure-as-Code experience, production support and incident response ownership, staff management/mentoring, familiarity with C++, Python, SQL, Azure, and real-time trading systems.
C++, Python, SQL, Linux, CentOS7, RHEL9, Microsoft Azure, Microsoft Azure Virtual Desktop (AVD), Claude Code, FIX Protocol, PostgreSQL, AERON