62 reliability engineering manager jobs at 41 companies in Buckhall, VA
4w
Save
Mark Applied
Hide
4w
Manager, Site Reliability Engineering
Reston or Austin
OnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOE1+ Mgmt8+ years in software engineering or infrastructure (or fewer with relevant degrees), 3–5 years automation/programming experience, data analysis skills, 1 year leadership experience preferred.
10+ YOEBachelor's in engineering and 10+ years' engineering experience; U.S. citizenship and ability to obtain Secret clearance; leadership in reliability/maintainability/safety, RAM/FMECA/FRACAS, MBE/modeling and proposal support.
AARP: Non-profit organization advocating for Americans aged 50 and older.
8+ YOE8+ years SRE/DevOps experience with leadership of enterprise-scale reliability, cloud-native architectures, CI/CD, and operational frameworks; bachelor\u0002s degree or equivalent; strong communication and decision-making skills.
CI/CD, Secure Software Development Lifecycle (SSDLC)
AARP: Nonprofit advocacy organization serving Americans aged 50 and older.
8+ YOEBachelor's or equivalent experience in CS/IT,8+ years SRE/DevOps/cloud operations with enterprise leadership,5+ years cloud-native and CI/CD experience,SSDLC and reliability program experience; U.S. work authorization required.
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
4+ YOEBachelor's in electrical engineering or related, 4+ years industrial/commercial engineering and commissioning experience in mission-critical facilities, experience in reliability engineering, physics-of-failure, root-cause analysis, statistical analysis, vendor management, and ability to travel domestically and internationally.
New York City or Boston or Chicago or Salt Lake City or Austin or Montreal or Canada or Washington or Dallas or United States or Vancouver or Toronto or Charlotte or Denver
$129k-$304k/yrRemoteFull Time
Chainlink Labs: Building decentralized oracle networks for blockchain smart contracts.
Deep experience building and operating production blockchain infrastructure, strong engineering judgment, leadership and coaching experience, stakeholder management, and ability to deliver reliable, scalable systems.
Bochum or Austin or Dubai or Geneva or London or Singapore or Tokyo or Washington, D.C.
€105k-€160k/yrOnsiteFull Time
Sonar: Provides automated tools for code quality and security analysis.
10+ YOERequires 10+ years in software engineering focused on SRE, cloud, or infrastructure; advanced AWS and IaC expertise; observability, resiliency, Agile, and cloud cost optimization experience.
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
J&J Worldwide ServicesNYSE: CBRE: Provides facilities maintenance and engineering services to federal agencies.
4+ YOERequires 4 years in condition monitoring or reliability engineering, 3 years with vibration, ultrasound, and infrared thermography, high school diploma, and ASNT Level 2 and Level 3 certifications.
Barbaricum: Provides technology and mission support to federal national security agencies.
10+ YOE10+ years SRE/systems administration experience, Bachelor\u0002s in CS/IT/related (Master's preferred), DoD Secret clearance, expertise in monitoring, automation, cloud (AWS, Microsoft Azure, Google Cloud), scripting (Python, Shell, PowerShell), and configuration management tools.
Ansible, Puppet, Chef, Python, Shell, Microsoft PowerShell, AWS, Microsoft Azure, Google Cloud, Windows, Linux
Ardent: Provides geospatial and digital transformation services to federal agencies.
Experience in 24x7 production monitoring and support, incident management and root cause analysis, hands-on AWS experience, ability to build monitoring/automation, leadership and strong communication skills.
Berkeley Research Group: Provides expert testimony and specialized business consulting services.
5+ YOEBachelor's in computer science or similar, 5+ years SRE or similar, programming in Golang/Ruby/Python, Kubernetes, cloud experience (Azure/AWS/GCP), observability tools, and incident management expertise.
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOE5+ years experience with Kubernetes and Linux, proficiency in Bash/Python, experience with infrastructure automation and large-scale server management; Top Secret/SCI clearance required or obtainable.
Medallia: Provides cloud-based software for experience management and analytics.
8+ YOEBachelor's in computer science or equivalent experience; 8+ years in SRE, platform engineering, DevOps, or infrastructure, or 5+ years at Staff scope. Requires AWS, Kubernetes, Terraform, Git, Linux, PostgreSQL, Python/Go, and on-call experience.
SWIFT: Provides secure messaging services for global financial transactions.
12+ YOE12+ years in SRE/DevOps/platform engineering; experience building enterprise-scale automation, Ansible, CI/CD, Linux/RHEL, ServiceNow integrations, Python, and strong cross-team leadership.
Ansible Automation Platform, CloudBees, Python, ServiceNow, Git, Microsoft Power BI, Linux, Red Hat Enterprise Linux (RHEL)
Cogent People: A government consulting and technology services firm delivering secure, scalable digital solutions for mission-critical federal and commercial programs.
Bachelor's degree or equivalent, experience in system reliability/DevOps/production support, observability and monitoring tools, incident management, cloud and automation, strong troubleshooting and communication skills.
WorkdayNASDAQ: WDAY: Provides cloud-based software for financial and human capital management.
5+ YOE5+ years managing large-scale cloud infrastructure with automation, CI/CD, Kubernetes, Terraform, security and strong collaboration; bachelor’s or equivalent and ability to obtain U.S. security clearance.
Terraform, Argo CD, Kubernetes, Amazon Web Services, C#, Python, Ruby, Rust, Go
Site Reliability Engineer (SRE) / Service Availability Manager
Bethesda, Maryland, United States
$96k-$145k/yrHybridFull Time
Marriott InternationalNASDAQ: MAR: Operates and franchises a global network of hotels and resorts.
5+ YOE5+ years IT experience, 3+ years IT operations and incident/change/release management, undergraduate degree or equivalent, on-call/24x7 availability, proficiency with Python and Shell, familiarity with Ansible, Jenkins, cloud platforms, IaC and containers.