48 infrastructure reliability engineer jobs at 24 companies in Lanham, MD
2mo
Save
Mark Applied
Hide
2mo
Infrastructure Reliability Engineer
Manassas or Sterling or Portland or Chicago or Dallas Fort Worth
OnsiteFull Time
STACK Infrastructure: Developer and operator of sustainable wholesale data center infrastructure.
5+ YOE5–8 years in critical infrastructure; strong fluency in electrical systems; RCA/forensic troubleshooting; bachelor’s in engineering or equivalent.
Power distribution equipment, Waveform analysis, Fault analysis tools
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
4+ YOEBachelor's in electrical engineering or related, 4+ years industrial/commercial engineering and commissioning experience in mission-critical facilities, experience in reliability engineering, physics-of-failure, root-cause analysis, statistical analysis, vendor management, and ability to travel domestically and internationally.
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yrRemoteFull Time
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
Applied Information Sciences: Delivers cloud transformation and software engineering services to organizations.
8+ YOE8+ years supporting enterprise cloud/infrastructure, bachelor\u0002s in CS or related (or equiv), IAT-2 and cloud certs, active Secret clearance, experience in regulated environments and secure engineering practices.
Senior Site Reliability Engineer, Fleet Infrastructure
Boston or Washington
$166k-$220k/yrOnsiteFull Time
Anduril Industries: Defense technology building autonomous military hardware and software.
8+ YOEBuild and operate high-availability observability and telemetry systems, collaborate with engineering teams, participate in on-call rotations, and meet U.S. Person access requirements.
Applied Information Sciences: An employee-owned technology firm that delivers architecture and infrastructure solutions and fosters continuous learning and inclusivity.
8+ YOE8+ years supporting enterprise cloud/infrastructure, bachelor's in CS/IS/Engineering (or equiv), IAT-2 and cloud certs, active Secret clearance, experience in regulated environments, secure engineering and documentation practices.
ManTech: Provides technology solutions for defense and intelligence agencies.
5+ YOEBachelor's in CS/Engineering, 5+ years SRE/DevOps experience, Kubernetes and cloud-native infrastructure expertise, observability and monitoring experience (OpenTelemetry, Prometheus, Grafana, Loki, Tempo), incident management and automation skills, TS/SCI with Poly for onsite work.
Knexus: Provider of applied artificial intelligence for government missions.
6+ YOE6+ years in infrastructure engineering; strong cloud expertise (GCP/AWS/Azure); security controls (NIST 800-53/800-171); DoD Cloud SRG; ATO/SSP experience; GCP certification desired; US citizen eligible for security clearance.
Google Cloud Platform, Amazon Web Services, Microsoft Azure, Kubernetes, IAM, SSP, ATO, NIST 800-53/800-171, DoD Cloud SRG, Google Cloud Professional certifications
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOE5+ years SRE/DevOps experience with Kubernetes and Linux, proficiency in Bash/Python and infrastructure automation, bachelor's in CS/IT/engineering or 7+ years experience, and ability to obtain/maintain Top Secret clearance.
GovCIO: Provides technology and mission support services to federal agencies.
5+ YOE5+ years engineering experience with Linux/Windows, Kubernetes, cloud infrastructure, Python/Go scripting, Terraform, CI/CD and monitoring; active Secret clearance required.
WorkdayNASDAQ: WDAY: Provides cloud-based software for financial and human capital management.
5+ YOE5+ years infrastructure experience, Kubernetes expertise, Terraform and Argo CD proficiency, 3+ years AWS experience, programming in GoLang or Python, ability to obtain US security clearance.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
Principal Site Reliability Engineer - ARINCDirect (Remote)
Arlington or United States
$108k-$205k/yrRemoteFull Time
RTXNYSE: RTX: RTX provides advanced aerospace and defense systems and services.
8+ YOESTEM degree with 8+ years relevant experience (or advanced degree with 5+ years, or 12+ years without degree). Must be authorized to work in the U.S. without sponsorship. Experience in Linux, Docker, Kubernetes, infrastructure automation (Saltstack, Ansible, Terraform), hardware (servers, switches, cabling), monitoring, incident response, and capac...
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOE1+ Mgmt8+ years in software engineering or infrastructure (or fewer with relevant degrees), 3–5 years automation/programming experience, data analysis skills, 1 year leadership experience preferred.
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
7+ YOE7+ years SRE/DevOps experience building and operating secure cloud infrastructure in air-gapped/government environments; experience with containers, Terraform, Python, monitoring, and AWS networking; active TS/SCI with polygraph required.
General Dynamics Information TechnologyNYSE: GD: Provides mission-critical IT services to government and defense organizations.
5+ YOE5+ years in site reliability engineering or DevOps; TS/SCI with CI Poly; automation tools; enterprise infrastructure; ITIL/ITSM knowledge; US citizenship.
Lockheed MartinNYSE: LMT: Designs and manufactures global security and aerospace systems.
Owning infrastructure-as-code and cloud platform (AWS/Azure) for Apriso; strong software development and operations experience, RDBMS/SQL skills, CI/CD and automation (GitLab/Ansible), Windows/Linux, scripting, and SOX-compliant change management.
Digital RealtyNYSE: DLR: Provides global data center, colocation, and interconnection solutions.
5+ YOEBachelor of Science in Electrical Engineering; 5 years facility management in high-reliability data centers; proficient in Microsoft Office; strong communication and technical writing; project and budget management; capable of extended hours.
Microsoft Office, Microsoft Project, Visio, AutoCAD