97 infrastructure reliability engineer jobs at 36 companies in Camp Springs, MD

3mo
Save
Mark Applied
Hide
Infrastructure Reliability Engineer
Manassas or Sterling or Portland or Chicago or Dallas Fort Worth
OnsiteFull Time
STACK Infrastructure
STACK Infrastructure: Developer and operator of sustainable wholesale data center infrastructure.
5+ YOE5–8 years in critical infrastructure; strong fluency in electrical systems; RCA/forensic troubleshooting; bachelor’s in engineering or equivalent.
Power distribution equipment, Waveform analysis, Fault analysis tools
2mo
Save
Mark Applied
Hide
Infrastructure Reliability Engineer, Infrastructure Reliability & Quality
Herndon, Virginia, United States
$117k-$160k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Multinational technology focused on e-commerce and cloud computing.
4+ YOEBachelor's in electrical engineering or related, 4+ years industrial/commercial engineering and commissioning experience in mission-critical facilities, experience in reliability engineering, physics-of-failure, root-cause analysis, statistical analysis, vendor management, and ability to travel domestically and internationally.
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Fleet Infrastructure
Boston or Washington
$166k-$220k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology developing AI-powered autonomous military systems.
8+ YOEBuild and operate high-availability observability and telemetry systems, collaborate with engineering teams, participate in on-call rotations, and meet U.S. Person access requirements.
Lattice OS, Docker, Kubernetes, AWS, GCP, Azure, ClickHouse, ClickStack, Victoria Metrics, Prometheus, Grafana, ELK
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
Vienna, Virginia, United States
RemoteFull Time
Knexus
Knexus: Private AI research and engineering delivering secure, tested systems and data-science solutions to U.S. government agencies.
6+ YOE6+ years in infrastructure engineering; strong cloud expertise (GCP/AWS/Azure); security controls (NIST 800-53/800-171); DoD Cloud SRG; ATO/SSP experience; GCP certification desired; US citizen eligible for security clearance.
Google Cloud Platform, Amazon Web Services, Microsoft Azure, Kubernetes, IAM, SSP, ATO, NIST 800-53/800-171, DoD Cloud SRG, Google Cloud Professional certifications
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
McLean, Virginia, United States
$129k-$190k/yr OnsiteFull Time
Medallia
Medallia: Private SaaS provider of customer, employee, citizen, and patient experience-management software for enterprises.
5+ YOERequires 5+ years leading reliability, platform, infrastructure, or cloud operations initiatives; Kubernetes, cloud platforms, Linux, Python/Go/Bash, Terraform, CI/CD, GitOps, networking, incident response, and on-call experience.
Kubernetes, AWS, OCI, GCP, Linux, Python, Go, Bash, Terraform, CI/CD, GitOps, DNS, TLS/SSL, ArgoCD, Prometheus, Grafana, Loki, OpenTelemetry
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
McLean, Virginia, United States
$129k-$190k/yr OnsiteFull Time
Medallia
Medallia: Private SaaS provider of customer, employee, citizen, and patient experience-management software for enterprises.
5+ YOERequires 5+ years leading reliability, platform, infrastructure, or cloud operations initiatives; production-scale Kubernetes, cloud, Linux, Python/Go/Bash, Terraform, CI/CD, GitOps, networking, and incident response experience.
Kubernetes, AWS, OCI, GCP, Linux, Python, Go, Bash, Terraform, CI/CD, GitOps, ArgoCD, Prometheus, Grafana, Loki, OpenTelemetry
2mo
Save
Mark Applied
Hide
Sr. Site Reliability Engineer (Starshield)
Hawthorne or Redmond or Washington
$165k-$230k/yr OnsiteFull Time
SpaceX
SpaceXNasdaq: SPCX: Designing, manufacturing, and launching advanced rockets and spacecraft.
5+ YOE5+ years SRE/DevOps experience with Kubernetes and Linux, proficiency in Bash/Python and infrastructure automation, bachelor's in CS/IT/engineering or 7+ years experience, and ability to obtain/maintain Top Secret clearance.
Kubernetes, Linux, Bash, Python, Bazel, Makefiles, Terraform, Ansible, TCP/IP
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer - US Federal (VDI & Infrastructure)
Reston, Virginia, United States
$147k-$221k/yr HybridFull Time
Workday
WorkdayNASDAQ: WDAY: Provides cloud-based software for financial and human capital management.
5+ YOE5+ years SRE/DevOps experience, 3+ years in high-security environments, US citizenship, ability to obtain TS/SCI w/CI Poly, expertise with Terraform, Argo CD, Microsoft Entra ID, Intune, Windows 365/AVD, and security monitoring tools.
Terraform, Argo CD, Microsoft Intune, Windows 365, Azure Virtual Desktop (AVD), Microsoft Entra ID, Microsoft Sentinel, Defender for Endpoint, Defender for Cloud Apps, Kubernetes, Python, Go, Ruby
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Arlington, Virginia, United States
$230k-$250k/yr HybridFull Time
GovCIO
GovCIO: Provider of IT modernization and digital transformation services.
5+ YOE5+ years engineering experience with Linux/Windows, Kubernetes, cloud infrastructure, Python/Go scripting, Terraform, CI/CD and monitoring; active Secret clearance required.
Linux, Windows, Kubernetes, Python, Go, Terraform, AWS, Azure
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Home services marketplace helping homeowners find and hire local professionals for repairs, maintenance, and improvements.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
2w
Save
Mark Applied
Hide
Sr Reliability / Quality Engineer, Data Center Infrastructure Products
Alexandria or Austin or Denver or United States or Seattle
$200k-$250k/yr HybridFull Time
Tract Capital
Tract Capital: Private alternative asset manager that builds digital infrastructure businesses, including data-center land and hyperscale campus platforms.
10+ YOEBachelor’s degree in engineering or related field and 10+ years of reliability, quality, or product validation experience with mission-critical electromechanical systems; expertise in testing, FMEA, root cause analysis, and supplier quality.
FMEA, HALT, HASS, UL, CE, IEC, ASHRAE
2mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer - ARINCDirect (Remote)
Arlington or United States
$108k-$205k/yr RemoteFull Time
RTX
RTXNYSE: RTX: RTX provides advanced aerospace and defense systems and services.
8+ YOESTEM degree with 8+ years relevant experience (or advanced degree with 5+ years, or 12+ years without degree). Must be authorized to work in the U.S. without sponsorship. Experience in Linux, Docker, Kubernetes, infrastructure automation (Saltstack, Ansible, Terraform), hardware (servers, switches, cabling), monitoring, incident response, and capac...
Docker, Kubernetes, Saltstack, Ansible, Terraform, GitOps, Python, Linux Shell (bash, awk, sed), Linux, PostgreSQL, SQL, AWS, DNS, DHCP, LDAP, NFS
1mo
Save
Mark Applied
Hide
Manager, Site Reliability Engineering
Reston or Austin
OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOE1+ Mgmt8+ years in software engineering or infrastructure (or fewer with relevant degrees), 3–5 years automation/programming experience, data analysis skills, 1 year leadership experience preferred.
1mo
Save
Mark Applied
Hide
Staff TDI Site Reliability Engineer, Okta Federal
Washington, District of Columbia, United States
$174k-$239k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Identity management and access control software provider.
7+ YOE7+ years SRE/DevOps experience building and operating secure cloud infrastructure in air-gapped/government environments; experience with containers, Terraform, Python, monitoring, and AWS networking; active TS/SCI with polygraph required.
EKS, ECS Fargate, Terraform, Python, Splunk, CloudWatch, Grafana, BGP, IPsec, VPC, TGW, VPC endpoints
3mo
Save
Mark Applied
Hide
Site Reliability Engineer - TS/SCI with Poly
Annapolis Junction, Maryland, United States
$111k-$150k/yr OnsiteFull Time
General Dynamics Information TechnologyNYSE: GD: Provides mission-critical IT services to government and defense organizations.
5+ YOE5+ years in site reliability engineering or DevOps; TS/SCI with CI Poly; automation tools; enterprise infrastructure; ITIL/ITSM knowledge; US citizenship.
PowerShell, Python, Ansible, Terraform, SolarWinds, SCOM, Splunk, Nagios, ELK
2w
Save
Mark Applied
Hide
Sr Network Reliability Engineer
Los Angeles or Florida or Denver or Reston or El Segundo
$198k-$277k/yr OnsiteFull Time
Blue Origin: Develops reusable rockets and systems for human spaceflight.
8+ YOEBachelor’s degree or equivalent experience and 8+ years designing scalable network infrastructure. Requires networking, automation, security, multi-vendor configuration, communication skills, and U.S. work authorization.
TCP/IP, IPv4, IPv6, MPLS, BGP, OSPF, IS-IS, DHCP, DNS, NTP, Python, Go, Rust, C++, Juniper, JNCIS-ENT, JNCIS-DevOps, JNCIP, Palo Alto Networks, NIST SP 800-series, NIST SP 800-57, FIPS 140-2, FIPS 140-3, VPN, VLAN
1d
Save
Mark Applied
Hide
Senior Site Reliability Engineer
United States or Canada or San Francisco or Washington, D.C. or Germany or Austria or Slovenia or Netherlands or New York City
$143k-$203k/yr RemoteFull Time
Planet Labs
Planet LabsNew York Stock Exchange: PL: Public benefit providing daily satellite imagery, geospatial data, and analytics to businesses and government agencies.
6+ YOE6+ years building cloud-native services; bachelor's degree in computer science or similar; Kubernetes, infrastructure-as-code, CI/CD, distributed systems, Python, Bash, and Jira experience required.
Kubernetes, Talos, RKE2, Proxmox, k3s, Terraform, Ansible, Helm, Kustomize, Jenkins, GitLab CI/CD, Argo CD, CircleCI, Alloy, Prometheus, Grafana, OpenTelemetry, Python, Bash, Jira, CUDA
2w
Save
Mark Applied
Hide
Lead DevOps (Site Reliability Engineer)
McLean or Wilmington
HybridFull Time
Anza Mortgage Insurance
Anza Mortgage Insurance: Provides mortgage insurance and credit risk protection for lenders.
Bachelor's degree in computer science or equivalent experience; expertise in AWS, Kubernetes, Terraform, Terragrunt, CI/CD, monitoring, incident management, Git, Docker, and infrastructure automation.
AWS, Terraform, Terragrunt, Argo CD, Argo Workflows, Kubernetes, GitHub Actions, Amazon EKS, AWS Fargate, Amazon Aurora, Docker, Git, Datadog, Cloudflare, GuardDuty, Security Hub, Camunda
2mo
Save
Mark Applied
Hide
Site Reliability Engineer, Steaming HUB - FreeWheel
Reston, Virginia, United States
OnsiteFull Time
Comcast
ComcastNASDAQ: CMCSA: Provides global telecommunications, media content, and entertainment services.
5+ YOE5+ years relevant experience; strong AWS and Kubernetes experience; Terraform, Ansible, Python/Go scripting; infrastructure, networking, and production incident management skills.
Amazon Web Services (AWS), Oracle Cloud Infrastructure (OCI), Amazon EKS, Kubernetes, Docker, Terraform, Ansible, Jenkins, Python, Go (Golang), Route 53, IAM, VPCs, Load Balancers
3d
Save
Mark Applied
Hide
Site Reliability Engineer, Streaming HUB - FreeWheel
Reston, Virginia, United States
$109k-$164k/yr OnsiteFull Time
Comcast
ComcastNASDAQ: CMCSA: A global media and technology.
5+ YOEBachelor's degree preferred; 5–7 years of relevant experience; strong AWS, Kubernetes, Terraform, networking, production infrastructure, automation, Python, and Go skills.
Amazon Web Services (AWS), Oracle Cloud Infrastructure (OCI), Amazon EKS, Kubernetes, Docker, Terraform, Ansible, Jenkins, Python, Go