41 infrastructure reliability engineer jobs at 24 companies in Rosedale, MD

1mo
Save
Mark Applied
Hide
Customer Reliability Engineer - Infrastructure
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yr RemoteFull Time
Astronomer
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
Apache Airflow, AWS, Azure, CI/CD, GCP, Infrastructure as Code (IaC), Kubernetes, Linux, Python
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Fleet Infrastructure
Boston or Washington
$166k-$220k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology building autonomous military hardware and software.
8+ YOEBuild and operate high-availability observability and telemetry systems, collaborate with engineering teams, participate in on-call rotations, and meet U.S. Person access requirements.
Lattice OS, Docker, Kubernetes, AWS, GCP, Azure, ClickHouse, ClickStack, Victoria Metrics, Prometheus, Grafana, ELK
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
Vienna, Virginia, United States
RemoteFull Time
Knexus
Knexus: Provider of applied artificial intelligence for government missions.
6+ YOE6+ years in infrastructure engineering; strong cloud expertise (GCP/AWS/Azure); security controls (NIST 800-53/800-171); DoD Cloud SRG; ATO/SSP experience; GCP certification desired; US citizen eligible for security clearance.
Google Cloud Platform, Amazon Web Services, Microsoft Azure, Kubernetes, IAM, SSP, ATO, NIST 800-53/800-171, DoD Cloud SRG, Google Cloud Professional certifications
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
McLean, Virginia, United States
$129k-$190k/yr OnsiteFull Time
Medallia
Medallia: Provides cloud-based software for experience management and analytics.
5+ YOERequires 5+ years leading reliability, platform, infrastructure, or cloud operations initiatives; Kubernetes, cloud platforms, Linux, Python/Go/Bash, Terraform, CI/CD, GitOps, networking, incident response, and on-call experience.
Kubernetes, AWS, OCI, GCP, Linux, Python, Go, Bash, Terraform, CI/CD, GitOps, DNS, TLS/SSL, ArgoCD, Prometheus, Grafana, Loki, OpenTelemetry
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
McLean, Virginia, United States
$129k-$190k/yr OnsiteFull Time
Medallia
Medallia: Provides enterprise experience management and customer feedback software.
5+ YOERequires 5+ years leading reliability, platform, infrastructure, or cloud operations initiatives; production-scale Kubernetes, cloud, Linux, Python/Go/Bash, Terraform, CI/CD, GitOps, networking, and incident response experience.
Kubernetes, AWS, OCI, GCP, Linux, Python, Go, Bash, Terraform, CI/CD, GitOps, ArgoCD, Prometheus, Grafana, Loki, OpenTelemetry
2mo
Save
Mark Applied
Hide
Sr. Site Reliability Engineer (Starshield)
Hawthorne or Redmond or Washington
$165k-$230k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOE5+ years SRE/DevOps experience with Kubernetes and Linux, proficiency in Bash/Python and infrastructure automation, bachelor's in CS/IT/engineering or 7+ years experience, and ability to obtain/maintain Top Secret clearance.
Kubernetes, Linux, Bash, Python, Bazel, Makefiles, Terraform, Ansible, TCP/IP
3w
Save
Mark Applied
Hide
Site Reliability Engineer
Arlington, Virginia, United States
$230k-$250k/yr HybridFull Time
GovCIO
GovCIO: Provides technology and mission support services to federal agencies.
5+ YOE5+ years engineering experience with Linux/Windows, Kubernetes, cloud infrastructure, Python/Go scripting, Terraform, CI/CD and monitoring; active Secret clearance required.
Linux, Windows, Kubernetes, Python, Go, Terraform, AWS, Azure
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
2mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer - ARINCDirect (Remote)
Arlington or United States
$108k-$205k/yr RemoteFull Time
RTX
RTXNYSE: RTX: RTX provides advanced aerospace and defense systems and services.
8+ YOESTEM degree with 8+ years relevant experience (or advanced degree with 5+ years, or 12+ years without degree). Must be authorized to work in the U.S. without sponsorship. Experience in Linux, Docker, Kubernetes, infrastructure automation (Saltstack, Ansible, Terraform), hardware (servers, switches, cabling), monitoring, incident response, and capac...
Docker, Kubernetes, Saltstack, Ansible, Terraform, GitOps, Python, Linux Shell (bash, awk, sed), Linux, PostgreSQL, SQL, AWS, DNS, DHCP, LDAP, NFS
1mo
Save
Mark Applied
Hide
Staff TDI Site Reliability Engineer, Okta Federal
Washington, District of Columbia, United States
$174k-$239k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
7+ YOE7+ years SRE/DevOps experience building and operating secure cloud infrastructure in air-gapped/government environments; experience with containers, Terraform, Python, monitoring, and AWS networking; active TS/SCI with polygraph required.
EKS, ECS Fargate, Terraform, Python, Splunk, CloudWatch, Grafana, BGP, IPsec, VPC, TGW, VPC endpoints
3mo
Save
Mark Applied
Hide
Site Reliability Engineer - TS/SCI with Poly
Annapolis Junction, Maryland, United States
$111k-$150k/yr OnsiteFull Time
General Dynamics Information TechnologyNYSE: GD: Provides mission-critical IT services to government and defense organizations.
5+ YOE5+ years in site reliability engineering or DevOps; TS/SCI with CI Poly; automation tools; enterprise infrastructure; ITIL/ITSM knowledge; US citizenship.
PowerShell, Python, Ansible, Terraform, SolarWinds, SCOM, Splunk, Nagios, ELK
1w
Save
Mark Applied
Hide
Lead DevOps (Site Reliability Engineer)
McLean or Wilmington
HybridFull Time
Anza Mortgage Insurance
Anza Mortgage Insurance: Provides mortgage insurance and credit risk protection for lenders.
Bachelor's degree in computer science or equivalent experience; expertise in AWS, Kubernetes, Terraform, Terragrunt, CI/CD, monitoring, incident management, Git, Docker, and infrastructure automation.
AWS, Terraform, Terragrunt, Argo CD, Argo Workflows, Kubernetes, GitHub Actions, Amazon EKS, AWS Fargate, Amazon Aurora, Docker, Git, Datadog, Cloudflare, GuardDuty, Security Hub, Camunda
2mo
Save
Mark Applied
Hide
Site Reliability Engineer - Software Ops and Scaling , One Material Handling System - Software, Controls and Science
Nashville or Arlington or Bellevue
$94k-$160k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
Experience automating, deploying, and supporting infrastructure; programming in Python, Ruby, Golang, Java, C++, C#, or Rust; Linux/Unix experience; CI/CD pipeline experience preferred.
Python, Ruby, Golang, Java, C++, C#, Rust, Linux/Unix, CI/CD
2mo
Save
Mark Applied
Hide
Systems Administrator / Site Reliability Engineer
Reston or Washington
OnsiteFull Time
Assured Consulting Solutions
Assured Consulting Solutions: An equal opportunity employer delivering technology solutions for government and industry.
8+ YOE8+ years cloud infrastructure management (AWS, Kubernetes); strong security framework knowledge; Red Hat/Linux admin; 8140 (Security+) and AWS certs; BS or higher or equivalent experience.
AWS, OpenShift, Kubernetes, Linux, Red Hat, PKI, Secrets Manager, DynamoDB, S3, RDS, ABAC
1mo
Save
Mark Applied
Hide
Sr. Director, Technology, Infrastructure & Stabilization
Oakland or United States or Washington or Ohio or District of Columbia or California or Long Beach or Arizona or Colorado or Connecticut or Florida or Georgia or Maryland or Minnesota or Nevada or Oregon or El Dorado Hills or San Diego or Woodland Hills or Alabama or Illinois or Virginia or Wisconsin or Texas or New York
$240k-$359k/yr HybridFull Time
Ascendiun
Ascendiun: Nonprofit parent overseeing health insurance and clinical service organizations.
12+ YOE6+ Mgmt12 years engineering experience, 6 years people management; experience with cloud-native architectures, CI/CD, observability, reliability, and stabilizing legacy systems; strong communication and cross-functional leadership.
2d
Save
Mark Applied
Hide
Site Reliability Engineering Lead (f/m/d)
Bochum or Austin or Dubai or Geneva or London or Singapore or Tokyo or Washington, D.C.
€105k-€160k/yr OnsiteFull Time
Sonar
Sonar: Provides automated tools for code quality and security analysis.
10+ YOERequires 10+ years in software engineering focused on SRE, cloud, or infrastructure; advanced AWS and IaC expertise; observability, resiliency, Agile, and cloud cost optimization experience.
AWS, Python, CDK, Terraform, Aurora DBs, OpenSearch, IAM, OUs, Account Vending
3d
Save
Mark Applied
Hide
Applied AI Lead - Infrastructure & Deployment
Beavercreek or Frederick or Fort Walton Beach or Dayton
OnsiteFull Time
Leonardo DRS
Leonardo DRSNASDAQ: DRS: Manufactures advanced electronic systems for defense and military applications.
5+ YOERequires 5+ years in infrastructure, platform, or site reliability engineering, including 3+ years in MLOps or model serving; hands-on self-hosted AI, security engineering, cloud development, bachelor's degree or equivalent, U.S. citizenship, and clearance eligibility.
MLOps, vLLM, GCC High, FedRAMP, CMMC, VRAM
2w
Save
Mark Applied
Hide
Staff Software Engineer, Platform
San Francisco or New York or Washington
$240k-$300k/yr OnsiteFull Time
Peregrine
Peregrine: Data integration and analytics platform for public safety agencies.
8+ YOE8+ years building and operating cloud infrastructure, hands-on ownership of platform/networking/traffic systems, strong security and reliability experience, on-call and incident response experience, degree or equivalent.
3d
Save
Mark Applied
Hide
Principal Platform Engineer
Herndon or Columbia
OnsiteFull Time
Clarity Innovations
Clarity Innovations: Provides software and data engineering for national security missions.
Requires Terraform, GitLab CI/CD, containerized and serverless workloads, databases, Linux infrastructure, Docker, Podman, Kubernetes, and secure, reliable infrastructure engineering.
Terraform, GitLab CI/CD, Docker, Podman, Kubernetes, Airflow, EMR, Ansible, Puppet, Chef, Helm, ArgoCD
1w
Save
Mark Applied
Hide
Senior Software Engineer, Data Infrastructure - AI Platform
United States or North America or San Francisco or Los Angeles or New York City or Washington, D.C. or London or Singapore
RemoteFull Time
TRM Labs
TRM Labs: An AI-powered intelligence helping agencies investigate crime and disrupt threat networks.
Requires U.S. citizenship, distributed OLAP or serving-layer operations experience, query tuning, data pipeline reliability, incident response, AI tool proficiency, and on-call ownership.
StarRocks, Trino, ClickHouse, Claude, Cursor, Slack, Otter.ai, Fireflies, Fathom, Cluey