52 systems reliability engineer jobs at 40 companies in Towson, MD

1mo
Save
Mark Applied
Hide
Senior Reliability Engineer
Laurel, Maryland, United States
$100k-$245k/yr OnsiteFull Time
Johns Hopkins University Applied Physics Laboratory
Johns Hopkins University Applied Physics Laboratory: Conducts research and engineering for national security and space.
5+ YOEBachelor's in engineering/math/physics, 5+ years reliability (RAM) engineering for complex weapon systems, experience with FMEA/FMECA/FTA/PRA, strong communication, ability to travel periodically, active Secret clearance and ability to obtain Top Secret.
Reliasoft, Windchill, Saphire, MADe, JMP, R, SAS, Matlab, Python, CAMEO, DOORS
1d
Save
Mark Applied
Hide
Reliability Engineer
Arlington, Virginia, United States
$80k-$160k/yr OnsiteFull Time
Decision Technologies
Decision Technologies: Provides engineering and technical support for defense programs.
4+ YOEBachelor's degree in a technical discipline, 4+ years RM&A experience on US Navy platforms or weapon systems, eligible for DoD Secret clearance, strong communication and presentation skills.
SIPR
4w
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
1mo
Save
Mark Applied
Hide
Senior Reliability Engineer
Washington, District of Columbia, United States
OnsiteFull Time
Barbaricum
Barbaricum: Provides technology and mission support to federal national security agencies.
10+ YOE10+ years SRE/systems administration experience, Bachelor\u0002s in CS/IT/related (Master's preferred), DoD Secret clearance, expertise in monitoring, automation, cloud (AWS, Microsoft Azure, Google Cloud), scripting (Python, Shell, PowerShell), and configuration management tools.
Ansible, Puppet, Chef, Python, Shell, Microsoft PowerShell, AWS, Microsoft Azure, Google Cloud, Windows, Linux
2w
Save
Mark Applied
Hide
Sr. Project Engineer- Electrical Systems Reliability
Towson, Maryland, United States
$90k-$145k/yr HybridFull Time
Stanley Black & Decker
Stanley Black & DeckerNYSE: SWK: Manufacturer of power tools, hand tools, and outdoor equipment.
5+ YOE5+ years reliability/quality experience in NPD for electrical systems; bachelor in electrical or mechanical engineering required; leadership, reliability analysis, and compliance experience required.
Weibull, FRACAS
2mo
Save
Mark Applied
Hide
Systems Administrator / Site Reliability Engineer
Reston or Washington
OnsiteFull Time
Assured Consulting Solutions
Assured Consulting Solutions: An equal opportunity employer delivering technology solutions for government and industry.
8+ YOE8+ years cloud infrastructure management (AWS, Kubernetes); strong security framework knowledge; Red Hat/Linux admin; 8140 (Security+) and AWS certs; BS or higher or equivalent experience.
AWS, OpenShift, Kubernetes, Linux, Red Hat, PKI, Secrets Manager, DynamoDB, S3, RDS, ABAC
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Intelligence Systems
Reston, Virginia, United States
$146k-$194k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology building autonomous military hardware and software.
Active U.S. Top Secret/SCI clearance; mid-level systems administration/DevOps/SRE experience with Linux, IaC (Terraform/Nix), CI/CD, scripting, GCP, containers/Kubernetes, and test automation; operational mindset.
Lattice OS, Terraform, Nix, GCP, Kubernetes, GCR, GitOps, Linux, CI/CD
2mo
Save
Mark Applied
Hide
Sr. Site Reliability Engineer
Washington, District of Columbia, United States
OnsiteFull Time
Tiger Analytics
Tiger Analytics: Provides AI and advanced analytics consulting for global enterprises.
Kubernetes, Docker; MLOps tools; Python, Bash; Go a plus; data systems; networking; IaC; CI/CD; monitoring; on-call readiness.
Kubernetes, Docker, Kubeflow, Vertex AI, MLflow, DVC, Python, Bash, Go, BigQuery, Pub/Sub, Pinecone, Milvus, Istio, Anthos, Terraform, Pulumi, GitHub Actions, ArgoCD, Prometheus, Grafana, Google Cloud Operations Suite, SLA/SLO, K8s (GKE), Vertex AI Endpoints, Kubeflow Pipelines
1mo
Save
Mark Applied
Hide
Customer Reliability Engineer - Infrastructure
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yr RemoteFull Time
Astronomer
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
Apache Airflow, AWS, Azure, CI/CD, GCP, Infrastructure as Code (IaC), Kubernetes, Linux, Python
2mo
Save
Mark Applied
Hide
Site Reliability Engineer II
Falls Church or South Carolina or Raleigh or Nashville or Louisiana or Pennsylvania or Plain City or South Bend or Orlando or Detroit
OnsiteFull Time
Kastle Systems
Kastle Systems: Managed security services provider for commercial and residential properties.
4+ YOE4+ years SRE/Platform experience owning production systems. Hands-on with Azure/AKS, Kubernetes, Terraform/OpenTofu/Pulumi, GitOps/ArgoCD, observability (Prometheus/Grafana/OpenTelemetry/ELK), Python/Go/Bash, and feature-flag/CI/CD practices.
ArgoCD, Flux, Crossplane, LaunchDarkly, Flagsmith, Terraform, OpenTofu, Pulumi, Prometheus, Grafana, OpenTelemetry, ELK, OpenSearch, Python, Go, Bash, C#, SQL, AKS, Azure Container Registry, Azure Monitor, Cosmos DB, Key Vault, Azure Front Door, GitOps
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Columbia, Maryland, United States
HybridFull Time
Cogent People
Cogent People: A government consulting and technology services firm delivering secure, scalable digital solutions for mission-critical federal and commercial programs.
Bachelor's degree or equivalent, experience in system reliability/DevOps/production support, observability and monitoring tools, incident management, cloud and automation, strong troubleshooting and communication skills.
AWS, Terraform, Splunk, Datadog, Prometheus, CI/CD
2mo
Save
Mark Applied
Hide
Jr. Site Reliability Engineer
Annapolis Junction, Maryland, United States
OnsiteFull Time
Entegra Systems
Entegra Systems: Provides engineering and intelligence analysis for government missions.
5+ YOEFive years software development; one year system engineering; distributed systems; scripting; TS/SCI Poly clearance.
Kubernetes, Red Hat Enterprise Linux, Windows, Solaris, Java, Python, Perl, Ruby, Hadoop, Hbase, Cassandra, CloudBase/Accumulo, CI/CD, Networking, SOA
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
1mo
Save
Mark Applied
Hide
Site Reliability / Operations Engineer (TS/SCI)
Herndon or Colorado Springs or Melbourne or Washington, DC
$82k-$132k/yr OnsiteFull Time
Vantor
Vantor: Providing AI-powered spatial intelligence and high-resolution Earth observation.
2+ YOEActive TS/SCI clearance, Security+ cert, Bachelor's in CS/IS/Engineering or equivalent, 2+ years systems/automation/DevOps experience, Linux (RHEL/CentOS/Ubuntu), Bash, ELK Stack (Elasticsearch, Logstash, Kibana).
Linux, RHEL, CentOS, Ubuntu, Bash, ELK Stack, Elasticsearch, Logstash, Kibana, Python
1mo
Save
Mark Applied
Hide
Site Reliability / Operations Engineer (TS/SCI)
Herndon or Colorado Springs or Melbourne or Washington
$82k-$132k/yr OnsiteFull Time
Maxar Intelligence
Maxar Intelligence: Providing satellite imagery and geospatial intelligence for global security.
2+ YOEActive TS/SCI clearance, Security+ certification, Bachelor's in CS/IS/Engineering or equivalent, 2+ years systems/DevOps experience, Linux and Bash proficiency, experience with ELK Stack (Elasticsearch, Logstash, Kibana), monitoring and troubleshooting skills.
Linux, RHEL, CentOS, Ubuntu, Bash, ELK Stack, Elasticsearch, Kibana, Logstash, Python, CI/CD
3mo
Save
Mark Applied
Hide
Lead Systems Engineer - Traffic Management
Durham or Miami or Palo Alto or Washington
$15k-$23k/yr HybridFull Time
Nubank
NubankNYSE: NU: Digital financial platform offering banking, credit, and investment services.
Experience operating large-scale cloud distributed systems, Kubernetes, service mesh (Istio/Linkerd/Envoy), AWS networking, IaC with Pulumi or Terraform, strong reliability/observability and AI-assisted workflows.
Kubernetes, Istio, Linkerd, Envoy, Finagle, AWS, ALB, NLB, VPC, Route 53, IAM, Pulumi, Terraform
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Rockville, Maryland, United States
$112k-$150k/yr HybridFull Time
Skyward IT Solutions
Skyward IT Solutions: Delivering digital modernization and AI solutions for federal agencies.
3+ YOE3+ years SRE/systems/cloud experience with AWS; hands-on Terraform/Ansible/CloudFormation, Jenkins or GitLab CI, Docker, Python, and observability tooling; bachelor\u0002s degree or equivalent; ability to obtain Public Trust clearance.
AWS, AWS CloudWatch, AWS Trusted Advisor, CloudFormation, Docker, GitLab CI, Jenkins, New Relic, Splunk, Terraform, Ansible, Python
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - CTJ - Top Secret
Redmond or Reston
$120k-$235k/yr HybridFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
2+ YOEDegree in CS/IT (MS+2yrs or BS+4yrs) or equivalent experience; active Top Secret clearance with ability/eligibility to maintain TS/SCI (with polygraph); experience with large-scale cloud/distributed systems; proficiency in C#, Go, Java, or Python; CI/CD and automation experience.
C#, Go, Java, Python, CI/CD
6d
Save
Mark Applied
Hide
Engineer (Reliability) - Hanover, PA
Hanover or Dillsburg or Gettysburg or York
OnsiteFull Time
FirstEnergy
FirstEnergyNYSE: FE: Provides electric power distribution and transmission services to customers.
Bachelor of Science in Engineering/Engineering Technology (ABET) required or PE/license alternative; FE exam a plus. Knowledge of electrical systems, NESC/NEC, switching/tagging, ability to analyze data and write technical reports, valid state driver’s license, willing to travel and work irregular hours.
2mo
Save
Mark Applied
Hide
Sr. Site Reliability Engineer III (6448)
Washington, District of Columbia, United States
$185k-$230k/yr HybridFull Time
MetroStar
MetroStar: Provides digital transformation and IT services to government agencies.
7+ YOE7+ years in software development, systems engineering, or operations; Top Secret clearance; DoD 8140 certification; Bachelor’s degree preferred; strong automation and cloud/container experience.
Ansible, GitLab CI/CD, Bash, Kubernetes, MinIO, S3-compatible storage, PortWorx, F5, VMware, CI/CD, Monitoring, Observability