37 platform reliability engineer jobs at 20 companies in Ilchester, MD

2w
Save
Mark Applied
Hide
Reliability Engineer
Arlington, Virginia, United States
$80k-$160k/yr OnsiteFull Time
Decision Technologies
Decision Technologies: Provides engineering and technical support for defense programs.
4+ YOEBachelor's degree in a technical discipline, 4+ years RM&A experience on US Navy platforms or weapon systems, eligible for DoD Secret clearance, strong communication and presentation skills.
SIPR
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
1mo
Save
Mark Applied
Hide
Platform Delivery & Reliability Engineer (Remote) - 29337
Colorado Springs or Columbia or San Antonio or Boise or Greenville or Augusta
$120k-$195k/yr RemoteFull Time
HII
HIINYSE: HII: Builds naval ships and provides global defense technology solutions.
7+ YOEMust obtain U.S. security clearance; 7+ years' relevant experience (varies by degree); deep Kubernetes, IaC, cloud, CI/CD, scripting, SRE practices, troubleshooting and delivery experience across distributed systems.
Kubernetes, Terraform, AWS, Azure, GCP, GitLab CI, Go, Python, Bash, Spark, Trino, Presto, Kafka, NiFi, Iceberg, Delta, Hudi, YouTrack, Nexus
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
McLean, Virginia, United States
$129k-$190k/yr OnsiteFull Time
Medallia
Medallia: Provides cloud-based software for experience management and analytics.
5+ YOERequires 5+ years leading reliability, platform, infrastructure, or cloud operations initiatives; Kubernetes, cloud platforms, Linux, Python/Go/Bash, Terraform, CI/CD, GitOps, networking, incident response, and on-call experience.
Kubernetes, AWS, OCI, GCP, Linux, Python, Go, Bash, Terraform, CI/CD, GitOps, DNS, TLS/SSL, ArgoCD, Prometheus, Grafana, Loki, OpenTelemetry
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
McLean, Virginia, United States
$129k-$190k/yr OnsiteFull Time
Medallia
Medallia: Provides enterprise experience management and customer feedback software.
5+ YOERequires 5+ years leading reliability, platform, infrastructure, or cloud operations initiatives; production-scale Kubernetes, cloud, Linux, Python/Go/Bash, Terraform, CI/CD, GitOps, networking, and incident response experience.
Kubernetes, AWS, OCI, GCP, Linux, Python, Go, Bash, Terraform, CI/CD, GitOps, ArgoCD, Prometheus, Grafana, Loki, OpenTelemetry
2mo
Save
Mark Applied
Hide
Site Reliability Engineer II
Falls Church or South Carolina or Raleigh or Nashville or Louisiana or Pennsylvania or Plain City or South Bend or Orlando or Detroit
OnsiteFull Time
Kastle Systems
Kastle Systems: Managed security services provider for commercial and residential properties.
4+ YOE4+ years SRE/Platform experience owning production systems. Hands-on with Azure/AKS, Kubernetes, Terraform/OpenTofu/Pulumi, GitOps/ArgoCD, observability (Prometheus/Grafana/OpenTelemetry/ELK), Python/Go/Bash, and feature-flag/CI/CD practices.
ArgoCD, Flux, Crossplane, LaunchDarkly, Flagsmith, Terraform, OpenTofu, Pulumi, Prometheus, Grafana, OpenTelemetry, ELK, OpenSearch, Python, Go, Bash, C#, SQL, AKS, Azure Container Registry, Azure Monitor, Cosmos DB, Key Vault, Azure Front Door, GitOps
1w
Save
Mark Applied
Hide
Data Platform Engineer - AI Platform
United States or North America or San Francisco or Los Angeles or New York City or Washington or London or Singapore
RemoteFull Time
TRM Labs
TRM Labs: An AI-powered intelligence technology helping agencies investigate crime and disrupt illicit activity.
U.S. citizenship, distributed OLAP or serving-layer operations experience, query tuning, data pipeline reliability, incident response, AI tool fluency, independent infrastructure ownership, and on-call readiness.
StarRocks, Claude, Trino, ClickHouse, Cursor, Slack, Otter.ai, Fireflies, Fathom, Cluey
4d
Save
Mark Applied
Hide
Principal Platform Engineer
Herndon or Columbia
OnsiteFull Time
Clarity Innovations
Clarity Innovations: Provides software and data engineering for national security missions.
Requires Terraform, GitLab CI/CD, containerized and serverless workloads, databases, Linux infrastructure, Docker, Podman, Kubernetes, and secure, reliable infrastructure engineering.
Terraform, GitLab CI/CD, Docker, Podman, Kubernetes, Airflow, EMR, Ansible, Puppet, Chef, Helm, ArgoCD
3d
Save
Mark Applied
Hide
Engineer 4 - Site Reliability Engineering
Reston, Virginia, United States
OnsiteFull Time
Comcast
ComcastNASDAQ: CMCSA: Provides global telecommunications, media content, and entertainment services.
8+ YOEBachelor's degree in computer science, software engineering, or related field; 8+ years as an SRE, DevOps, or data operations engineer; cloud, data platforms, automation, programming, monitoring, and troubleshooting expertise.
AWS, GCP, Azure, Kafka, Hadoop, Spark, Cassandra, HDFS, AWS S3, MySQL, PostgreSQL, Ansible, Terraform, Kubernetes, Docker, Python, Go, Java, Scala, Prometheus, Grafana, ELK Stack, Aerospike, Snowflake
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
San Antonio or McLean or United States
$80k-$133k/yr HybridFull Time
Guidehouse
Guidehouse: Provides management and technology consulting services to diverse organizations.
4+ YOEBA/BS or equivalent experience, 4+ years IT/admin/software/platform experience with AWS, 1+ years cloud deployment experience, proficiency with CI/CD and IaC tools (Terraform, Ansible, GitLab, Artifactory, Packer), scripting (Python, PowerShell, Bash), Windows/Linux, Agile, and ability to obtain Public Trust.
CI/CD, IaC, Terraform, Ansible Automation Platform, GitLab, Artifactory, Packer, Python, PowerShell, Bash, Windows, Linux, SDLC, Scrum, Kanban, SAFe
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Boston or Miami or New Jersey or New York City or Princeton or Raleigh or Washington or Toronto or North America
$127k-$249k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud (AWS, Google Cloud Platform, Microsoft Azure) expertise; Linux and networking knowledge.
Argo Workflows, ArgoCD, Kubernetes, Python, Go, AWS, Google Cloud Platform (GCP), Microsoft Azure, Linux
2w
Save
Mark Applied
Hide
Staff Software Engineer, Platform
San Francisco or New York or Washington
$240k-$300k/yr OnsiteFull Time
Peregrine
Peregrine: Data integration and analytics platform for public safety agencies.
8+ YOE8+ years building and operating cloud infrastructure, hands-on ownership of platform/networking/traffic systems, strong security and reliability experience, on-call and incident response experience, degree or equivalent.
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (SRE) / Service Availability Manager
Bethesda, Maryland, United States
$96k-$145k/yr HybridFull Time
Marriott International
Marriott InternationalNASDAQ: MAR: Operates and franchises a global network of hotels and resorts.
5+ YOE5+ years IT experience, 3+ years IT operations and incident/change/release management, undergraduate degree or equivalent, on-call/24x7 availability, proficiency with Python and Shell, familiarity with Ansible, Jenkins, cloud platforms, IaC and containers.
Python, Shell, Ansible, Jenkins, AWS, Azure, GCP, ServiceNow
1mo
Save
Mark Applied
Hide
Software Engineer, Agent Platform
Washington or Costa Mesa
$220k-$292k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology building autonomous military hardware and software.
Strong backend engineering experience building production platforms; deep expertise in LLM agent framework design and evaluation; familiarity with model post-training workflows (SFT, RL); emphasis on reliability and partnering with ML/product teams.
Lattice OS, Langchain, Deepagents, Claude SDK, Kubernetes, Docker
3d
Save
Mark Applied
Hide
Cloud Platforms Engineer
Fairfax, Virginia, United States
$180k-$210k/yr HybridFull Time
ECS
ECSNYSE: ASGN: Provides advanced technology and engineering services to government agencies.
5+ YOEBachelor’s degree and 5+ years of related experience required, with active DoD Secret clearance, cloud infrastructure, Terraform, automation, reliability, disaster recovery, and troubleshooting expertise.
Microsoft Azure Government, Terraform, Infrastructure as Code (IaC), Site Reliability Engineering, DevOps
1w
Save
Mark Applied
Hide
Cloud Engineering – Site Reliability – Infrastructure/Application
Fairfax, Virginia, United States
$90k-$198k/yr HybridFull Time
CGI
CGINYSE: GIB: Provides information technology and business consulting services.
3+ YOEBachelor's degree in computer science, information technology, or related field; 3+ years in infrastructure automation and cloud technologies; IaC, cloud platforms, CI/CD, scripting, and communication skills.
Terraform, Ansible, CloudFormation, Google Cloud Platform, AWS, Azure, GitLab Ultimate, Python, Bash, PowerShell, AWS API Gateway, Azure DevOps, GitLab
1mo
Save
Mark Applied
Hide
Software Development Engineer 4
Portland or San Francisco or Washington or Boston
$141k-$173k/yr OnsiteFull Time
WEX
WEXNYSE: WEX: Provides global payment processing and business information management services.
Staff/lead level engineer with deep experience architecting autonomous/agentic AI systems, cross-platform architecture, security-by-design, CI/CD, cloud reliability, and mentoring senior engineers.
Helm Charts, Argo Workflows, CI/CD, LLMs
2mo
Save
Mark Applied
Hide
Senior Staff, Software Engineer
Sterling, Virginia, United States
HybridFull Time
Asurion: Provides insurance and repair services for consumer electronics.
10+ YOE10+ years building APIs, backend services, distributed systems and data-intensive platforms; strong Node.js and TypeScript experience; data modeling, database design, API design, observability, and reliability engineering.
Node.js, TypeScript, PostgreSQL, MySQL, DynamoDB, MongoDB, Redis, Elasticsearch/OpenSearch, Neo4j, Kafka, CDC, CI/CD, schema registry
2w
Save
Mark Applied
Hide
Senior Systems Development Engineer L4
Bowie, Maryland, United States
$94k-$131k/yr OnsiteFull Time
Inovalon
Inovalon: Cloud platform for data-driven healthcare analytics and insights.
5+ YOE5+ years cloud/systems engineering experience; expertise in AWS/Azure/GCP/OCI/Snowflake, infrastructure as code, scripting, and platform governance; bachelor's degree required; strong communication and reliability focus.
AWS, Azure, GCP, OCI, Snowflake, Microsoft Azure DevOps, Microsoft Azure Pipelines, Terraform, Linux Shell, Microsoft PowerShell, Python, Microsoft Active Directory, GPO, RDS, Microsoft Windows, Linux
1mo
Save
Mark Applied
Hide
Senior Lead Software Engineer-AI Foundation Services
Plano or Jersey City or Wilmington or McLean
$171k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years software engineering experience building cloud-native AI/ML platform services with Kubernetes, CI/CD, and infrastructure-as-code; proficiency in Python/Java/Go; strong production reliability and secure-by-design practices.
Kubernetes, CI/CD, infrastructure-as-code, Python, Java, Go, GPU