63 platform reliability engineer jobs at 18 companies in Camp Springs, MD

3w
Save
Mark Applied
Hide
Reliability Engineer
Arlington, Virginia, United States
$80k-$160k/yr OnsiteFull Time
Decision Technologies
Decision Technologies: Private U.S. defense contractor providing engineering, intelligence, program management, and technical support to government customers.
4+ YOEBachelor's degree in a technical discipline, 4+ years RM&A experience on US Navy platforms or weapon systems, eligible for DoD Secret clearance, strong communication and presentation skills.
SIPR
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: Private supply-chain AI software serving corporations, government agencies, and banks with risk and compliance technology.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
1mo
Save
Mark Applied
Hide
Platform Delivery & Reliability Engineer (Remote) - 29337
Colorado Springs or Columbia or San Antonio or Boise or Greenville or Augusta
$120k-$195k/yr RemoteFull Time
Mission Technologies
Mission TechnologiesNYSE: HII: Defense technology division delivering integrated all-domain solutions to defense, federal, and commercial customers.
7+ YOEMust obtain U.S. security clearance; 7+ years' relevant experience (varies by degree); deep Kubernetes, IaC, cloud, CI/CD, scripting, SRE practices, troubleshooting and delivery experience across distributed systems.
Kubernetes, Terraform, AWS, Azure, GCP, GitLab CI, Go, Python, Bash, Spark, Trino, Presto, Kafka, NiFi, Iceberg, Delta, Hudi, YouTrack, Nexus
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
McLean, Virginia, United States
$129k-$190k/yr OnsiteFull Time
Medallia
Medallia: Private SaaS provider of customer, employee, citizen, and patient experience-management software for enterprises.
5+ YOERequires 5+ years leading reliability, platform, infrastructure, or cloud operations initiatives; Kubernetes, cloud platforms, Linux, Python/Go/Bash, Terraform, CI/CD, GitOps, networking, incident response, and on-call experience.
Kubernetes, AWS, OCI, GCP, Linux, Python, Go, Bash, Terraform, CI/CD, GitOps, DNS, TLS/SSL, ArgoCD, Prometheus, Grafana, Loki, OpenTelemetry
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
McLean, Virginia, United States
$129k-$190k/yr OnsiteFull Time
Medallia
Medallia: Private SaaS provider of customer, employee, citizen, and patient experience-management software for enterprises.
5+ YOERequires 5+ years leading reliability, platform, infrastructure, or cloud operations initiatives; production-scale Kubernetes, cloud, Linux, Python/Go/Bash, Terraform, CI/CD, GitOps, networking, and incident response experience.
Kubernetes, AWS, OCI, GCP, Linux, Python, Go, Bash, Terraform, CI/CD, GitOps, ArgoCD, Prometheus, Grafana, Loki, OpenTelemetry
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer – Managed Patching
Manassas, Virginia, United States
$122k-$226k/yr OnsiteFull Time
Swift
Swift: Member-owned Belgian cooperative providing secure financial messaging services to banks, financial institutions, and corporates worldwide.
12+ YOE12+ years in SRE/DevOps/platform engineering; experience building enterprise-scale automation, Ansible, CI/CD, Linux/RHEL, ServiceNow integrations, Python, and strong cross-team leadership.
Ansible Automation Platform, CloudBees, Python, ServiceNow, Git, Microsoft Power BI, Linux, Red Hat Enterprise Linux (RHEL)
2mo
Save
Mark Applied
Hide
Senior DevOps/Platform Engineer
San Francisco or San Jose or Seattle or Los Angeles or San Diego or Portland or Reno or Ontario or Bakersfield or Phoenix or Riverside or Sacramento or Dallas or Houston or Irving or Chicago or Cleveland or Columbus or Cincinnati or Detroit or Newark or Dulles or Indianapolis or Austin or Brooklyn or Las Vegas or Kansas City or Philadelphia or Pensacola or South San Francisco or St. Louis
$120k-$147k/yr HybridFull Time
Jitsu
Jitsu: Privately held U.S. last-mile delivery provider serving e-commerce brands and high-volume shippers.
5+ YOE5+ years as a DevOps or Site Reliability Engineer, with deep CI/CD, cloud infrastructure, Terraform, GitOps, Kubernetes, Helm, databases, monitoring, reliability, and infrastructure security experience.
Jenkins, GitHub Actions, Terraform, ArgoCD, Google Cloud Platform, GCP, Kubernetes, Helm, PostgreSQL, CloudSQL, MongoDB, Cassandra, Redis, IAM, AIOps, AI, LLMs, Java
2w
Save
Mark Applied
Hide
Data Platform Engineer - AI Platform
United States or North America or San Francisco or Los Angeles or New York City or Washington or London or Singapore
RemoteFull Time
TRM Labs
TRM Labs: Blockchain intelligence for detecting crypto-related financial crime.
U.S. citizenship, distributed OLAP or serving-layer operations experience, query tuning, data pipeline reliability, incident response, AI tool fluency, independent infrastructure ownership, and on-call readiness.
StarRocks, Claude, Trino, ClickHouse, Cursor, Slack, Otter.ai, Fireflies, Fathom, Cluey
1w
Save
Mark Applied
Hide
Principal Platform Engineer
Herndon or Columbia
OnsiteFull Time
Clarity Innovations
Clarity Innovations: Private national-security software and data engineering serving U.S. defense, intelligence, and federal agencies.
Requires Terraform, GitLab CI/CD, containerized and serverless workloads, databases, Linux infrastructure, Docker, Podman, Kubernetes, and secure, reliable infrastructure engineering.
Terraform, GitLab CI/CD, Docker, Podman, Kubernetes, Airflow, EMR, Ansible, Puppet, Chef, Helm, ArgoCD
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
San Antonio or McLean or United States
$80k-$133k/yr HybridFull Time
Guidehouse
Guidehouse: Global consulting firm serving public and commercial sectors.
4+ YOEBA/BS or equivalent experience, 4+ years IT/admin/software/platform experience with AWS, 1+ years cloud deployment experience, proficiency with CI/CD and IaC tools (Terraform, Ansible, GitLab, Artifactory, Packer), scripting (Python, PowerShell, Bash), Windows/Linux, Agile, and ability to obtain Public Trust.
CI/CD, IaC, Terraform, Ansible Automation Platform, GitLab, Artifactory, Packer, Python, PowerShell, Bash, Windows, Linux, SDLC, Scrum, Kanban, SAFe
3d
Save
Mark Applied
Hide
Principal Site Reliability Engineer - Remote
United States or Eden Prairie or Minnesota or Washington
$135k-$231k/yr RemoteFull Time
UnitedHealth Group
UnitedHealth GroupNYSE: UNH: Diversified health care helping people live healthier lives.
10+ YOERequires 10+ years in software, platform, DevOps, or SRE roles; 3+ years in senior technical leadership; 5+ years with cloud and containers; and experience with observability and production automation.
Azure, AWS, OpenTelemetry, Prometheus, Grafana, Datadog, Terraform, Pulumi, Ansible, Helm, Kubernetes, RAG
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Boston or Miami or New Jersey or New York City or Princeton or Raleigh or Washington or Toronto or North America
$127k-$249k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Unified data platform for building modern applications.
6+ YOE6+ years software development experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud (AWS, Google Cloud Platform, Microsoft Azure) expertise; Linux and networking knowledge.
Argo Workflows, ArgoCD, Kubernetes, Python, Go, AWS, Google Cloud Platform (GCP), Microsoft Azure, Linux
3w
Save
Mark Applied
Hide
Staff Software Engineer, Platform
San Francisco or New York or Washington
$240k-$300k/yr OnsiteFull Time
Peregrine Technologies
Peregrine Technologies: The leading data integration platform helping public safety agencies make better decisions in the moments that matter.
8+ YOE8+ years building and operating cloud infrastructure, hands-on ownership of platform/networking/traffic systems, strong security and reliability experience, on-call and incident response experience, degree or equivalent.
2mo
Save
Mark Applied
Hide
Software Engineer, Agent Platform
Washington or Costa Mesa
$220k-$292k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology developing AI-powered autonomous military systems.
Strong backend engineering experience building production platforms; deep expertise in LLM agent framework design and evaluation; familiarity with model post-training workflows (SFT, RL); emphasis on reliability and partnering with ML/product teams.
Lattice OS, Langchain, Deepagents, Claude SDK, Kubernetes, Docker
1w
Save
Mark Applied
Hide
Cloud Platforms Engineer
Fairfax, Virginia, United States
$180k-$210k/yr HybridFull Time
ECS Federal
ECS FederalEFOR: Federal segment of Everforth delivering technology and engineering solutions.
5+ YOEBachelor’s degree and 5+ years of related experience required, with active DoD Secret clearance, cloud infrastructure, Terraform, automation, reliability, disaster recovery, and troubleshooting expertise.
Microsoft Azure Government, Terraform, Infrastructure as Code (IaC), Site Reliability Engineering, DevOps
2w
Save
Mark Applied
Hide
Cloud Engineering – Site Reliability – Infrastructure/Application
Fairfax, Virginia, United States
$90k-$198k/yr HybridFull Time
CGI Federal
CGI Federal: U.S. technology and professional services serving federal agencies with IT consulting and mission solutions.
3+ YOEBachelor's degree in computer science, information technology, or related field; 3+ years in infrastructure automation and cloud technologies; IaC, cloud platforms, CI/CD, scripting, and communication skills.
Terraform, Ansible, CloudFormation, Google Cloud Platform, AWS, Azure, GitLab Ultimate, Python, Bash, PowerShell, AWS API Gateway, Azure DevOps, GitLab
2mo
Save
Mark Applied
Hide
Software Development Engineer 4
Portland or San Francisco or Washington or Boston
$141k-$173k/yr OnsiteFull Time
WEX
WEXNYSE: WEX: Public fintech and payments helping businesses manage fleet fuel, employee benefits, and corporate payments.
Staff/lead level engineer with deep experience architecting autonomous/agentic AI systems, cross-platform architecture, security-by-design, CI/CD, cloud reliability, and mentoring senior engineers.
Helm Charts, Argo Workflows, CI/CD, LLMs
2mo
Save
Mark Applied
Hide
Senior Staff, Software Engineer
Sterling, Virginia, United States
HybridFull Time
Asurion
Asurion: Global provider of device protection, insurance, and technical support services.
10+ YOE10+ years building APIs, backend services, distributed systems and data-intensive platforms; strong Node.js and TypeScript experience; data modeling, database design, API design, observability, and reliability engineering.
Node.js, TypeScript, PostgreSQL, MySQL, DynamoDB, MongoDB, Redis, Elasticsearch/OpenSearch, Neo4j, Kafka, CDC, CI/CD, schema registry
3w
Save
Mark Applied
Hide
Senior Systems Development Engineer L4
Bowie, Maryland, United States
$94k-$131k/yr OnsiteFull Time
Inovalon
Inovalon: Healthcare software serving payers, providers, pharmacies, and life sciences organizations with data and analytics.
5+ YOE5+ years cloud/systems engineering experience; expertise in AWS/Azure/GCP/OCI/Snowflake, infrastructure as code, scripting, and platform governance; bachelor's degree required; strong communication and reliability focus.
AWS, Azure, GCP, OCI, Snowflake, Microsoft Azure DevOps, Microsoft Azure Pipelines, Terraform, Linux Shell, Microsoft PowerShell, Python, Microsoft Active Directory, GPO, RDS, Microsoft Windows, Linux
1mo
Save
Mark Applied
Hide
Senior Lead Software Engineer-AI Foundation Services
Plano or Jersey City or Wilmington or McLean
$171k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services and investment banking firm.
5+ YOE5+ years software engineering experience building cloud-native AI/ML platform services with Kubernetes, CI/CD, and infrastructure-as-code; proficiency in Python/Java/Go; strong production reliability and secure-by-design practices.
Kubernetes, CI/CD, infrastructure-as-code, Python, Java, Go, GPU