76 ai infrastructure engineer jobs at 60 companies in Springfield, VA

1mo
Save
Mark Applied
Hide
AI Infrastructure Engineer - Emerging Technologies
Ashburn or Dallas or Columbus
RemoteFull Time
Cologix
Cologix: Provides network-neutral interconnection and hyperscale edge data center services.
5+ YOEBachelor's in engineering or related field, 5+ years in data center/AI/HPC/power or cooling engineering, strong understanding of AI compute and infrastructure, ability to translate technical analysis into infrastructure strategy, authorized to work in the U.S. (no sponsorship).
NVIDIA GPU architectures, ARM, Battery Energy Storage Systems (BESS), UPS
4w
Save
Mark Applied
Hide
Staff AI Infrastructure Engineer
Austin or Reston
HybridFull Time
Seekr
Seekr: Transparent AI platform for enterprise and government decision-making.
8+ YOE8+ years building distributed systems and cloud-native AI infrastructure; strong Python and systems programming skills; Kubernetes, GPU inference, and platform engineering experience; leadership and architecture experience.
Python, Go, Rust, C++, vLLM, SGLang, TensorRT-LLM, Triton Inference Server, Ray Serve, Kubernetes, Helm, Argo CD, Docker, Prometheus, Grafana, OpenTelemetry, Infrastructure-as-Code, GitOps, CI/CD, AWS, Azure, Oracle Cloud Infrastructure, Google Cloud Platform
1mo
Save
Mark Applied
Hide
AI Infrastructure & Security Engineer
Alexandria, Virginia, United States
$110k-$140k/yr RemoteFull Time
HRCI
HRCI: Provides global certifications and learning for HR professionals.
3+ YOE3+ years software engineering and Azure experience, CI/CD and DevOps with Azure DevOps/GitHub, containerization, Python scripting, production AI coding agent experience, security-focused instincts, and ability to communicate across teams.
Claude, Claude Code, CLAUDE.md, Cowork, Azure, Azure Container Apps, Azure DevOps, Azure Repos, GitHub, CI/CD, npm, NuGet, Office 365, eCommerce, CRM, Python, Windows Server, Linux, DNS, TCP/IP, firewalls, VPNs, MCP
1w
Save
Mark Applied
Hide
Principal Platform Engineer, AI & Infrastructure
Denver or Long Beach or Washington
$255k-$375k/yr HybridFull Time
True Anomaly
True Anomaly: Develops autonomous spacecraft and software for space security missions.
12+ YOE12+ years in software/DevOps/SRE/cloud with deep AI platform and cloud infrastructure expertise; ability to obtain Top-Secret clearance; hands-on with LLM APIs, agent frameworks, and IaC.
CrewAI, Pydantic AI, Terraform, OpenTelemetry, AWS, Azure, GCP
1mo
Save
Mark Applied
Hide
Gen AI Infrastructure Engineer
Woodbridge or New York City or Atlanta or Boston or Chicago or Dallas or Delaware or Denver or Garden City or Cayman Islands or Greenwich or Houston or Los Angeles or Miami or Naples or Nevada or Palm Beach or San Diego or San Francisco or Seattle or Stuart or Washington
$160k-$200k/yr HybridFull Time
Bessemer Trust
Bessemer Trust: Wealth management and family office services for affluent clients.
7+ YOE7+ years in DevOps/platform/cloud infrastructure engineering; strong AWS (IAM, CloudFormation, Lambda, API Gateway, VPC, CloudWatch, SSM, Secrets Manager, ECR); AWS CDK/CloudFormation/Terraform; CI/CD (Bitbucket/GitHub Actions, OIDC); container and datastore operations; security fundamentals.
AWS Bedrock, AgentCore, Lambda, API Gateway, AWS CDK, CloudFormation, Terraform, Bitbucket Pipelines, GitHub Actions, OIDC, IAM, VPC, CloudWatch, SSM, Secrets Manager, ECR, Neo4j, Neptune, Redis, Milvus, VectorDB
1mo
Save
Mark Applied
Hide
AI Operations & Infrastructure Engineer
Fort Meade, Maryland, United States
OnsiteFull Time
Invictus International Consulting
Invictus International Consulting: Provides cybersecurity and technology services for national security systems.
Hands-on experience administering NVIDIA GPU and DPU technologies, AI software stacks, containerization and orchestration (Docker, Kubernetes, Slurm), NVIDIA certifications, and an active TS/SCI clearance with CI Polygraph.
GPUs, DPUs, InfiniBand, Ethernet, Docker, Kubernetes, Slurm, NVIDIA Base Command Manager, Pyxis, Enroot, Run:Ai, NGC CLI, HPL, NCCL, NVIDIA Nemo, ClusterKit, BlueField, MIG, BMC, TPM
2mo
Save
Mark Applied
Hide
Sr. AI Engineer, Platform Infrastructure, Special Programs
Palo Alto or Washington
$220k-$350k/yr HybridFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOEBachelor's in CS/engineering; 5+ years Go/Python; 5+ years Kubernetes or similar tooling; ability to obtain Top Secret/SCI clearance.
Go, Python, Kubernetes, Pulumi, Terraform, Linux, CI/CD, Build tooling, NVIDIA CUDA, NVIDIA drivers
1mo
Save
Mark Applied
Hide
Senior Lead AI Engineer (GenAI Platform, Agentic Infrastructure)
New York or San Francisco or McLean or Cambridge or San Jose or Plano
$209k-$286k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's in CS/AI/EE/CE +6 years or Master's +4 years; 6+ years programming with Python/Go/Scala/Java; experience deploying scalable AI on cloud; LLM, inference, similarity search, VectorDBs, guardrails, model evaluation, and optimization experience; leadership and research literacy.
Python, Go, Scala, Java, C++, C#, Golang, AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, AWS, Google Cloud, Azure
4w
Save
Mark Applied
Hide
Senior AI Architect/Engineer
Arlington or Washington or Pentagon or Springfield or Chantilly or Tysons Corner
$105k-$200k/yr OnsiteFull Time
Technomics
Technomics: Provides cost analysis and decision support services for government.
8+ YOE8+ years in cloud architecture or AI/ML infrastructure, strong Azure experience, knowledge of RAG, embeddings, vector DBs, model APIs, containerized deployments, security and governance for regulated clouds.
Microsoft Azure, Azure GCC High, RAG, embeddings, vector databases, model APIs, containers, APIs
2mo
Save
Mark Applied
Hide
Senior AI ML Engineer - Remote
San Francisco or Minneapolis or Washington, D.C. or United States
$120k-$215k/yr RemoteFull Time
UnitedHealth Group
UnitedHealth GroupNYSE: UNH: Provides health insurance and technology-enabled health care services.
4+ YOE2+ MgmtBachelor's degree or 4+ years equivalent, 4+ years Python, 4+ years cloud infrastructure (AWS/Azure/GCP), 4+ years AI/ML infrastructure experience, 2+ years team lead, 1+ year LLM experience.
Python, AWS, Azure, GCP, Large Language Models (LLMs), GitHub, GitHub Actions, Docker, Terraform, CI/CD
1mo
Save
Mark Applied
Hide
Network Engineer, AI Infrastructure Repair
Sarpy County or Mesa or Aiken or Ashburn or Forest City or Menlo Park or Polk County or Montgomery or Prineville or Fort Worth or Richmond or Chandler or New Albany or Huntsville or Rosemount or Eagle Mountain or Crook County or Houston or United States
$193k-$271k/yr OnsiteFull Time
Meta
MetaNASDAQ: META: Develops social networking platforms and virtual reality technologies.
12+ YOEBachelor's in CS or equivalent, 12+ years network engineering experience at data center/HPC scale, expertise in high-speed fabrics (RDMA, InfiniBand, optical interconnects), program leadership in network fault management and repair automation.
RDMA over Converged Ethernet, InfiniBand, optical interconnects
2d
Save
Mark Applied
Hide
Senior AI Engineer
Tysons, Virginia, United States
RemoteFull Time
Cellebrite
CellebriteNASDAQ: CLBT: Provider of digital forensic and investigative intelligence software solutions.
Design and implement AWS-based AI infrastructure and services for ML workflows; strong Python and Typescript, AWS (SageMaker, Lambda, Bedrock), ML libraries, GitHub Actions, and observability experience required.
Python, Typescript, GitHub Actions, CDK, SageMaker, Lambda, Bedrock, PyTorch, Pandas, NumPy, GitHub, Amazon EKS, Kubernetes, Datadog
3w
Save
Mark Applied
Hide
Software Engineer (AI Infrastructure)
Columbia, Maryland, United States
OnsiteFull Time
BigBear.ai
BigBear.aiNYSE: BBAI: Provides AI-powered decision intelligence software for national security.
8+ YOETS/SCI w/Poly clearance, 8+ years experience (or BS+4+), AWS, Kubernetes, Python, observability (OpenTelemetry, Grafana, Prometheus), CI/CD, IaC, systems integration and production-scale infrastructure experience.
AWS, Kubernetes, Python, OpenTelemetry, Grafana, Prometheus, vLLM, LiteLLM, LangChain, Infrastructure-as-Code (IaC), APM, CI/CD, Retrieval Augmented Generation (RAG)
2mo
Save
Mark Applied
Hide
AI Agent & Infrastructure - Lead Engineer
Rockville or Tysons Corner
HybridFull Time
FINRA
FINRA: Regulates brokerage firms and exchange markets in the United States.
7+ YOE7+ years in software engineering; strong OO design; cloud, DevOps, CI/CD; AI agent, SDLC, observability; Java, Python, JavaScript/TypeScript; AWS; experience with monitoring tools.
Java, Python, JavaScript/TypeScript, Angular, Spring Boot, AWS, Prometheus, Grafana, CloudWatch
3w
Save
Mark Applied
Hide
Software Engineer (AI Infrastructure)
Columbia, Maryland, United States
OnsiteFull Time
BigBear.ai
BigBear.aiNYSE: BBAI: AI-powered decision intelligence for complex mission environments.
8+ YOE8+ years experience (or Bachelor's +4 years), TS/SCI with polygraph, proven production systems experience, AWS cloud engineering, Kubernetes, Python, observability (OpenTelemetry, Grafana, Prometheus), CI/CD familiarity.
AWS, Kubernetes, Python, OpenTelemetry, Grafana, Prometheus, vLLM, LiteLLM, LangChain, APM
2mo
Save
Mark Applied
Hide
Software Engineer (AI Infrastructure)
Laurel, Maryland, United States
$115k-$160k/yr OnsiteFull Time
Visionist
Visionist: Provides software engineering and data analytics for national security.
8+ YOE8+ years of software development; production systems; cloud AWS; Kubernetes; Python; observability; CI/CD; strong communication.
AWS, Kubernetes, Python, OpenTelemetry, Grafana, Prometheus, CI/CD, Infrastructure as Code
1d
Save
Mark Applied
Hide
Member of Technical Staff - AI Cloud Infrastructure
Oakland or Boston or Washington or California or Massachusetts or District of Columbia
HybridFull Time
Emerald AI
Emerald AI: Managing data center power with AI-driven workload orchestration.
7+ YOERequires 7+ years in infrastructure or platform engineering, managed cloud or AI platform architecture, Kubernetes, Slurm, parallel filesystems, Linux, Terraform, Ansible, Python or Go, and GPU networking.
Kubernetes, Slurm, Lustre, GPFS, Weka, VAST, BeeGFS, Linux, Terraform, Ansible, Python, Go, InfiniBand, RoCE, RDMA, NVIDIA SuperPOD, GPUDirect Storage, NCCL, DCGM, S3, Ceph, MinIO, Emerald Conductor
1mo
Save
Mark Applied
Hide
Full Stack AI Software Engineer
Annapolis Junction, Maryland, United States
OnsiteFull Time
Tiber Technologies: Providing specialized engineering and security services for federal missions.
U.S. citizen with active clearance and polygraph; bachelor\u0002s in CS/engineering; experienced in designing, implementing, and operating production-grade AI infrastructure.
1mo
Save
Mark Applied
Hide
AI Solutions Architect - East Region
Hartford or Washington or Tampa or Atlanta or Boston or Charlotte or Trenton or New York or Philadelphia or Pittsburgh
$185k-$235k/yr OnsiteFull Time
World Wide Technology
World Wide Technology: Global technology solutions provider and systems integrator.
10+ YOESenior technical pre-sales architect with 10+ years of experience designing AI/ML infrastructure and platforms, hands-on NVIDIA and cloud experience, strong presentation and whiteboarding skills, and a bachelor’s degree in computer science, engineering, or related field.
NVIDIA DGX, NVIDIA HGX, CUDA, NVIDIA AI Enterprise (NVAIE), NeMo, Omniverse, AWS, Microsoft Azure, GCP, Dell, HPE, Cisco, NetApp, Pure Storage, Vast Data, Advanced Technology Center (ATC), MLOps
3w
Save
Mark Applied
Hide
AI Native Software Engineer II
Seattle or Arlington
$144k-$180k/yr HybridFull Time
Remitly
RemitlyNASDAQ: RELY: Provides digital cross-border money transfer services for global customers.
Proficiency with AI coding agents and prompt engineering, full-stack application knowledge (frontend, backend, data, infrastructure), familiarity with web architectures, ability to read usage analytics, and strong communication skills.
AI coding agents