Normal Computing
Posted 2mo ago

Infrastructure Software Engineer

Normal Computing
New York City or Copenhagen or London or San Francisco
$185k-$285k/yrHybridFull Time
Responsibilities
  • building infrastructure
  • designing APIs
  • improving reliability
Requirements
  • 4+ years in infrastructure or backend systems
  • Strong backend engineering, APIs
  • Distributed systems
  • Observability
  • Practical Docker/Kubernetes and persistence (Postgres
  • Redis) experience
  • Ownership and cross-team communication
Technical tools mentioned
DockerKubernetesPostgresRedis/Valkeyobject storageEDA

Job description

Normal Computing | Incredible Opportunities

The Normal Team builds foundational software and hardware that help move technology forward, supporting the semiconductor industry, critical AI infrastructure, and the broader systems that power our world. We work as one team across New York, San Francisco, Copenhagen, and London.

Your Role in Our Mission

We’re looking for an Infrastructure Software Engineer to build the production systems behind Normal’s AI products.

This is an application engineering role focused on infrastructure-shaped software: orchestration services, execution runtimes, internal APIs, persistence layers, observability, and developer experience. You’ll help define the runtime layer for a new class of AI products: systems where agents execute long-running work, coordinate across distributed environments, interact with code and tools, and need to be reliable enough for real customer workflows.

This role sits between product engineering, AI engineering, and platform engineering. You will not primarily be managing Terraform, Helm charts, CI/CD, or company-wide SaaS infrastructure. Instead, you’ll own the application-level infrastructure that powers long-running AI workflows: session lifecycle, sandboxed execution, workload orchestration, persistence, observability, reliability, and the internal interfaces other engineers build on.

The systems you build will be used directly by product, AI, research, and platform teams as new capabilities move from early ideas into production. Developer experience matters: APIs should be understandable, failure modes should be debuggable, and abstractions should make the right thing easy.

This is a highly cross-functional role for someone who enjoys ambiguity, cares about clean abstractions, and wants to help shape how a frontier AI company builds and operates production systems. Strong engineering judgment and ownership matter more than rigid specialization.

On any given day, you might design the runtime architecture for a new AI product capability, build the orchestration layer for long-running autonomous workflows, improve how workloads are scheduled and isolated across distributed environments, or create the systems abstractions that let engineers turn ambitious AI prototypes into reliable production products.

Responsibilities

  • Build and maintain production software infrastructure for Normal’s AI products, especially orchestration, execution, and runtime systems.

  • Design internal backend services and APIs used by product engineers, AI engineers, execution services, and other internal systems.

  • Improve the operational maturity of rapidly evolving systems through better state management, failure handling, metrics, tracing, and debugging tools.

  • Work with Kubernetes-backed execution environments, including container lifecycle, scheduling behavior, autoscaling, resource isolation, and runtime reliability.

  • Build developer-facing tools and abstractions that make it easier for other engineers to use and extend the systems you own.

  • Turn promising prototypes into durable production systems by designing clear abstractions, hardening critical paths, and creating operational patterns that scale with the product.

  • Collaborate closely with product, AI, research, and platform engineers to define the right interfaces between product features, AI workloads, and production infrastructure.

  • Lead design discussions for core runtime and orchestration systems, including API boundaries, state management, execution models, and operational tradeoffs.

What We’re Looking For

  • 4+ years of experience in infrastructure software, backend infrastructure, production infrastructure, platform engineering, distributed systems, or related areas.

  • Strong software engineering fundamentals, including backend programming, APIs, data modeling, concurrency, debugging, and testing.

  • Experience building or operating production services where reliability, observability, and maintainability matter.

  • Practical experience with Docker and Kubernetes, including debugging containerized workloads, deployments, networking, resource limits, and lifecycle issues.

  • Comfort working with persistence systems such as Postgres, Redis/Valkey, object storage, or similar production data stores.

  • Experience building orchestration systems, job schedulers, workflow engines, sandboxes, developer platforms, or distributed execution systems.

  • Experience designing internal APIs and developer-facing abstractions that other engineers can use confidently.

  • Strong systems thinking: you can reason about state machines, failure modes, retries, queues, leases, scheduling, and long-running workflows.

  • Pragmatism in fast-moving environments: you know when to improve an abstraction, when to delete one, and when to ship the simple version.

  • Ownership mindset: you care whether the systems you build work in production and are usable by other engineers.

  • Clear communication and good technical judgment across product, AI, and infrastructure boundaries.

Nice to Have

  • Deep Kubernetes experience, such as controllers/operators, networking, storage, scheduling, autoscaling, or resource isolation.

  • Experience with AI agent infrastructure, ML infrastructure, model orchestration, or LLM-based product systems.

  • Background in production infrastructure, reliability engineering, or infrastructure software at meaningful scale.

  • Experience in high-growth startups or engineering teams where ownership boundaries are still being defined.

  • Experience with Chips, EDA or Device Verification

Equal Employment Opportunity Statement

Normal Computing is an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, veteran status, or any other legally protected status.

Accessibility Accommodations

Normal Computing is committed to providing reasonable accommodations to individuals with disabilities. If you need assistance or an accommodation due to a disability, please let us know at [email protected].

Privacy Notice

By submitting your application, you agree that Normal Computing may collect, use, and store your personal information for employment-related purposes in accordance with our Privacy Policy.

About Normal Computing

Applied AI and silicon hardware/software serving semiconductor design and manufacturing institutions.

Similar jobs

Infrastructure Software Engineer roles near New York City, New York
1d
Save
Mark Applied
Hide
Software Engineer, Infrastructure
New York City, New York, United States
$160k-$300k/yr OnsiteFull Time
Outtake
Outtake: Private cybersecurity using autonomous AI agents to detect and dismantle digital impersonation threats for enterprises.
Staff-scope experience, distributed infrastructure design, production scaling, relational databases, asynchronous systems, search and retrieval, large-scale data processing, and cloud infrastructure.
1d
Save
Mark Applied
Hide
Senior Infrastructure Software Engineer
New York City or San Francisco or Seattle or London
$180k-$220k/yr HybridFull Time
Lightning AI
Lightning AI: Private AI software and GPU-cloud helping developers and enterprises build, train, deploy, and run AI systems.
8+ YOERequires 8+ years of software or infrastructure engineering experience, production backend development in Python or similar languages, Linux expertise, scalable infrastructure automation, APIs, containerization, orchestration, HPC, and bare-metal fundamentals.
PyTorch Lightning, Python, Linux, PXE, iPXE, BMC, Redfish, IPMI, Dell, GPU, SONiC, Palo Alto, Juniper Networks, VAST
5d
Save
Mark Applied
Hide
Software Engineer, Infrastructure
New York City, New York, United States
$155k-$185k/yr HybridFull Time
Sixfold
Sixfold: Private insurtech developing AI underwriting software for insurers, managing general agents, and reinsurers.
3+ YOE3+ years of software engineering experience, including 2+ years in infrastructure, platform engineering, DevOps, or SRE; production Kubernetes, Terraform, cloud, CI/CD, on-call, and programming experience required.
Terraform, Azure, AWS, Kubernetes, Helm, Argo CD, GitHub Actions, Datadog, Codex, Claude Code
6d
Save
Mark Applied
Hide
Senior Software Infrastructure Engineer
New York City, New York, United States
$160k-$250k/yr OnsiteFull Time
Ultra
Ultra: Private AI networking platform that connects users with-founders, investors, experts, and talent.
7+ YOE7+ years in infrastructure or security engineering, production authn/authz ownership, Python proficiency, SOC2 compliance experience, scalable data infrastructure design, and secure AI tooling expertise.
Python, SOC2, LLMs, RBAC
1w
Save
Mark Applied
Hide
Software Engineer, Infrastructure
New York City or San Francisco
$160k-$300k/yr OnsiteFull Time
Hebbia
Hebbia: Private generative AI platform helping finance, legal, and professional-services teams research, analyze, and create documents.
5+ YOE5+ years building production cloud infrastructure, deep IaC expertise, AWS networking/IAM/account architecture, container orchestration, CI/CD, scripting in Python or Go, observability, security, and compliance experience.
AWS, Python, Go, CI/CD, IAM
2w
Save
Mark Applied
Hide
Infrastructure Software Engineer, Fleet & Automation
Houston or New York City or San Francisco or Seattle
OnsiteFull Time
Nscale
Nscale: Vertically integrated AI infrastructure and GPU cloud platform.
5+ YOEBachelor's degree or equivalent experience, 5+ years building large-scale infrastructure applications, and expertise in Python, Linux, networking, distributed systems, and infrastructure tooling.
C, C++, Java, Python, Linux, TCP/IP, BGP, Ansible, Terraform, DCIMs, NetBox, OpenStack, MAAS, Ironic, IPMI, NVIDIA GPUs, InfiniBand, NCCL, SLURM, Prometheus, Grafana, OpenTelemetry, Kubernetes, Docker
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Infrastructure
New York, New York, United States
$175k-$215k/yr OnsiteFull Time
CLEAR
CLEARNYSE: YOU: Provides biometric identity verification and expedited travel security services.
6+ YOE6+ years in infrastructure/platform development with AWS networking, software-defined networking, Python, Kubernetes (EKS/ECS), IaC (Terraform/Pulumi/CloudFormation), observability (Splunk/Datadog), networking appliances (Palo Alto), Nginx, API Gateway, and CDN/edge solutions.
Python, Kubernetes, EKS, ECS, Istio, Splunk, Datadog, AWS, VPC, Route53, ALB, ELB, NATGW, AWS Network Firewall, AWS WAF, Terraform, Pulumi, CloudFormation, Bash, Palo Alto, Nginx, API Gateway, Cloudflare, Fastly
2mo
Save
Mark Applied
Hide
Staff Software Engineer, Infrastructure
Canada or United States or Seattle or Paris or New York City
$238k-$382k/yr RemoteFull Time
Docker
Docker: Privately held container application platform helping developers build, share, and run applications.
8+ YOE8+ years backend/infrastructure engineering; strong Go; experience with Kubernetes/EKS, cloud platforms, networking, and reliability; Bachelor's or equivalent; experience leading cross-team technical initiatives and strong written/verbal communication.
Go, Terraform, Argo CD, EKS, Envoy Gateway, Grafana Cloud, OpenTelemetry, Prometheus, Grafana, GitHub Actions, Kubernetes, Linux, GitOps, Covey Scout