Brain Co.
Posted 3mo ago

AI Platform Engineer, Backend (Agentic Engineering)

Brain Co.
San Francisco, California, United States
HybridFull Time
Responsibilities
  • design sandboxing
  • build foundations
  • enable orchestration
Requirements
  • 5+ years building backend systems in production
  • Deep proficiency in Python, TypeScript, Go, or Rust
  • Strong fundamentals in distributed systems
  • Designed APIs and services
  • Built shared infrastructure
  • Strong developer experience
Technical tools mentioned
PythonTypeScriptGoRustKubernetesOAuth/OIDCSecrets managementCI/CDVM isolation

Job description

About Brain Co.

Brain Co. is an applied AI startup co-founded by Jared Kushner and Elad Gil, and backed by leading Silicon Valley builders including Patrick Collison and Andrej Karpathy.

We are building AI applications for the world’s most important institutions, delivering impact on real-world problems across governments, healthcare systems, and critical industries.

Our progress so far:

  • Automated construction permitting for a sovereign government → 80% faster, unlocking $375M+ in value

  • Optimized supply chains for a leading global energy company → 30% lower cost, 99% reliability, preventing $100M+ in losses

  • Streamlined hospital patient care across national health systems → 40% better outcomes, 80% less admin work

Company momentum:

  • Raised a $55M Series A from leading investors

  • Built a team of 70+ AI experts from Tesla, Google DeepMind, NVIDIA, and Databricks

At Brain Co., we focus on applying frontier AI to real institutional challenges, working alongside governments, healthcare systems, and critical industries to modernize how essential services operate.

We are looking for leaders who want to help bring new technology into institutions that impact millions of people.

About the role:

You'll join the team that builds and enables agentic workflows across Brain Co. For every engineer, operator, and business team internally, and for the production AI systems we deploy to governments, healthcare systems, and critical industries. This is a platform role at the center of the company's agent-first strategy: you'll build foundational systems used by every engineering team, and the bar is product-grade because the entire company depends on them.

What you’ll work on:

  • Own the foundations of how LLMs are used across the company: cost visibility and controls, data privacy, identity and access, routing, and the security posture around all provider traffic.

  • Design the sandboxing, orchestration, audit, and guardrail layers that product teams build their agents on, so verticals don't need to invent their own abstraction.

  • Solve the hard problems: prompt-injection defenses, scoped credentials, kill switches, multi-tenant isolation (including VM-level pod isolation), and runaway-cost controls.

  • Design the orchestration, isolation, and resource models that make this viable: cold-start vs. always-on tradeoffs, credential and token lifecycle, fan-out and fan-in patterns, fairness and quota enforcement across tenants, and the observability needed to debug at that volume.

  • Make AI-assisted development a first-class platform layer: coding agents that review and ship code, automate CI, refactor at scale, and run as background workers across the codebase, together with the canonical scaffolding and guardrails that govern them.

  • Build the systems that let every team; engineering, operations, and the business, run their own agents reliably and safely against the tools they already use, with the right credentials, scheduling, memory, and audit underneath.

  • End-to-end ownership: architecture, implementation, rollout, observability, on-call, and iteration based on internal user feedback.

  • Partner closely with security, infrastructure, and product teams to make agent deployments safe by default.

You Might Be a Great Fit If You…

  • Have 5+ years building backend systems in production, with deep proficiency in at least one of Python, TypeScript, Go, or Rust.

  • Bring strong fundamentals in distributed systems: consistency, idempotency, retries, failure modes, queueing, scheduling.

  • Have designed and operated APIs and services that other engineers depend on.

  • Have a proven track record building shared infrastructure, internal platforms, or developer-facing services that real users adopted.

  • Have strong intuition for developer experience, long-term maintainability, and where to draw abstraction boundaries.

  • Are comfortable owning the full lifecycle: writing the design doc, shipping the MVP, hardening it, and driving adoption across the company.

  • Have owned services with real uptime and operational responsibility, and are comfortable with observability stacks, incident response, and SLOs.

  • Bring cloud-native experience: Kubernetes, infrastructure-as-code, OAuth/OIDC, secrets management.

Ways you might stand out:

  • Experience building or operating LLM infrastructure: gateways, inference systems, prompt routing, cost attribution, evaluation harnesses.

  • Experience with agent frameworks, tool-use systems, or sandboxed code execution.

  • Security instincts around prompt injection, supply-chain risk in agent ecosystems, and credential scoping for autonomous systems.

  • Background in multi-tenant, regulated, or government deployments (HIPAA, SOC2).

  • Open-source contributions to AI infrastructure, agent tooling, or developer platforms.

Why Join Us

  • Collaborate with industry veterans from Tesla, DeepMind, Databricks, and more

  • Accelerate your career with ownership based on impact, not tenure

  • Earn competitive compensation + meaningful equity in a high-growth company

  • Thrive in a culture built on speed, curiosity, and impact

Benefits

  • Competitive salary plus equity

  • Daily lunches

  • Commuter benefits

  • 401(k)

  • Medical, Dental and Vision

  • Unlimited PTO

About Brain Co.

Builds AI software platforms for governments and large enterprises.

Year founded
2024
Employees
40
Organization type
Private
Latest investment
Raised $30.00M Series A (2025) — led by Affinity Partners, Gil Capital
Headquarters
US

Similar jobs

AI Platform Engineer roles near San Francisco, California
2d
Save
Mark Applied
Hide
Principal AI Platform Engineer - AI Experiences, Copilot Engineering
Mountain View, California, United States
OnsiteFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
N/A
1w
Save
Mark Applied
Hide
Staff AI Platform Engineer
San Francisco or Los Angeles or Denver or Austin or Chicago or New York or Seattle or Toronto or Santa Barbara or San Diego
$180k-$273k/yr RemoteFull Time
Invoca
Invoca: AI platform for conversation intelligence and revenue execution.
7+ YOERequires 7+ years in AI/ML platform engineering, strong Python and/or TypeScript, distributed systems, APIs, evaluation, observability, and AI infrastructure experience; bachelor's degree or equivalent practical experience.
Python, TypeScript, JavaScript, React, Kubernetes, Docker, Terraform, LangChain, LangGraph, LlamaIndex, LiveKit, REST, gRPC, MCP, A2A
2w
Save
Mark Applied
Hide
AI Platform Engineer
San Francisco, California, United States
$190k-$310k/yr OnsiteFull Time
Applied Compute
Applied Compute: Developing specialized AI agents for corporate enterprise workflows.
Strong full-stack engineering skills, production web application experience, real-time and asynchronous systems experience, rapid application development ability, and familiarity with LLM-powered applications and agent architectures.
SDK, LLM
2w
Save
Mark Applied
Hide
AI Platform Engineer
San Francisco, California, United States
OnsiteFull Time
Sapiom
Sapiom: Financial infrastructure for autonomous AI agents.
8+ YOE8+ years building large-scale backend or distributed systems; staff/principal level technical leadership; strong Python, Go or Typescript skills; experience with developer platforms, event-driven architectures, and modern AI/LLM systems.
Python, Go, Typescript, LangGraph, OpenAI Agents SDK, CrewAI, AutoGen, Temporal, vLLM, SGLang, TensorRT-LLM, LLMs
1mo
Save
Mark Applied
Hide
Senior AI Platform Engineer Contract
San Francisco, California, United States
OnsiteContract
Nextdata
Nextdata: Provides an operating system for building autonomous data products.
Hands-on experience building backend services, agentic AI apps, RAG/retrieval systems, SQL and vector search integration, API development, and production controls (access, audit, testing).
LangChain, LangGraph, Python, SQL, Databricks, Snowflake, BigQuery, Spark, DuckDB, Postgres, OAuth, OIDC, SAML, SSO, RBAC, ABAC, SCIM, MCP
1mo
Save
Mark Applied
Hide
Senior Platform AI Engineer
Santa Clara or United States
$184k-$357k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
8+ YOE8+ years designing and operating production platform/backend infrastructure,5+ years ML infrastructure,BS/MS/PhD or equivalent in CS/EE/CE,strong Python and compiled-language skills,experience with Kubernetes Jobs and job queues.
Python, C, C++, Go, Java, Rust, Kubernetes Jobs, Celery, Sidekiq, Temporal, container runtimes, EDA
1mo
Save
Mark Applied
Hide
Senior Platform AI Engineer
Santa Clara or California
$184k-$357k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOE8+ years building and operating production platform/backend infrastructure; 5+ years ML infrastructure; strong Python and a compiled language; experience with job queues, sandboxed execution, and production reliability.
Python, C, C++, Go, Java, Rust, Kubernetes Jobs, Celery, Sidekiq, Temporal, container runtimes
1mo
Save
Mark Applied
Hide
AI Platform Engineer
Milpitas, California, United States
$136k-$232k/yr OnsiteFull Time
KLA
KLANASDAQ: KLAC: Provides process control and yield management for semiconductor manufacturing.
5+ YOEDegree in CS/CE or related, 5+ years systems/DevOps/ML infrastructure experience, hands-on AI/GPU cluster and Linux/Kubernetes expertise, Python/Bash scripting, and knowledge of storage, networking, and security.
Linux, Kubernetes, Docker, Slurm, Ray, TensorFlow, PyTorch, GPU, TPU, Bash, Python, CI/CD, Infrastructure as Code (IaC), TPM