Intercontinental Exchange
Posted 2y ago

Lead Engineer, Platform Engineering - AI

Intercontinental Exchange
Atlanta or New York or Pleasanton or Chicago or Jacksonville
$134k-$155k/yrOnsiteFull Time
Responsibilities
  • defining strategy
  • leading team
  • managing budget
Requirements
  • 8+ years in IT infrastructure/platform engineering
  • 3+ years technical leadership
  • Kubernetes
  • GPU/CUDA and MCP experience
  • Vector DBs
  • RAG and agentic AI pipelines
  • Vendor management, and strong executive communication
Technical tools mentioned
KubernetesNVIDIACUDAMCPs

Job description

Job Purpose

We are on a mission as a team.  We are problem solvers and partners, always starting with our customers to solve their challenges and create opportunities.  Our start-up roots keep us nimble, flexible, and moving fast.  We take ownership and make decisions. We all work for one company and work together to drive growth across the business.  We engage in robust debates to find the best path, and then we move forward as one team.  We take pride in what we do, acting with integrity and passion, so that our customers can perform better.  We are experts and enthusiasts - combining ever-expanding knowledge with leading technology to consistently deliver results, solutions and opportunities for our customers and stakeholders.   Every day we work toward transforming global markets.

 

The AI Platform Engineering Lead drives the AI Platform Operations team, guiding platform strategy, governance, and stakeholder engagement. They align technical execution with business goals, ensuring cost-effective, secure, and scalable AI/ML solutions. The lead defines the architecture, standards, and governance across AI/ML infrastructure, workflow automation, and agentic AI capabilities, including RAG pipelines, vector store infrastructure, agent memory frameworks, and MCP server strategy. They establish design patterns and best practices that span LLM, MCP, and agentic capabilities, ensuring the platform scales securely and operates with consistency across the enterprise. The lead collaborates across teams to set security standards, manage resources, maintain compliance, and align platform capabilities with enterprise architecture standards.

 

Responsibilities

  • Define platform strategy, roadmap, and capability evolution
  • Establish governance frameworks, policies, and exception processes
  • Manage team budget, CapEx planning, and vendor relationships
  • Build and lead the AI Platform Operations team
  • Define the architecture, standards, and governance for AI/ML infrastructure, including GPU cluster design, compute resource planning, security controls, and observability across the platform
  • Drive the strategy, standards, and governance for AI-enabled workflow automation across LLM, MCP, and agentic capabilities, ensuring the platform scales securely and operates with consistency across the enterprise
  • Define the architecture and design standards for vector store infrastructure supporting RAG pipelines, agent memory, and semantic search across the enterprise
  • Establish design patterns and best practices for RAG workflow implementation, including ingestion strategies, chunking approaches, embedding model selection, and retrieval optimization
  • Architect agent memory frameworks, defining standards for short-term context, long-term persistent memory, and episodic memory patterns across AI platform workloads
  • Drive the architecture and governance of Agentic AI systems, including multi-agent orchestration design and tool-use pipeline standards
  • Define the strategy and architecture for hosting and managing MCP servers across the platform, including deployment topology, security boundaries, and integration standards
  • Establish governance frameworks and policies for MCP server lifecycle management, versioning, and access control
  • Evaluate and select MCP server tooling and vendors; manage relationships and roadmap alignment
  • Serve as executive liaison for platform matters
  • Own major incident management and executive communication
  • Drive continuous improvement and platform maturity initiatives
  • Align platform capabilities with enterprise architecture standards
  • Respond to and assist in production operations in a 24/7 environment
  • Provide technical analysis, resolve problems, and propose solutions
  • Provide support to, and coordinate with, developers, operations staff, release engineers, and end-users
  • Educate and mentor team members and operations staff

 

Knowledge and Experience

  • 8+ years in IT infrastructure or platform engineering roles
  • 3+ years in technical leadership or management positions
  • 1+ years hands-on experience with Kubernetes in production
  • Direct experience with GPU infrastructure (NVIDIA preferred)
  • 2+ years experience using CUDA
  • 1+ years experience using MCPs
  • 2+ years experience with vector databases and embedding infrastructure
  • 2+ years experience with RAG pipeline design and deployment
  • 2+ years experience with agent memory patterns (in-context, external stores, retrieval-augmented memory)
  • 1+ years experience with agentic AI systems using orchestration frameworks
  • 2+ years experience with semantic search, embedding models, and ANN search techniques
  • 3+ years working with workflow/orchestrion automation tools
  • Experience managing teams of 5+ technical staff
  • Demonstrated success in vendor management and contract negotiation
  • Strong executive communication and presentation skills
  • Understanding of AI/ML workloads and infrastructure requirements
  • Experience with enterprise monitoring and observability tools
  • Ability to work in a service-oriented team environment
  • Project Management, organization, and time management
  • Customer focused, and dedicated to the best possible user experience
  • Communicate effectively with both technical and business resources
  • Fluent speaking, reading, and writing in English

 

Illinois Base Salary Range 

The expected base salary for this role, if located in Illinois, is between $133,900 - 154,500 USD.  The base salary range does not include Intercontinental Exchange’s incentive compensation.  While we provide this range as general guidance, at ICE we compensate employees based on the skillset and experience of the individual. Regular full-time ICE employees are eligible for a suite of competitive employee benefits, including healthcare coverage (medical, dental and vision), a 401(k) plan, life insurance, time off, and paid leave for qualifying circumstances. 

 

New York Base Salary Range 

The expected base salary for this role, if located in New York, is between $149,400 - 180,000 USD.  The base salary range does not include Intercontinental Exchange’s incentive compensation.  While we provide this range as general guidance, at ICE we compensate employees based on the skillset and experience of the individual. Regular full-time ICE employees are eligible for a suite of competitive employee benefits, including healthcare coverage (medical, dental and vision), a 401(k) plan, life insurance, time off, and paid leave for qualifying circumstances. 

 

California Base Salary Range 

The expected base salary for this role, if located in California, is between $149,400 - 180,000 USD.  The base salary range does not include Intercontinental Exchange’s incentive compensation.  While we provide this range as general guidance, at ICE we compensate employees based on the skillset and experience of the individual. Regular full-time ICE employees are eligible for a suite of competitive employee benefits, including healthcare coverage (medical, dental and vision), a 401(k) plan, life insurance, time off, and paid leave for qualifying circumstances. 

#LI-SH3

#LI-ONSITE

About Intercontinental Exchange

Global provider of financial exchanges, clearing houses, and technology.

Similar jobs

Platform Engineer - AI roles near Atlanta, Georgia
2w
Save
Mark Applied
Hide
Principal AI Platform Engineer
Atlanta, Georgia, United States
$142k-$203k/yr OnsiteFull Time
Capgemini
CapgeminiEuronext Paris: CAP: Provides global IT consulting and digital transformation services.
Experience leading AI platform or infrastructure teams, hands-on LLM inference/serving, cloud (AWS/Azure/GCP) with Kubernetes and IaC, regulated-industry delivery, research fluency, and platform-as-product mindset.
Lang Smith, Lang Graph Platform, model gateway, vector search, PostgreSQL, pg vector, ClickHouse, S3-compatible object storage, Neo4j Enterprise, Open Telemetry, Grafana, Kubernetes, Helm, Argo CD, AWS, Azure, GCP
1mo
Save
Mark Applied
Hide
AI Platform Engineer
Atlanta or Austin or Houston or Dallas or Miramar or Orlando or Tallahassee or West Palm Beach or Fort Lauderdale or Phoenix or Charlotte
HybridFull Time
Greenberg Traurig
Greenberg Traurig: Provides global legal representation and corporate advisory services.
7+ YOE7+ years platform engineering experience with 3+ years deploying AI/ML workloads in cloud; experience with Azure/AWS/GCP, Terraform, containers, Python/PowerShell, and AI architecture patterns.
Azure AI Foundry, Azure OpenAI, AWS Bedrock, Google Vertex AI, Terraform, Docker, Kubernetes, Python, PowerShell, Semantic Kernel, LangChain, REST API
3mo
Save
Mark Applied
Hide
Sr Advanced AI Platform Engineer
Atlanta, Georgia, United States
HybridFull Time
Honeywell
HoneywellNASDAQ: HON: Manufactures aerospace products, building technologies, and industrial control systems.
8+ YOE8+ years in software/data/ML platform engineering; strong Python and systems language; cloud-native data platforms; ML/AI pipelines; LangChain/LangGraph/LangSmith; edge AI; knowledge graphs.
Python, Go, Rust, C++, Databricks, BigQuery, Azure Data Lake, Kubernetes, LangChain, LangGraph, LangSmith, MLflow, Azure Machine Learning Studio, ML-Ops, FastAPI
3mo
Save
Mark Applied
Hide
Principal - Business Consulting
Atlanta or Dallas or Los Angeles or Seattle or California or Georgia or Texas or Washington or United States
$154k-$193k/yr FieldFull Time
Infosys
InfosysNYSE: INFY: Provides IT consulting, software development, and business outsourcing services.
7+ YOEBachelor's or equivalent, 7+ years in cloud architecture or AI platform engineering, hands-on with agentic platforms/AI gateways, Responsible AI and governance (OPA/Rego), cloud-native (Azure AKS/AWS EKS), CI/CD and IaC (Terraform/Helm), consulting experience, travel up to 75%.
Azure AKS, AWS EKS, Azure OpenAI Service, AWS Bedrock, Microsoft Foundry, MoveWorks, Open Policy Agent (OPA), Rego, OpenAI GPT, Anthropic Claude, Meta Llama, LangChain, CrewAI, Haystack, Semantic Router, AutoGen, LlamaIndex, DSPy, MCP Protocol, MCP servers, OAuth2, OpenID Connect, Entra ID, Okta, OpenAPI, JSON schemas, Terraform, Helm, CI/CD
1w
Save
Mark Applied
Hide
Data & AI Platform Engineer
San Ramon or Salt Lake City or Philadelphia or Chicago or Pasadena or Bellevue or Irvine or Duluth or Los Angeles or Boca Raton or Brunswick or Atlanta or Boise or El Segundo or Denver or Dallas or New York City or Los Angeles or Garden City or Nashville or St. Louis or Woodland Hills or San Jose or San Francisco or Austin
$105k-$143k/yr OnsiteFull Time
Armanino
Armanino: Provides accounting, tax, and consulting services to diverse organizations.
2+ YOERequires 2+ years in data, platform, analytics engineering, or cloud operations; 1+ year accounting experience; cloud data, BI, CI/CD, scripting, SQL, governance, and AI/ML familiarity.
Microsoft Fabric, Snowflake, Databricks, Power BI, Tableau, Git, Python, PowerShell, SQL, Azure Key Vault, Azure OpenAI, AI Search
4mo
Save
Mark Applied
Hide
Lead Engineer, Platform Engineering - AI
Atlanta, Georgia, United States
$134k-$155k/yr OnsiteFull Time
Intercontinental Exchange
Intercontinental ExchangeNYSE: ICE: Operates global financial exchanges, clearing houses, and mortgage platforms.
8+ YOE3+ MgmtLead platform engineering role focusing on AI/ML infrastructure, governance, and operations with 8+ years IT infrastructure/engineering and 3+ years leadership.
Kubernetes, CUDA, MCPs, Vector databases, Embedding infrastructure, GPU infrastructure, LLM, RAG pipelines, Observability, Security controls, Versioning