Capital One
Posted 2w ago

Lead Data Engineer (Python, AWS, Spark, Kafka, SQL, Snowflake, Databricks, GenAI)

Capital One
New York City or McLean or Richmond or New York or Virginia
$179k-$246k/yrOnsiteFull Time
Responsibilities
  • designing solutions
  • developing pipelines
  • mentoring engineers
Requirements
  • Bachelor's degree
  • 4+ years in application development
  • 2+ years in big data, and 1+ year in cloud computing
  • Python, SQL
  • Data modeling, and data engineering experience required
Technical tools mentioned
PythonAWSApache SparkApache KafkaSQLSnowflakeDatabricksMicrosoft AzureGoogle CloudRDBMSNoSQLMapReduceHadoopHiveEMRMySQLMongoDBCassandraAmazon RedshiftUNIXLinuxClaude CodeGitHub CopilotETL

Job description

Overview



Lead Data Engineer (Python, AWS, Spark, Kafka, SQL, Snowflake, Databricks, GenAI)

Do you love building and pioneering in the technology space? Do you enjoy solving complex business problems in a fast-paced, collaborative,inclusive, and iterative delivery environment?

At Capital One, you'll be part of a big group of makers, breakers, doers and disruptors, who solve real problems and meet real customer needs. We are seeking Data Engineers who arepassionate about data, data modeling, data quality, governance and permissible use, and building data pipelines with emerging technologies. As a Capital One Lead Data Engineer, you’ll have the opportunity to be on the forefront of driving a major transformation within Capital One. 

The Marketing and Messaging team is responsible for delivering hyper-personalized messages and experiences that will delight the customer, attract prospects and drive increasing business value. The team builds scalable platforms that deliver omnichannel messages in owned and paid Adtech channels.

What You’ll Do:

  • Collaborate with and across Agile teams to design, develop, test, implement, and support technical solutions using data movement tools and technologies

  • As a lead developer, work with a team of developers who have deep experience in data movement, distributed computing, and full-stack systems

  • Utilize programming languages like Python, SQL, and Open Source RDBMS and NoSQL databases and cloud-based data warehousing services such as Snowflake and Databricks

  • Optimize information systems for end-users and downstream application consumers by using sound data design practices

  • Share your passion for staying on top of tech trends, experimenting with and learning new technologies, participating in internal & external technology communities, and mentoring other members of the engineering community

  • Collaborate with digital product managers, and deliver robust cloud-based solutions that drive powerful experiences to help millions of Americans achieve financial empowerment

  • Perform unit tests and conduct reviews with other team members to make sure your code is rigorously designed, elegantly coded, and effectively tuned for performance

Basic Qualifications: 

  • Bachelor’s Degree 

  • At least 4 years of experience in application development (Internship experience does not apply)

  • At least 2 years of experience in big data technologies 

  • At least 1 year experience with cloud computing (AWS, Microsoft Azure, Google Cloud)

Preferred Qualifications:

  • Master's Degree

  • 7+ years of experience in application development including Python, SQL, Spark, ETL tools, or AWS Glue

  • 4+ years of experience with a public cloud (AWS, Microsoft Azure, Google Cloud)

  • 4+ years experience with Distributed data/computing tools (MapReduce, Hadoop, Hive, EMR, Kafka, Spark, or MySQL)

  • 4+ year experience working on real-time data and streaming applications 

  • 4+ years of experience with NoSQL implementation (Mongo, Cassandra) 

  • 4+ years of data warehousing experience (Redshift or Snowflake) 

  • 4+ years of experience with UNIX/Linux including basic commands and shell scripting

  • 4+ years of experience with data modeling for data warehousing

  • 2+ years of experience with Agile engineering practices

  • Experience leveraging interactive AI tooling (Claude Code, GitHub Copilot) to accelerate software delivery

At this time, Capital One will not sponsor a new applicant for employment authorization, or offer any immigration related support for this position (e.g. H1B, F-1 OPT, F-1 STEM OPT, F-1 CPT, J-1, TN, E-3, and O-1, or any other forms of work authorization that require immigration support from an employer).

The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked.

McLean, VA: $197,300 - $225,100 for Lead Data Engineer


New York, NY: $215,200 - $245,600 for Lead Data Engineer


Richmond, VA: $179,400 - $204,700 for Lead Data Engineer









Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate’s offer letter.

This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan.

Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at theCapital One Careers website. Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level.

This role is expected to accept applications for a minimum of 5 business days.

No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections 4901-4920; New York City’s Fair Chance Act; Philadelphia’s Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries.

If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1-800-304-9102 or via email at [email protected]. All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations.

For technical support or questions about Capital One's recruiting process, please send an email to [email protected]

Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site.

Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).

About Capital One

A diversified financial services providing banking and credit products.

Similar jobs

Data Engineer roles near New York City, New York
49m
Save
Mark Applied
Hide
Lead Data Engineer
Teaneck, New Jersey, United States
OnsiteFull Time
Cognizant
CognizantNASDAQ: CTSH: Provides IT consulting and technology services to global enterprises.
Extensive experience designing scalable Big Data pipelines and Azure Cloud solutions using Spark, Python, SQL, Kafka, Azure Synapse, Databricks, and distributed data-processing frameworks; technical leadership required.
Spark, PySpark, Scala, Java, Python, SQL, Hive, Kafka, Azure Synapse Analytics, Azure Data Lake Storage, Azure Data Factory, Databricks, HDFS, MapReduce, Impala, Tez, Sqoop, Oozie, HBase, Cassandra, MongoDB, Storm, Knox, Ranger, Flume, NiFi, Kerberos, Sentry, Cloudera Manager, Cloudera Navigator, Ambari, Azure Cloud
12h
Save
Mark Applied
Hide
Data Engineer
Plano or Teaneck
$65k/yr OnsiteFull Time
Cognizant
CognizantNasdaq: CTSH: Provides global information technology and business process outsourcing services.
Bachelor’s or Master’s degree in a related field; Python and SQL skills; data pipelines, orchestration, cloud, data architecture, and analytical and communication skills required.
Python, SQL, Spark, AWS, Azure, GCP, Airflow, Prefect, Snowflake, Databricks, BigQuery, Docker, Kubernetes, GitHub, Electronic Medical Records (EMR)
12h
Save
Mark Applied
Hide
Data Engineer
New York City, New York, United States
$100k-$115k/yr HybridFull Time
News Corp
News CorpNasdaq: NWSA: Global media, publishing, and digital information services.
2+ YOEBachelor's or master's degree in computer science or related field, 2+ years of software engineering experience, and proficiency in Python, SQL, PySpark, cloud services, data pipelines, and data warehouses.
Vertex AI, AWS Lambda, Amazon DynamoDB, Google Cloud Functions, Amazon S3, Kubernetes, AWS Glue, Google BigQuery, Amazon Web Services (AWS), Google Cloud Platform (GCP), Python, SQL, PySpark, Apache Airflow, Google Cloud Dataflow, Snowflake, Amazon Redshift, Git, JIRA, Linux, Shell, SSH, crontab, CloudWatch, Datadog, Splunk, REST, GraphQL, Docker
13h
Save
Mark Applied
Hide
Data Engineer II - Digital and Technology Partners - Hybrid/Remote
New York City, New York, United States
$90k-$135k/yr HybridFull Time
Mount Sinai Health System
Mount Sinai Health System: Provides comprehensive hospital, clinical, and medical research services.
4+ YOEBachelor's degree in computer science or related discipline, 4+ years of relevant professional development experience, and proficiency with Azure, databases, big data technologies, programming languages, and Agile practices.
Microsoft Azure, Azure Data Factory, Databricks, SQL, NoSQL, Oracle, PostgreSQL, MySQL, MongoDB, Hadoop, Spark, Kafka, HL7, Mirth, CI/CD, Git, Scala, Python, Java, Node.js, Django, PHP, SOA, AWS, JIRA
13h
Save
Mark Applied
Hide
Software Engineering - Data, Lakehouse and AI Data Platform Engineer - Vice President - New York
New York City, New York, United States
$130k-$250k/yr OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities, and investment management firm.
7+ YOE7–12+ years’ experience; bachelor’s or master’s degree or equivalent; Python or Java, SQL, software engineering fundamentals, distributed processing, data modeling, quality, and production pipeline experience.
Python, Java, SQL, Apache Spark, JSON, Avro, Parquet, ANSI SQL, Kafka, Snowflake, Apache Iceberg, Databricks, Hadoop, Sybase IQ, CI/CD, Kubernetes, EAP
14h
Save
Mark Applied
Hide
Data Engineer III - Digital and Technology Partners - Hybrid/Remote
New York City or United States
$109k-$164k/yr HybridFull Time
Mount Sinai Health System
Mount Sinai Health System: Integrated academic medical system providing patient care and research.
5+ YOEBachelor's degree in computer science or related field, 5+ years of development experience, Azure data engineering, SQL/NoSQL databases, two programming languages, RESTful services, big data, CI/CD, Git, and Agile experience.
Microsoft Azure, Azure Data Factory, Databricks, Oracle, PostgreSQL, MySQL, MongoDB, Scala, Python, Java, Node.js, Django, PHP, Hadoop, Spark, Kafka, Git, Microsoft JIRA, HL7, Mirth, SQL, NoSQL, RESTful, SOA, CI/CD
14h
Save
Mark Applied
Hide
Senior Data Engineer
New York City, New York, United States
$125k-$140k/yr HybridFull Time, Contract
Ekimetrics
Ekimetrics: Provides data science and AI-powered marketing and sustainability solutions.
4+ YOEBachelor’s or Master’s degree in a quantitative field and 4–6 years’ experience; requires Databricks, Python, SQL, Spark, Bash, Azure, ETL/ELT, CI/CD, and stakeholder management.
Databricks, Delta Lake, Apache Spark, Asset Bundles, Unity Catalog, Cluster Management, Python, SQL, Bash, CI/CD, Azure, Google Cloud Platform (GCP), Amazon Web Services (AWS), Power BI, Tableau, ETL, ELT, APIs, CSV, JSON
22h
Save
Mark Applied
Hide
Data Engineer
Mesa or Plano or Teaneck
$65k/yr HybridFull Time
Cognizant
CognizantNasdaq: CTSH: Provides IT consulting and digital business process services.
Bachelor’s or master’s degree in a related field; Python and SQL; data orchestration, cloud, data lake/warehouse, containerization, CI/CD, ETL/ELT, modeling, and architecture knowledge.
Python, SQL, Spark, Airflow, Prefect, Snowflake, Databricks, BigQuery, Docker, Kubernetes, GitHub, AWS, Azure, GCP, JSON, XML, CI/CD