This job has expired

This job posting is no longer active and is not accepting applications. Explore similar roles below!

JPMorgan Chase
Posted 1mo ago

Data Engineer III - Python, Databricks

JPMorgan Chase
Houston or Plano
OnsiteFull Time
Responsibilities
  • designing pipelines
  • developing pipelines
  • implementing security
Requirements
  • 3+ years applied experience
  • Strong Python
  • Databricks, AWS
  • Advanced SQL
  • NoSQL familiarity
  • Experience across data lifecycle, and use of AI coding assistants for code generation and validation
Technical tools mentioned
PythonDatabricksSQLNoSQLAWSAWS GLUEAurora PostgresMongoDBCursorGitHub CopilotClaudeDockerCI/CDJavaScriptJava

Job description

Join a dynamic team where your unique skills will help build innovative solutions and contribute to a winning culture. You’ll have opportunities for career growth, collaborate with talented professionals, and make a real impact on our business objectives.  

As a Data Engineer III at JPMorgan Chase within our agile team, you will design and deliver reliable data collection, storage, access, and analytics solutions that are secure, stable, and scalable.. You will develop, test, and maintain essential data pipelines and architectures, supporting various business functions to achieve the firm’s goals. Working with us, you will use your skills to drive innovation and help shape our team culture. Together, we focus on excellence, collaboration, and continuous improvement.

Job responsibilities

  • Develop workflows and ELT pipelines using Python and Databricks.
  • Support review of controls to ensure sufficient protection of enterprise data.
  • Implement data security using entitlements frameworks.
  • Update logical or physical data models based on new use cases.
  • Use SQL frequently and understand NoSQL databases
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate data pipeline/design analysis and documentation, validating outputs and handling data according to sensitivity and security requirements.
  • Applies reuse-first, AI-assisted practices to strengthen SDLC-quality routines for data pipelines (e.g., test generation and control validation), ensuring traceability/auditability and alignment to resiliency and security expectations.

 

Required qualifications, capabilities, and skills

  • Formal training or certification on software engineering concepts and 3 years applied experience.
  • Good working knowledge of AWS, Databricks, and Python, Experience across the data lifecycle.
  • Advanced at SQL, including joins and aggregations, Working understanding of NoSQL databases.
  • Significant experience with statistical data analysis and ability to determine appropriate tools and data patterns for analysis.
  • Utilize AWS Cloud Services for developing, deploying, and managing applications at scale.
  • Proficiency in AI Coding Assistants, Daily use of tools like Cursor, GitHub Copilot, and Claude to accelerate code generation, documentation, and refactoring.
  • Effective Prompt Engineering,  Providing AI models with context, clear goals, relevant source material, and - defined output expectations to generate accurate, usable code.
  • Critical Evaluation & Validation, Ability to identify hallucination patterns, security vulnerabilities, and logic errors in AI-generated code, ensuring safety before production deployment.
  • Demonstrated experience using enterprise-authorized AI capabilities within the work environment to support data engineering workflows with strong validation habits and awareness of data sensitivity.
  • Ability to review and validate AI-assisted outputs (e.g., query suggestions, test ideas, or model change summaries) before use, escalating when uncertain and following data handling requirements.

 

Preferred qualifications, capabilities, and skills

  • Familiarity with the Standardized data layer practices (Medallion architecture)
  • Exposure to Aurora Postgres and MongoDB
  • Experience developing and supporting AWS GLUE Jobs, Federated Data Lake
  • Skills in designing efficient data models including normalization, denormalization, and schema design and an understanding around relational and star schemas.
  • Augmented Development Workflow: Integrating tools into CI/CD pipelines, containerization (e.g., Docker), and leveraging AI to quickly bridge language gaps (e.g., transitioning between Python, JavaScript, or Java).

 

 

About Company

J.P. Morgan’s Commercial & Investment Bank is a global leader across banking, markets, securities services and payments. Corporations, governments and institutions throughout the world entrust us with their business in more than 100 countries. The Commercial & Investment Bank provides strategic advice, raises capital, manages risk and extends liquidity in markets around the world. 

Company


JPMorganChase, one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world’s most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management.

We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process. 

We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants’ and employees’ religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation.

JPMorgan Chase & Co. is an Equal Opportunity Employer, including Disability/Veterans

About JPMorgan Chase

Global financial services firm providing banking and investment solutions.

Similar jobs

Data Engineer roles near Houston, Texas
12h
Save
Mark Applied
Hide
Senior Data Engineer
Stafford, Texas, United States
OnsiteFull Time
Eurofins Scientific
Eurofins ScientificEuronext Paris: ERF: Global provider of analytical testing and laboratory services.
Bachelor’s or master’s degree or equivalent experience; expert Python, SQL, Git, cloud, orchestration, databases, scientific data, and MLOps experience; analytical chemistry knowledge; unrestricted US work authorization; English proficiency.
Machine Learning (ML), Python, Pandas, NumPy, AWS, Azure, GCP, Azure Data Lake Storage (ADLS), Azure Compute VMs, Azure Functions, Apache Airflow, Prefect, Dagster, SQL, PostgreSQL, Git, Azure ML Studio, MLflow, Kubeflow, Sagemaker, CSV, JSON, LIMS/ELN
1d
Save
Mark Applied
Hide
Sr Data Engineer_ITDE1I
Houston, Texas, United States
$102k-$170k/yr HybridFull Time
ENGIE
ENGIEEuronext Paris: ENGI: Produces and distributes electricity, gas, and renewable energy.
5+ YOEBachelor’s degree or equivalent experience; 5+ years building cloud data pipelines; 2+ years in energy trading, asset management, or energy operations; Python, SQL, ETL/ELT, orchestration, modeling, APIs, and stakeholder backlog management.
AWS, Azure, GCP, Python, SQL, ETL, ELT, APIs, CI/CD
1d
Save
Mark Applied
Hide
Data Engineer
Atlanta or Chicago or Fairview Heights or Houston or Arlington
$89k-$148k/yr OnsiteFull Time
Guidehouse
Guidehouse: Provides management and technology consulting services to diverse organizations.
5+ YOEBachelor's degree and 5 years of data engineering or software development experience. Requires U.S. citizenship or green card, ability to obtain Public Trust, Java, Python, SQL, PySpark, ETL, databases, Databricks, APIs, and CI/CD experience.
SQL, R, Python, Git, Docker, Kubernetes, Java, PySpark, RESTful, SOAP, Co-Pilot, CodeX, MySQL, PostgreSQL, SQL Server, MongoDB, Databricks, GitHub, Jira, Confluence, Kibana, Tableau, Spring Boot
3d
Save
Mark Applied
Hide
Lead Analyst Data Engineer
Houston, Texas, United States
OnsiteFull Time
Boardwalk Pipelines
Boardwalk Pipelines: Transporting and storing natural gas and liquids via pipelines.
7+ YOEBachelor's degree in a related field and 7–10 years of cloud data architecture experience, including 5+ years with AWS. Requires AWS, Databricks, SQL, Python, governance, and leadership expertise.
Google Chrome, Safari, Firefox, Microsoft Edge, AWS Glue, Amazon Redshift, Amazon Athena, AWS Lake Formation, Amazon SageMaker, Amazon Bedrock, AWS Step Functions, Databricks, Alation, SQL, T-SQL, Python, PySpark, Pandas, DAX, Power Query (M), PL/SQL, SQL Server, Oracle, PostgreSQL, Azure SQL, Teradata, Power BI, Git, Visual Studio Code, SSMS
6d
Save
Mark Applied
Hide
Enterprise Data Engineer
Ridgeland or Alabama or Houston or Memphis or Florida or Atlanta
RemoteFull Time
Trustmark
TrustmarkNASDAQ: TRMK: Provides retail and commercial banking, wealth, and insurance services.
4+ YOEBachelor's degree in data or computer science or equivalent certification, 4 years with modern ETL platforms, database, data warehousing, modeling, SQL, Python, and advanced analytical skills.
IBM Datastage, Informatica, Snowpipe, Azure, AWS, SQL, Python
1w
Save
Mark Applied
Hide
Google Senior Data Engineer
Albany or Arlington or Atlanta or Austin or Beaverton or Bentonville or Boston or Carmel or Charlotte or Chicago or Cincinnati or Cleveland or Columbus or Culver City or Denver or Des Moines or Detroit or Hartford or Houston or Irving or Kirkland or Miami or Milwaukee or Minneapolis or Morristown or Mountain View or Nashville or New York City or Oklahoma City or Overland Park or Philadelphia or Pittsburgh or Raleigh or Redmond or Sacramento or San Diego or San Francisco or Scottsdale or Seattle or St. Louis or St. Petersburg or Walnut Creek or United States
$80k-$266k/yr HybridFull Time
Accenture
AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
5+ YOERequires 5+ years in data engineering, analytics, or ML; 4+ years with GCP; 5+ years with SQL and pipelines; 3+ years with Python or AI tools; and a bachelor's degree or equivalent.
Google Cloud Platform (GCP), BigQuery, Looker, Vertex AI, Gemini Foundation Models, Gemini Enterprise, Dataflow, Dataproc, Pub/Sub, Cloud Storage, Looker Studio, Model APIs, Embeddings, Dataplex, IAM, SQL, Python, Git
1w
Save
Mark Applied
Hide
Data Engineer 1
The Woodlands, Texas, United States
OnsiteFull Time
Accelerated Mobile Power
Accelerated Mobile Power: Provides mobile power solutions using gas turbines and generators.
3+ YOEBachelor's degree in MIS, computer science, engineering, IT, or related field, or equivalent experience; 3–5+ years in a data-focused role; Databricks, PySpark, SparkSQL, SQL, and visualization experience.
Databricks, Delta Lake, Delta Live Tables (DLT), Databricks Jobs, Auto Loader, Change Data Capture (CDC), Spark, PySpark, SparkSQL, SQL, Microsoft SQL Server, Power BI, Tableau, Spotfire, Sigma
1w
Save
Mark Applied
Hide
Data Engineer 1
The Woodlands, Texas, United States
OnsiteFull Time
Beusa Energy
Beusa Energy: Provides oil exploration, fracturing, and power generation services.
3+ YOEBachelor's degree in MIS, computer science, engineering, IT, or related field, or equivalent experience; 3-5+ years in a data-focused role; strong Databricks, PySpark, SparkSQL, SQL, and visualization skills.
Databricks, Delta Lake, Delta Live Tables (DLT), Databricks Jobs, Autoloader, CDC, Spark, PySpark, SparkSQL, SQL, Microsoft SQL Server, Power BI, Tableau, Spotfire, Sigma, Z-Ordering, Auto Optimize, OPTIMIZE, VACUUM, MERGE, Change Data Feed, APPLY CHANGES INTO
This job has expired