Castleton Commodities International
Posted 1mo ago

Data Engineer – Data Science & Technology

Castleton Commodities International
London, England, United Kingdom
OnsiteFull Time
Responsibilities
  • executing architecture
  • managing ingestion
  • standardizing data
Requirements
  • Bachelor's degree in a related field
  • 3+ years data engineering experience with SQL
  • Data architecture and ETL/ELT pipelines
  • Hands-on Python, Pandas, NumPy
  • Snowflake or relational DB experience
  • Strong data mapping
  • Standardization, and communication skills
Technical tools mentioned
SQLSnowflakePythonPandasNumPyETL/ELT

Job description

At Castleton Commodities International (CCI), we are redefining how data and technology shape the future of energy trading. Our Data Science & Technology team is at the forefront of this transformation, developing systems and innovative tools that empower our front-office teams to better understand market dynamics, forecast prices, and manage risk.

Our commitment to excellence is reflected in a robust, modern, open-source technology stack, empowering us to solve complex, high-impact challenges at scale. From cloud-native infrastructure to machine learning platforms, real-time analytics, and proprietary libraries and APIs, we are passionate about using technology to drive innovation and create a competitive edge in global energy commodity trading and investing. Candidates will gain direct exposure to our commercial investment teams, offering a unique opportunity to see how cutting-edge technology drives strategy and performance in energy markets.

Our Data Engineering team is integral to our mission, ensuring the seamless ingestion and management of data across our platforms. The Data Engineer will play an integral role as the team implements new data management platforms, creates new data ingestion pipelines, and sources new data sets. The role will assist with all aspects of data – from data architecture design to on-going data management and will have significant exposure to our Commercial investing teams globally. This position will play an integral role as the firm continues to expand its use of data for advanced analytics and other commercial purposes.

Responsibilities

  • Execute data architecture and data management projects for both new and existing data sources.

  • Help transition existing data sets, databases, and code to a new technology stack.

  • Manage end to end data ingestion process and publishing to investing teams.

  • Own the process of mapping, standardizing, and normalizing data.

  • Ad hoc research on project topics such as vendor trends, usage best practices, big data trends, artificial intelligence, vendors, etc. 

  • Help transition existing data sets, databases, and code to a new technology stack.

  • Assess data loads for tactical errors and build out appropriate workflows, as well as create data quality analysis that identifies larger issues in data.

  • Actively manage vendors and capture changes in data input proactively.

  • Properly prioritize and resolve data issues based on business usage.

  • Assist with managing strategic initiatives around big data projects for the commercial (trading) business.

  • Partner with commercial teams to gain understanding of current data flow, data architecture, investment process as well as gather functional requirements.

  • Assess gaps in current datasets and remediate.

Qualifications:

  • Bachelor’s degree in Computer Science, Mathematics, Physics, Business Intelligence, or related field of study.

  • 3+ years of Data Engineering experience specializing in SQL programming, data architecture, and dimension modeling.

  • Experience in energy commodities or financial services is nice to have but not required.

  • Experience in mapping, standardizing, and normalizing data.

  • Experience with ETL/ELT frameworks to write pipelines to load millions/billions of records.

  • Advanced skills in writing highly optimized SQL code.

  • Experience with relational databases; Snowflake highly preferred.

  • Hands-on experience developing data solutions in Python, Pandas, Numpy, etc.

  • Ability to communicate and interact with a wide range of data users – from very technical to non-technical.

  • Team player who is execution focused, with the ability to handle a rapidly changing set of projects. and priorities, while maintaining strong professional presence.

  • Ability to work effectively in a fast-paced, dynamic and high-intensity environment including an open-floor plan, with timely responsiveness and the ability to work beyond normal business hours when required.   

 

Employee Programs & Benefits:

CCI offers competitive benefits and programs to support our employees, their families and local communities. These include:

  • Competitive comprehensive medical, dental, retirement and life insurance benefits

  • Employee assistance & wellness programs

  • Parental and family leave policies

  • CCI in the Community: Each office has a Charity Committee and as a part of this program employees are allocated 2 days annually to volunteer at the selected charities.

  • Charitable contribution match program

  • Tuition assistance & reimbursement

  • Quarterly Innovation & Collaboration Awards

  • Employee discount program, including access to fitness facilities

  • Competitive paid time off

  • Continued learning opportunities

Visit  https://www.cci.com/careers/life-at-cci/# to learn more!

#LI-CD1

About Castleton Commodities International

Invests in and manages global energy and commodity assets.

Similar jobs

Data Engineer roles near London, England
15h
Save
Mark Applied
Hide
Senior Data Engineer - UK
London or United Kingdom
HybridFull Time
CluePoints
CluePoints: Provides AI-driven risk-based quality management software for clinical trials.
8+ YOE8+ years of hands-on data engineering and enterprise warehousing; Microsoft Fabric implementation experience; Kimball modelling, Azure, advanced T-SQL, Python or PySpark, and technical leadership skills.
Microsoft Fabric, Azure SQL, Microsoft Azure Data Factory, Microsoft Power BI, Microsoft Azure DevOps, Microsoft Azure Key Vault, T-SQL, Python, PySpark, Microsoft Purview, OneLake, Bicep, Terraform, PowerShell, Dataflows Gen2, Delta, Direct Lake, Git, CI/CD, RLS, OLS, Fabric/Azure APIs
18h
Save
Mark Applied
Hide
Senior Data Engineer (Contract - 3 days per week)
London, England, United Kingdom
OnsitePart Time, Contract
VCCP
VCCP: A global integrated creative and marketing agency network.
Senior data engineering expertise with Microsoft Fabric, Python, SQL, data ingestion, platform architecture, CI/CD, automated testing, version control, and production data delivery experience.
Microsoft Fabric, OneLake, Lakehouses, Warehouses, Pipelines, Data Factory, Dataflows, Python, PySpark, Pandas, SQL, REST, GraphQL, Git, Azure DevOps, GitHub Actions, Google Cloud Platform (GCP), BigQuery, Cloud Storage, Dataflow
1d
Save
Mark Applied
Hide
Staff Data Engineer
London or New York City
$260k-$274k/yr HybridFull Time
MoonPay
MoonPay: Global payments infrastructure for digital currency and blockchain.
Significant experience designing large-scale data platforms and distributed systems, leading cross-team initiatives, building shared capabilities, and applying strong architectural, reliability, governance, security, and cost judgment.
GCP, Pub/Sub, Kafka, Apache Beam, Bigtable, Redis, Kubernetes, Airflow, BigQuery, Terraform, Python, SQL, Claude, ChatGPT, Gemini
1d
Save
Mark Applied
Hide
Senior Data Engineer (Microsoft Fabric)
Edinburgh or Leeds or Manchester or London or Bulgaria
HybridFull Time
CreateFuture
CreateFuture: Builds digital products and provides technology consulting services.
Senior data engineering experience with Microsoft Fabric, Python, PySpark, production data pipelines, CI/CD management, teamwork, stakeholder collaboration, and technical guidance.
Microsoft Fabric, Python, PySpark, Azure OneLake, Azure, Azure DevOps, Azure ML
1d
Save
Mark Applied
Hide
Data Engineer - Axion
Sunnyvale or Washington or San Diego or Fort Walton Beach or Ann Arbor or London or Stuttgart or Munich or Stockholm or Bangalore or Seoul or Tokyo
$200k-$275k/yr OnsiteFull Time
Applied Intuition
Applied Intuition: Developing software and simulation infrastructure for autonomous vehicles.
5+ YOERequires 5+ years of relevant experience, modern ML infrastructure, large-scale GPU jobs, data software, microservices or databases, U.S. citizenship, and eligibility for security clearance.
React, TypeScript, Python, Golang, Docker, Kubernetes, Opensearch, Postgres, LLMs, VLMs, GPU
1d
Save
Mark Applied
Hide
Data Engineer
Telford or Worthing
OnsiteFull Time
Capgemini
CapgeminiEuronext Paris: CAP: Provides global IT consulting and digital transformation services.
Experienced data engineer with strong SQL, data modeling, ETL/ELT, databases, cloud platforms, programming, scripting, data engineering fundamentals, and consultancy skills; SC clearance required.
SQL, Talend, Pentaho DI, Informatica, AWS Glue, SAS, Oracle, Cloudera, AWS, Python, Bash
1d
Save
Mark Applied
Hide
Senior Data Engineer (Remote within the UK)
London, England, United Kingdom
£70k-£80k/yr RemoteFull Time
Chip
Chip: Mobile app for automated personal savings and wealth management.
Experienced data engineer with deep Databricks expertise, strong Python, PySpark, SQL, Git, Terraform, data modelling, governance, architecture, database, and mentoring skills.
Databricks, Infrastructure as Code, Unity Catalog, Git, Python, PySpark, SQL, Redshift, Postgres, MongoDB, Terraform, Docker, Kubernetes, AWS, GCP
1d
Save
Mark Applied
Hide
Software Engineer - Data, Lakehouse and AI Data Platform Engineer - Vice President - London
London, England, United Kingdom
OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities, and investment management firm.
Bachelor’s or master’s degree or equivalent experience; strong Python or Java, SQL, data modeling, production pipelines, distributed processing, data quality, and software engineering practices.
Python, Java, SQL, Apache Spark, JSON, Avro, Parquet, Kafka, Snowflake, Apache Iceberg, Databricks, Hadoop, Sybase IQ, Kubernetes, CI/CD