This job has expired

This job posting is no longer active and is not accepting applications. Explore similar roles below!

Goldman Sachs
Posted 3w ago

Software Engineering, Lakehouse and AI Data Platform Engineer, Analyst, Singapore

Goldman Sachs
Singapore
OnsiteFull Time
Responsibilities
  • building pipelines
  • curating datasets
  • ensuring quality
Requirements
  • Design and build production data pipelines and curated datasets using Python/Java
  • SQL and distributed processing
  • Apply data modelling
  • Quality
  • Reconciliation and CI/CD practices
Technical tools mentioned
PythonJavaSQLApache SparkJSONAvroParquetCI/CD

Job description

The Opportunity

Join a team building the data foundations that support the firm’s AI and analytics capabilities. This role sits within the engineering effort to develop a modern Lakehouse and AI data platform that enables reliable, well-governed and high-performing data use across the firm.

At Goldman Sachs, engineering teams are positioned at the centre of the business, building scalable systems, solving complex technical problems and turning data into action. In data engineering roles, the emphasis is on designing, building and maintaining large-scale data platforms, delivering production pipelines, improving reliability and quality, and partnering closely with users of the platform.

This is a delivery-focused role for engineers who want to build robust data assets in production, work with modern data technologies, and grow over time within the firm. You will contribute to the data models, pipelines and platform capabilities that underpin analytics, operational decision-making and emerging AI use cases, and may also help extend platform tooling where additional functionality is needed.

 

Role Summary

As a Data Engineer in the Lakehouse and AI Data Platform team, you will design, build, test and support data pipelines and curated datasets on the firm’s modern data platform. You will work across ingestion, transformation, modelling, optimisation and data quality, helping to deliver data products that are reliable, scalable and fit for purpose.  Where there are gaps in platform functionality, you may also contribute to shared tooling or framework components that improve how the platform is used and operated.

The role is suited to engineers who are comfortable writing code, working with SQL and distributed data processing, and solving practical delivery problems in a team environment. More experienced candidates may also contribute to technical design, platform standards and the shaping of delivery approaches across a wider set of use cases.

 

Key Responsibilities

Pipeline Engineering

  • Build, enhance and support batch and streaming data pipelines on the Lakehouse and AI data platform.
  • Refactor or modernise existing data flows where needed to improve reliability, performance and maintainability.
  • Where needed, build reusable tooling to improve delivery, consistency and operational support.
  • Ensure data pipelines are production-ready, well tested and operationally supportable.

Data Modelling and Curation

  • Develop raw, refined and curated datasets that support analytics, reporting and AI use cases.
  • Apply sound data modelling principles to represent business entities, relationships and historical change accurately.
  • Work with consumers to shape data products that are usable, well documented and aligned to business needs.

Data Quality and Reconciliation

  • Implement controls to validate completeness, accuracy and consistency of data across pipelines and datasets.
  • Use reconciliation approaches to build confidence in production outputs and investigate breaks where they arise.
  • Contribute to clear standards for testing, monitoring and issue resolution.
  • Contribute to practical improvements in testing, monitoring or reconciliation tooling where these strengthen platform reliability and day-to-day delivery.

Delivery and Partnership

  • Work closely with engineers, platform teams and data consumers to deliver agreed outcomes to time and quality expectations.
  • Communicate clearly on progress, risks, dependencies and design choices, including where delivery would benefit from improvements to shared platform tooling.
  • For more senior candidates, take a broader role in technical leadership, task breakdown and support for junior engineers.

 

Skills and Experience

Required

  • Bachelor’s or master’s degree in a relevant discipline, or equivalent practical experience, with evidence of strong quantitative skills or data engineering expertise.
  • Strong hands-on programming experience in Python or Java.
  • Good working knowledge of SQL, including troubleshooting, optimisation and data analysis.
  • Ability to learn new tools, internal platforms and delivery workflows quickly.
  • Familiarity with software engineering fundamentals, including version control, testing, release discipline and CI/CD practices.

Data Engineering Capability

  • Understanding of temporal data modelling, including the handling of historical state and change over time.
  • Knowledge of schema design, schema evolution and data compatibility considerations.
  • Understanding of partitioning, clustering and other techniques used to improve data performance at scale.
  • Ability to make sensible design choices across normalised and denormalised models, and between natural and surrogate keys.
  • Practical approach to data quality, reconciliation and root-cause analysis.
  • Experience building or supporting production data pipelines in a collaborative engineering environment.
  • Experience working with distributed data processing frameworks such as Apache Spark.
  • Working knowledge of common data formats such as JSONAvro and Parquet.

 

 

 

ABOUT GOLDMAN SACHS

At Goldman Sachs, we commit our people, capital and ideas to help our clients, shareholders and the communities we serve to grow. Founded in 1869, we are a leading global investment banking, securities and investment management firm. Headquartered in New York, we maintain offices around the world.

We believe who you are makes you better at what you do. We're committed to fostering and advancing diversity and inclusion in our own workplace and beyond by ensuring every individual within our firm has a number of opportunities to grow professionally and personally, from our training and development opportunities and firmwide networks to benefits, wellness and personal finance offerings and mindfulness programs. Learn more about our culture, benefits, and people at GS.com/careers.

We’re committed to finding reasonable accommodations for candidates with special needs or disabilities during our recruiting process. Learn more: https://www.goldmansachs.com/careers/footer/disability-statement.html

 

© The Goldman Sachs Group, Inc., 2025. All rights reserved.

Goldman Sachs is an equal opportunity employer and does not discriminate on the basis of race, color, religion, sex, national origin, age, veterans status, disability, or any other characteristic protected by applicable law.

About Goldman Sachs

Provides investment banking, securities, and wealth management services globally.

Similar jobs

Data Engineer roles
16h
Save
Mark Applied
Hide
Data Engineer Graduate (Data Platform - Global Live) - 2027 Start
San Jose or Los Angeles County or Los Angeles or Singapore or New York City or London or Dublin or Paris or Berlin or Dubai or Jakarta or Seoul or Tokyo
$128k-$317k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Bachelor's degree in computer science, computer engineering, or related field; data solutions experience; coding experience with SQL and Python; analytical problem-solving skills; and data warehousing knowledge.
SQL, Python, LLM
1d
Save
Mark Applied
Hide
Data Engineer -- Regulatory Reporting & Portfolio Intelligence
Bangalore or San Francisco or Singapore
HybridFull Time
Nium
Nium: Provides infrastructure for real-time cross-border payments and card issuance
3+ YOE3–5 years in data engineering, ETL/ELT, or systems integration; strong SQL and Python; REST/SOAP, XML/JSON, cloud data platforms, orchestration, data modeling, quality testing, and regulatory reporting knowledge.
SQL, Python, XML, REST API, SOAP, JSON, AWS Redshift, Snowflake, Google BigQuery, Amazon S3, Google Cloud Storage, Apache Airflow, Prefect, Kafka, AWS
3d
Save
Mark Applied
Hide
Data Engineer / Senior Data Engineer
Singapore, Central Singapore, Singapore
OnsiteFull Time
PatSnap
PatSnap: AI-powered platform for intellectual property and R&D intelligence.
3+ YOEBachelor's degree in computer science, data engineering, software engineering, or related fields; 3+ years of data development; Java or Python; Flink, Spark, Linux, data privacy, and compliance knowledge.
Java, Python, Flink, Spark, Linux
6d
Save
Mark Applied
Hide
Senior Data Engineer, Product Analytics, Payments
Singapore, Singapore, Singapore
OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
5+ YOEBachelor's degree or equivalent experience; 5 years coding in Python and SQL, designing data pipelines, machine learning operations, and data architecture; 3 years stakeholder partnership preferred.
Python, SQL, Large Language Models (LLMs), AI, Gemini Command-line Interface, Cider Agents, LLM Extensions, Agent Development Kit (ADK) Agents
1w
Save
Mark Applied
Hide
Staff Data Engineer - AI Platform
United States or North America or San Francisco or Los Angeles or New York City or Washington or London or Singapore
RemoteFull Time
TRM Labs
TRM Labs: An AI-powered intelligence helping agencies investigate crime and disrupt illicit activity.
Experience operating distributed OLAP or serving-layer systems, tuning queries, owning pipeline reliability, responding to incidents, using AI tools, and taking on-call responsibility. U.S. citizenship required.
StarRocks, Claude, Trino, ClickHouse, Cursor, Slack, Otter.ai, Fireflies, Fathom, Cluey
1w
Save
Mark Applied
Hide
Data Engineer
Singapore, Singapore, Singapore
OnsiteFull Time
Bambu Lab
Bambu Lab: Manufacturer of desktop 3D printers and 3D printing accessories.
Extensive data warehouse modeling and development experience; big data expertise with Hadoop, Spark, Iceberg, or Flink; Java, Scala, or Python proficiency; and data privacy compliance knowledge.
Hadoop, Spark, Iceberg, Flink, Java, Scala, Python
1w
Save
Mark Applied
Hide
Data Engineer Intern, Technology (Jan - Jun 2027)
Singapore, Singapore, Singapore
OnsiteInternship
Temasek
Temasek: Global investment managing a diversified portfolio.
Currently pursuing engineering, computer science, or computer engineering degree; coding proficiency, SQL and database knowledge, software engineering concepts, and strong communication skills required. Cloud, big data, CI/CD, and visualization experience preferred.
SQL, AWS, airflow, spark, git, bitbucket, Tableau, Qlik
1w
Save
Mark Applied
Hide
Data Engineer
Singapore, Singapore, Singapore
OnsiteFull Time
Accenture
AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
2+ YOEBachelor's degree in a relevant field, 2+ years of data engineering experience, SQL and Python proficiency, ETL/ELT pipeline experience, cloud platform experience, and data engineering best practices.
SQL, Python, AWS, Microsoft Azure, Google Cloud Platform (GCP), Apache Spark, Tableau, Pandas Python Library
This job has expired