This job has expired

This job posting is no longer active and is not accepting applications. Explore similar roles below!

Citi
Posted 1d ago

Data Engineer, Java, Apache Spark, Hive - Assistant Vice President

Citi
Pune, Maharashtra, India
HybridFull Time
Responsibilities
  • building pipelines
  • developing aggregations
  • designing APIs
Requirements
  • Requires 5+ years of relevant experience with Core Java
  • Apache Spark
  • Big data technologies, and Hive
  • Strong SQL, OLAP
  • Systems analysis
  • Project delivery, and stakeholder communication skills
Technical tools mentioned
JavaScalaApache SparkHiveApache PinotApache DruidTrinoSQLNatural Language Processing (NLP)OLAP

Job description

Discover your future at Citi

Working at Citi is far more than just a job. A career with us means joining a team of approximately 219,000 dedicated people from around the globe. At Citi, you’ll have the opportunity to grow your career, give back to your community and make a real impact.

Job Overview

The Opportunity: Are You Ready to Build the Future of Risk Analytics?


Want to solve one of the most challenging big data problems in finance today? Are you excited by the idea of working directly with front-office and risk quants on massive-scale analytics for critical regulations like FRTB? Do you want to tame billions of rows of complex financial data with Apache Spark and deliver insights to senior leaders in a fraction of a second?
Citi is looking for an elite, hands-on technologist to engineer the next-generation analytics platform for our Market Risk organization. This isn't just another data job. You will be the architect and builder of a high-performance system that ingests, processes, and serves petabytes of risk calculation data, making it instantly accessible and understandable to the entire firm.

Your Role and Impact
As a Senior Technologist, you will be at the heart of the action, designing the systems that power our most sophisticated risk-based calculations. You will take the raw, trade-level Present Value (PV) outputs from our Historical VaR and FRTB Expected Shortfall engines and transform them into a strategic data asset.
Your impact will be immediate and far-reaching. You will build the data pipelines that handle immense volumes, create intelligent APIs that democratize data access, and leverage cutting-edge OLAP and NLP technologies to provide unparalleled drill-down and analytical capabilities. You will be the go-to expert who empowers senior stakeholders in the Markets and Risk organizations to make faster, smarter decisions.

Key Responsibilities

  • Architect and build robust, scalable data pipelines to ingest and process billions of trade-level PV calculations from various stress engines.

  • Develop and optimize large-scale aggregation jobs using Apache Spark, ensuring high performance and efficiency.

  • Design and deliver a suite of "intelligent data APIs" that provide flexible, on-demand access to both aggregated and non-aggregated risk data for teams across the firm.

  • Integrate Natural Language Processing (NLP) capabilities to create intuitive, query-based interfaces for data exploration, lowering the barrier to entry for complex analytics.

  • Load and model massive aggregated datasets into high-performance OLAP engines like Apache Pinot, Apache Druid, and Trino.

  • Build powerful, interactive analytical tools and dashboards on top of the OLAP layer, providing summary views and lightning-fast drill-down capabilities.

  • Partner directly with senior stakeholders in the Front Office, Quantitative teams, and Risk Management to understand their analytical needs and deliver innovative solutions.

What We're Looking For

  • A true passion for data, analytics, and solving complex problems at massive scale.

  • A degree in a quantitative or technical field such as Computer Science, Financial Mathematics, or Financial Engineering.

  • Expert-level, hands-on experience with big data technologies, particularly Apache Spark.

  • Proven experience with high-performance OLAP databases such as Apache Pinot, Apache Druid, or Trino.

  • Strong programming skills in Java and/or Scala, and expert-level SQL.

  • Strong background in fundamental computer science concepts, including data structures and algorithms.

  • A mindset for 'AI-first' development, constantly looking for ways to embed intelligence into systems.

  • Experience or a strong interest in applying Natural Language Processing (NLP) to data access and analytics.

  • A knack for solving 'needle in a haystack' problems, with a talent for debugging complex data and access issues.

  • Experience in the financial industry with an understanding of market risk, derivatives, and risk calculations (VaR, Stress Testing, PV) is highly desirable.

  • Exceptional problem-solving skills and the ability to work independently and lead technical projects.

  • Excellent communication skills, with the confidence to collaborate with senior business and quantitative stakeholders.    

Qualifications:

  • 5+ years of relevant experience using Core Java, Apache Spark, Big Data Technologies HDFC, Hive etc.

  • Experience in systems analysis and programming of software applications

  • Experience in managing and implementing successful projects

  • Working knowledge of consulting/project management techniques/methods

  • Ability to work under pressure and manage deadlines or unexpected changes in expectations or requirements

------------------------------------------------------

Job Family Group:

Technology

------------------------------------------------------

Job Family:

Applications Development

------------------------------------------------------

Time Type:

Full time

------------------------------------------------------

Most Relevant Skills

Please see the requirements listed above.

------------------------------------------------------

Other Relevant Skills

For complementary skills, please see above and/or contact the recruiter.

------------------------------------------------------

Citi is an equal opportunity employer, and qualified candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other characteristic protected by law.

If you are a person with a disability and need a reasonable accommodation to use our search tools and/or apply for a career opportunity review Accessibility at Citi.

View Citi’s EEO Policy Statement and the Know Your Rights poster.

About Citi

Global financial services organization enabling growth and economic progress.

Similar jobs

Data Engineer roles near Pune, Maharashtra
23h
Save
Mark Applied
Hide
Data Engineer-Data Platforms-SnowFlake
Pune or Bangalore
HybridFull Time
IBM
IBMNYSE: IBM: Global technology and consulting focusing on cloud and AI.
Bachelor's degree required; master's preferred. Experience with Snowflake, data engineering, cloud computing, and implementing data and AI use cases; technical solution delivery experience required.
Snowflake
1d
Save
Mark Applied
Hide
Associate Data Engineer
Pune, Maharashtra, India
OnsiteFull Time
Ecolab
EcolabNYSE: ECL: Provider of water, hygiene, and infection prevention solutions.
3+ YOERequires 3–5 years in data engineering, analytics, or business intelligence; advanced SQL and Python; Snowflake, Azure, ETL/ELT, data governance, Agile/Scrum, and data architecture experience.
Snowflake, SQL, Microsoft Power BI, Microsoft Azure, DBT, SQL Server, Microsoft Logic Apps, Microsoft App Services, Microsoft Data Factory, Lakehouse, Warehouse, Python, PySpark, Five Tran, Streamlit, Flask, Node.js, Microsoft Power Apps, Graph API, GIT, SAP, GEP, SCRUM, Agile
1d
Save
Mark Applied
Hide
Data Engineer
Bengaluru or Pune or New Delhi
OnsiteFull Time
Merkle
Merkle: Merkle is a dentsu-owned global experience consultancy delivering data, technology, marketing, and customer-experience services to brands.
3+ YOERequires 3–7 years of data engineering experience, including 1+ year with GCP; Node.js, Python, SQL, data modeling, warehousing, CI/CD, and GCP pipeline expertise. Bachelor's or master's degree required.
Google Cloud Platform (GCP), Node.js, Python, BigQuery, Cloud Functions, Cloud Run, Cloud Build, Dataform, Pub/Sub, Eventarc, Cloud Storage, Cloud Composer, SQL, Git, JIRA, Microsoft Office, Snowflake, CI/CD
1d
Save
Mark Applied
Hide
Sr. Data Engineer
Pune or Coimbatore
OnsiteFull Time
Avantor
AvantorNYSE: AVTR: Global provider of mission-critical products for life sciences.
5+ YOEDegree in supply chain management or APICS CPIM with 5–7+ years in global supply chain and inventory management. Requires analytics, project management, planning systems, ETL, and data modeling expertise.
Microsoft Excel, Power BI, Power Query/M, DAX, R, Python, SQL, Codex, SAP APO/IBP, OM Partners, Snowflake, SAP, .csv, .xlsx, .txt, .parquet, .rds
2d
Save
Mark Applied
Hide
Databricks Data Engineer - Contract-to-Hire
Pune, Maharashtra, India
HybridFull Time, Contract
Customized Energy Solutions
Customized Energy Solutions: Private energy services and technology helping utilities, generators, suppliers, developers, and investors navigate deregulated energy markets.
4+ YOEBachelor's degree and 4–7 years of data engineering experience. Requires Databricks, Spark, PySpark or Scala, SQL, SQL Server migration, SSIS ETL modernization, data warehousing, Delta Lake, and performance tuning.
Databricks, Apache Spark, Delta Lake, Unity Catalog, Microsoft SQL Server, SSIS, SSRS, SSAS, Databricks notebooks, Databricks Workflows, Databricks Jobs, PySpark, Scala, SQL, Databricks SQL, Power BI, MLflow
2d
Save
Mark Applied
Hide
Azure Senior Data Engineer
Pune or Bengaluru or National Capital Region
OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesBSE: 532540: Global leader in IT services, consulting, and business solutions.
8+ YOEBachelor of Technology and 8–15 years of experience required. Must design data pipelines using PySpark and Databricks, manage large datasets and Cosmos DB, and support automation, testing, and deployment.
PySpark, Databricks, Cosmos DB, Python, Azure
2d
Save
Mark Applied
Hide
Data Engineer, Java, Apache Spark, Hive - Assistant Vice President
Pune, Maharashtra, India
HybridFull Time
Citi
CitiNYSE: C: Global financial services organization enabling growth and economic progress.
5+ YOE5+ years using Core Java, Apache Spark, and big data technologies; strong SQL and OLAP experience; technical degree preferred; experience with systems analysis, project delivery, and financial risk analytics.
Java, Apache Spark, Hive, Scala, SQL, Apache Pinot, Apache Druid, Trino, Natural Language Processing (NLP), OLAP
3d
Save
Mark Applied
Hide
Senior Data Engineer
Pune, Maharashtra, India
OnsiteFull Time
Snowflake
SnowflakeNYSE: SNOW: Cloud-based data platform for AI, analytics, and data applications.
5+ YOEBachelor's degree or equivalent experience and 5–8 years building production data pipelines, models, and platform infrastructure; expert SQL, strong Python, Snowflake, dbt, and Apache Airflow skills required.
SQL, Python, Snowflake, Snowpark, Snowflake Cortex, dbt, Apache Airflow, MLflow, Feature Stores, vector databases, LLM, AWS, Azure, GCP, Kafka, Snowpipe Streaming, schema registry
This job has expired