Capgemini
Posted 2mo ago

AI Data Engineer (Data Engineering, Cloud Platform, Python) (Warszawa, PL)

Capgemini
Warsaw, Masovian Voivodeship, Poland
HybridFull Time
Responsibilities
  • developing pipelines
  • processing data
  • supporting AI
Requirements
  • 3+ years data engineering experience
  • Strong Python and SQL
  • Experience with PySpark/Spark
  • Kafka/Hadoop
  • Databricks/Snowflake/BigQuery/Redshift
  • Cloud (AWS/Azure/GCP)
  • Airflow, dbt
  • Docker
  • Kubernetes, CI/CD
  • Strong analytical and communication skills
Technical tools mentioned
PythonSQLPySparkApache SparkKafkaHadoopDatabricksSnowflakeBigQueryRedshiftAWSAzureGoogle Cloud PlatformAirflowdbtDockerKubernetesCI/CDLLM FrameworksVector Databases

Job description

At Capgemini Invent, we believe difference drives change. As inventive transformation consultants, we blend our strategic, creative and scientific capabilities, collaborating closely with clients to deliver cutting-edge solutions. Join us to drive transformation tailored to our client's challenges of today and tomorrow. Informed and validated by science and data. Superpowered by creativity and design. All underpinned by technology created with purpose.

Your role

As an AI Data Engineer, you will develop and maintain scalable data pipelines and AI-ready cloud infrastructure to support analytics, machine learning, and business intelligence solutions. You will work closely with engineering and AI teams to ensure reliable, secure, and high-quality data processing across multiple systems and platforms.

Your project

You will contribute to enterprise data transformation initiatives focused on modernizing data architecture, building cloud-native platforms, and supporting AI/ML applications. The project includes integration of multiple data sources, automation of workflows, and optimization of data processing systems.

Your client

Our client is an innovative organization focused on leveraging data and AI technologies to improve business operations and customer experiences. The company offers a collaborative environment with opportunities to work on modern cloud and AI technologies.

Your tasks

  • Develop and maintain ETL/ELT pipelines for enterprise data platforms
  • Process and transform large-scale structured and unstructured datasets
  • Support AI and machine learning data preparation workflows
  • Integrate data from APIs, databases, and cloud services
  • Monitor and improve data quality, reliability, and pipeline performance
  • Collaborate with Data Scientists, Analysts, and Engineering teams
  • Support deployment and automation of cloud-based data solutions
  • Troubleshoot and resolve production data issues

Your profile

  • 3+ years of experience in Data Engineering or related roles
  • Strong programming skills in Python and SQL
  • Experience with PySpark, Apache Spark, Kafka, or Hadoop
  • Familiarity with Databricks, Snowflake, BigQuery, or Redshift
  • Knowledge of AWS, Azure, or Google Cloud Platform
  • Experience with Airflow, dbt, Docker, and Kubernetes
  • Understanding of CI/CD pipelines and version control tools
  • Familiarity with Vector Databases and LLM Frameworks is a plus
  • Strong analytical and communication skills
  • Ability to work in a fast-paced and collaborative environment

What You'll love about working here

  • Well-being culture: medical care with Medicover, private life insurance, and Sports card. But we went one step further by creating our own Capgemini Helpline offering therapeutical support if needed and the educational podcast "Let's talk about wellbeing" which you can listen to on Spotify.
  • Access to over 70 training tracks with certification opportunities (e.g., GenAI, Excel, Business Analysis, Project Management) on our NEXT platform. Dive into a world of knowledge with free access to Education First languages platform, TED Talks and Udemy Business materials and trainings.
  • Continuous feedback and ongoing performance discussions thanks to our performance management tool GetSuccess supported by a transparent performance management policy.
  • Enjoy hybrid working model that fits your life - after completing onboarding, connect work from a modern office with ergonomic work from home, thanks to home office package (including laptop, monitor, and chair). Ask your recruiter about the details.

Get to know us

Capgemini is committed to diversity and inclusion, ensuring fairness in all employment practices. We evaluate individuals based on qualifications and performance, not personal characteristics, striving to create a workplace where everyone can succeed and feel valued.

Do you want to get to know us better? Check our Instagram — @capgeminipl or visit our Facebook profile — Capgemini Polska. You can also find us on YouTube. 

About Capgemini

Capgemini is an AI-powered global business and technology transformation partner, delivering tangible business value. We imagine the future of organizations and make it real with AI, technology and people. With our strong heritage of nearly 60 years, we are a responsible and diverse group of over 420,000 team members in more than 50 countries. We deliver end-to-end services and solutions with our deep industry expertise and strong partner ecosystem, leveraging our capabilities across strategy, technology, design, engineering and business operations.

About Capgemini

Provides global IT consulting and digital transformation services.

Similar jobs

AI Data Engineer roles near Warsaw, Masovian Voivodeship
3w
Save
Mark Applied
Hide
Data & AI Engineer_Full Stack Developer with Italian
Warsaw or Krakow or Gdansk or Wroclaw
HybridFull Time
DXC Technology
DXC TechnologyNYSE: DXC: Global provider of IT services and business technology solutions.
Mid-level Data & AI engineer/full-stack developer with Python, SQL, cloud and GenAI experience and strong Italian for client communication.
Python, SQL, Azure, AWS, GCP, Databricks, Snowflake, Azure Data Factory, Git, GitHub Copilot, Cursor, Claude Code, Airflow, dbt, Matillion, Docker, Kubernetes, Azure DevOps, GitHub Actions, JavaScript, TypeScript, React, Angular, Vue, REST, GraphQL, LLMs, embeddings, RAG
2mo
Save
Mark Applied
Hide
AI Data Engineer (Snowflake + AI)
Warsaw, Mazowieckie, Poland
HybridFull Time
Nordea
NordeaNasdaq Helsinki: NDA: Leading Nordic financial services group.
8+ YOE8+ years in data or ML engineering with production Snowflake experience; strong SQL and Snowpark; 1+ year with LLMs in production; experience with Snowflake Cortex, Anthropic/OpenAI/Azure OpenAI; Airflow, CI/CD ownership.
Snowflake, Snowpark (Python/Scala), OpenAI API, Anthropic API, Azure OpenAI, Cortex, Airflow, Bitbucket, CI/CD, Streamlit
3mo
Save
Mark Applied
Hide
AI Data Engineer
Warsaw, Masovian Voivodeship, Poland
zł10k-zł15k/mo HybridFull Time
Telemedi
Telemedi: Platform providing remote and home-based medical consultations.
Experience building data pipelines and automated reporting; practical knowledge of LLMs and agentic workflows; independent with business mindset.
SQL, Python, TypeScript, Power BI, Metabase, Streamlit, LLM
4mo
Save
Mark Applied
Hide
AI Data Engineer
Warsaw, Mazowieckie, Poland
zł120k-zł222k/yr OnsiteFull Time
IQVIA
IQVIANYSE: IQV: Provides clinical research and data analytics for pharmaceutical companies.
3+ YOEMid-level data engineer with 3+ years experience; strong Python, SQL, NoSQL; data pipelines, ETL/ELT, governance; cloud; vector stores; CI/CD.
Python, Java, Scala, SQL, NoSQL, Data Warehousing, Lakehouse, Cloud (AWS/Azure/GCP), CI/CD, Docker, Kubernetes, Vector databases
9mo
Save
Mark Applied
Hide
Data & AI Engineer with MS Fabric or Databricks - Junior/Mid/Senior - ICH Europe
Warsaw or Kraków
HybridFull Time
Accenture
AccentureNYSE: ACN: Global provider of management consulting and technology services.
2+ YOEMinimum 2 years working with relational/analytical databases and SQL, 1 year designing or implementing ETL/ELT using Databricks or Microsoft Fabric, data modelling experience, programming in Python/R/SAS, Spark exposure, GenAI experience, English proficiency, willingness to travel across Europe.
Databricks, Microsoft Fabric, Python, R, SAS, SQL, Spark, GCP, Azure, AWS, Snowflake, Hive, NiFi, HBase, HDFS, Kafka, Kudu, GenAI
1y
Save
Mark Applied
Hide
Data Scientist / AI Engineer
Warsaw, Mazowieckie, Poland
RemoteFull Time
Addepto
Addepto: Leading AI consulting and data engineering delivering scalable AI solutions
6+ YOE6+ years in designing and implementing scalable AI solutions; lead end-to-end ML projects; strong Python and ML libraries; cloud deployment (AWS/Azure); LLMs; SQL/NoSQL; CI/CD; MLOps; Kubernetes/Docker; advanced degree.
Python, Scikit-Learn, PyTorch, TensorFlow, AWS, Azure, SQL, NoSQL, MongoDB, Snowflake, Databricks, CI/CD, GitHub, GitHub Actions, MLOps, Kubernetes, Docker, Big Data, Spark, Hadoop, Kafka, LLMs, NLP, GenAI
9mo
Save
Mark Applied
Hide
Data and AI Solutions Engineer (Warszawa, PL, 00-124)
Warsaw, Poland
OnsiteFull Time
EY
EY: Global firm providing audit, tax, and professional consulting services.
2+ YOEStrong English, 2-6 years Data & AI experience, bachelor's degree, data engineering, Python/SQL, cloud/DataOps, and AI/LLM experience.
Azure OpenAI, Databricks, Microsoft Fabric, Python, PySpark, SQL, Power BI, Snowflake, Git, Jira, Confluence, Azure DevOps, Kubernetes, REST API