Feathersoft
Posted 1mo ago

Data Engineer

Feathersoft
Kerala, India
OnsiteFull Time
Responsibilities
  • designing ETL
  • processing data
  • supporting cloud
Requirements
  • 3+ years experience building ETL pipelines using Python
  • PySpark
  • SQL and AirFlow
  • Experience with Redshift/PostgreSQL
  • Cloud platforms (AWS/Azure/GCP)
  • Vectorization tools (LangChain/LlamaIndex), Git, and strong problem-solving
Technical tools mentioned
PythonPySparkSQLAirFlowRedshiftPostgreSQLAWSAzureGCPLangChainLlamaIndexGit

Job description

Key Responsibilities:

  • Design, develop, and maintain ETL (Extract, Transform, Load) processes to ensure the seamless integration of raw data from various sources into our data lakes or warehouses.
  • Utilize Python, PySpark, SQL and AirFlow etc., to process, analyze, and store large-scale datasets efficiently.
  • Write and maintain SQL queries for data retrieval, transformation, and storage in relational databases like Redshift or PostgreSQL.
  • Support cloud-based data platforms such as AWS, Azure, or GCP, with a focus on orchestrating AI retraining cycles, versioning, and automated pipeline monitoring.
  • Familiarity in converting unstructured data into vectors using frameworks like LangChain or LlamaIndex and storing them.
  • Collaborate with cross-functional teams, including data scientists, ML engineers, and domain experts to design and implement scalable solutions.
  • Troubleshoot and resolve performance issues, data quality problems, and errors in data pipelines.
  • Document processes, code, and best practices for future reference and team training.

 



Requirements

Additional Information:

  • Experience level 3+ years.
  • Strong understanding of data governance, security, and compliance principles is preferred.
  • Ability to work independently and as part of a team in a fast-paced environment.
  • Excellent problem-solving skills with the ability to identify inefficiencies and propose solutions.
  • Experience with version control systems (e.g., Git) and scripting languages for automation tasks.


About Feathersoft

Provides custom software development and IT consulting services.

Similar jobs

Data Engineer roles in Kerala
3d
Save
Mark Applied
Hide
Data Engineer
Cochin, Kerala, India
OnsiteFull Time
IBS Software
IBS Software: Provides SaaS solutions for the travel and logistics industries.
Requires SQL query optimization, SQL databases, data warehousing, ETL, AWS Glue, Amazon S3, Python, data investigation, and end-to-end ownership of technical solutions.
SQL, Amazon Redshift, PostgreSQL, MySQL, ETL, AWS, AWS Glue, Amazon S3, Python
3d
Save
Mark Applied
Hide
Lead I - Data Engineering
Thiruvananthapuram, Kerala, India
OnsiteFull Time
UST
UST: Global provider of digital transformation and IT services.
5+ YOERequires 5+ years of data engineering experience with Azure Data Factory, Azure Databricks, PySpark, Python, SQL, data warehousing, data lakes, ETL/ELT, Git, and CI/CD.
Microsoft Azure Data Factory, Microsoft Azure Databricks, PySpark, Python, SQL, Microsoft SQL Server, PostgreSQL, Oracle, Microsoft Azure Data Lake Storage Gen2, Git, CI/CD, Microsoft Azure Synapse Analytics, Microsoft Fabric, Microsoft Azure Key Vault, Microsoft Azure Monitor, Kafka, Event Hubs, Power BI, Terraform
4d
Save
Mark Applied
Hide
Associate Data Engineer
Kochi, Kerala, India
OnsiteFull Time
Gallagher
GallagherNYSE: AJG: Provides insurance brokerage, risk management, and human resources consulting.
3+ YOEUniversity degree and at least 3 years of relevant experience or equivalent. Requires SQL, Python, cloud data, data transformation, distributed warehouses, data lakes, and stakeholder management.
SQL, Python, Azure Data Factory, Databricks, Spark, AWS
5d
Save
Mark Applied
Hide
Data Engineer
Kochi, Kerala, India
OnsiteFull Time
SOTI
SOTI: Software for managing and securing mobile and IoT devices.
3+ YOERequires 3–5 years in data engineering, advanced Python and PySpark, SQL, C#, ETL/ELT, Microsoft Fabric, and Azure data services including ADF, ADLS, Synapse, Databricks, and Azure SQL.
SQL, Python, PySpark, C#, Microsoft Fabric, Lakehouse, Warehouse, Pipelines, Notebooks, Azure Data Factory (ADF), Azure Data Lake Storage (ADLS), Azure Synapse Analytics, Databricks, Azure SQL, SQL Server, Microsoft Azure, Microsoft Power BI
1w
Save
Mark Applied
Hide
EY - GDS Consulting - AI And DATA -Azure Databricks-Senior
Thiruvananthapuram or India
OnsiteFull Time
EY
EY: Global firm providing audit, tax, and professional consulting services.
5+ YOERequires 5–8 years of data or platform engineering experience, Python, PySpark, SQL, Delta Lake, data modeling, cloud security, and a relevant bachelor's or master's degree; Databricks certification preferred.
Delta Lake, Lakeflow, Databricks SQL, Unity Catalog, Spark, Delta Live Tables, Lakeflow Spark Declarative Pipelines, Auto Loader, Mosaic AI, AI/BI Genie, Agent Bricks, vector search, MLflow, Databricks Assistant, Genie Code, Python, PySpark, SQL, Git, CI/CD, infrastructure as code, RAG, MCP, LLM
1w
Save
Mark Applied
Hide
Lead Data Engineer - Data Migration
Kochi, Kerala, India
OnsiteFull Time
Milestone Technologies
Milestone Technologies: Global provider of IT managed services and digital solutions.
8+ YOE3+ Mgmt8+ years data engineering experience with 3+ years leading enterprise-scale ServiceNow to Salesforce data migrations; strong SQL, ETL/ELT, Python, ServiceNow and Salesforce expertise.
Azure Data Factory, Informatica, Talend, SSIS, AWS Glue, MuleSoft, Python, SQL, Shell Scripting, ServiceNow, Salesforce Agentforce, Service Cloud, Salesforce Bulk APIs, Salesforce Data Loader, SQL Server, Oracle, PostgreSQL, Snowflake, Azure SQL, Microsoft Azure, AWS, Google Cloud Platform
2w
Save
Mark Applied
Hide
Data Engineer-Data Platforms-AWS
Kochi, Kerala, India
HybridFull Time
IBM
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
Design, build, and operate batch and real-time data pipelines on AWS; work with RedShift, Aurora, DynamoDB and data migration; use Airflow, dbt, Spark and Python/Scala; up to 20% travel.
AWS EMR, AWS Glue, Glue Catalog, Kinesis, Managed Streaming for Apache Kafka, AWS RedShift, Aurora, DynamoDB, AWS DMS, Apache Airflow, dbt, Spark, Python, Scala, AWS Glue Databrew, Lambda, RedShift Spectrum
3w
Save
Mark Applied
Hide
Data Engineer
Chennai or Kochi or Coimbatore
HybridFull Time
Orion Innovation
Orion Innovation: Provides digital transformation, software engineering, and cloud services.
10+ YOE10+ years Data Engineering experience with hands-on Databricks, PySpark, and SQL/T-SQL; ability to migrate legacy ETL logic to Databricks, strong analytical and communication skills.
Databricks, PySpark, SQL, T-SQL, SQL Server, Oracle, Azure Data Factory, Delta Lake, Unity Catalog, Azure Synapse Analytics, Azure DevOps