This job has expired

This job posting is no longer active and is not accepting applications. Explore similar roles below!

📋 External Recruiting Agencies

Jobgether is a remote job matching platform and talent marketplace; the job listing explicitly states it is posted on behalf of an anonymous partner company, making Jobgether a recruiting intermediary rather than the direct employer.

This company was flagged and excluded from default search results. Proceed with caution.

J
Posted 1w ago

Data Engineer

Jobgether
India
$150k-$170k/yrRemoteFull Time
Responsibilities
  • building pipelines
  • developing workflows
  • monitoring systems
Requirements
  • Requires 3+ years in data engineering or related work
  • Strong Python and SQL
  • Airflow and dbt expertise
  • API integrations
  • Production pipeline ownership
  • Data modeling
  • Monitoring, and communication skills
Technical tools mentioned
PythonSQLAirflowdbtBigQueryGCPCloud Composer

Job description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Data Engineer based in India.

This is an opportunity to build and own the data infrastructure that powers critical business decisions, reporting, operations, internal tools, and AI initiatives. You’ll take complex business needs and turn them into reliable, scalable data products from source systems through production. The role combines hands-on engineering with end-to-end ownership of pipelines, data quality, and production operations. You’ll work across Python, SQL, Airflow, dbt, APIs, and modern cloud data infrastructure. You’ll collaborate closely with Finance, Operations, Growth, and Technology teams in a fast-moving environment where priorities evolve quickly. Success means creating trusted systems that are resilient, well-documented, and increasingly efficient through thoughtful use of AI-powered development tools.



Accountabilities
  • Design, build, and own custom data pipelines that integrate source systems with the data warehouse, including complex or poorly documented APIs.
  • Develop and operate Airflow DAGs for scheduled and event-driven workflows, ensuring reliable and scalable data processing.
  • Build dbt models and governed semantic layers that provide trusted data for analysts, internal applications, and AI tools.
  • Create reusable patterns for incremental synchronization, retries, replay, backfills, and event processing to improve engineering efficiency.
  • Implement monitoring, data-quality checks, reconciliation, alerting, and logging to identify issues before they affect reporting or downstream systems.
  • Diagnose and resolve production incidents involving missing, delayed, duplicated, or inaccurate data, taking ownership through resolution.
  • Partner with Finance, Operations, Growth, and Technology stakeholders to translate ambiguous business needs into durable data solutions.
  • Use AI coding agents to accelerate implementation, testing, debugging, refactoring, and review while maintaining responsibility for technical quality and correctness.
  • Document architectural decisions, system ownership, operational procedures, and recovery processes to ensure systems remain maintainable and supportable.
  • Establish scalable engineering patterns and best practices that help the broader team deliver reliable data products faster.
  • Requirements

    • 3+ years of experience in data engineering, analytics engineering, or a closely related engineering role, with demonstrated ownership beyond execution-focused responsibilities.
    • Strong production-level Python and advanced SQL skills, with an emphasis on clean, maintainable, and well-tested code.
    • Hands-on experience designing, building, and operating Airflow workflows for custom data pipelines.
    • Strong knowledge of dbt and dimensional modeling, including grain, keys, facts, dimensions, incremental models, testing, documentation, and schema evolution.
    • Experience developing custom integrations using third-party APIs, webhooks, files, or event streams rather than relying exclusively on managed connectors.
    • Deep understanding of API ingestion, including authentication, pagination, rate limits, incremental synchronization, retries, and historical backfills.
    • Experience owning business-critical pipelines throughout their lifecycle, including design, production support, monitoring, alerting, logging, reconciliation, CI/CD, and runbooks.
    • Experience working in a high-growth environment where priorities, business requirements, and source systems change frequently; DTC or subscription experience is a plus.
    • Broad, hands-on understanding of the modern data stack rather than specialization in only one technical area.
    • Strong analytical and problem-solving abilities, with a habit of investigating discrepancies until the underlying cause is identified.
    • Ability to translate business decisions and requirements into appropriate technical solutions while challenging unclear definitions or conflicting expectations.
    • Strong ownership, communication, documentation, and collaboration skills, with the ability to work independently in a remote environment.
    • Fluency with AI coding agents for implementation, testing, debugging, refactoring, and review, combined with the judgment to validate and correct AI-generated output.
    • Experience with BigQuery, GCP, or Cloud Composer is a plus.
    • Benefits

      • Competitive base salary of $150,000–$170,000, plus a bonus, depending on experience and location.
      • Comprehensive benefits package designed to support employee and family well-being.
      • Strong healthcare and benefits coverage.
      • Generous paid time off.
      • Wellness-focused perks and access to company products.
      • Fully remote, high-trust working environment.
      • Opportunities for significant ownership, impact, and career growth in a fast-scaling organization.
      • Biannual company off-sites providing opportunities to connect with colleagues in person.
      • Opportunity to work with modern data infrastructure and AI-powered engineering tools.


How Jobgether works:
We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.
We appreciate your interest and wish you the best!
 Why Apply Through Jobgether? 
 
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.
 
 
#LI-CL1

Similar jobs

Data Engineer roles
6h
Save
Mark Applied
Hide
Data Engineer-Data Platforms-AWS
Pune, Maharashtra, India
HybridFull Time
IBM
IBMNYSE: IBM: Global technology and consulting focusing on cloud and AI.
Bachelor's degree required; master's preferred. Experience with AWS data services, batch and real-time pipelines, data layers, open-source technologies, and data service management.
AWS EMR, AWS Glue, Glue Catalog, Kinesis, AWS Redshift, Aurora, DynamoDB, AWS DMS, Managed Streaming for Apache Kafka, Apache Airflow, dbt, Spark, Python, Scala, AWS Glue Databrew, Lambda, Redshift Spectrum
6h
Save
Mark Applied
Hide
Data Engineer II
Hyderabad, Telangana, India
OnsiteFull Time
Electronic Arts
Electronic Arts: Global leader in interactive entertainment, publishing video games and online services.
Bachelor's degree in computer science or equivalent training; experience with data pipelines, databases, lakehouse architectures, ETL platforms, cloud providers, LLMs, AI frameworks, and visualization tools.
MySQL, MongoDB, Cassandra, Parquet, Iceberg, Avro, BigQuery, BigLake, DeltaLake, Snowflake, Redshift, Airflow, Python, Java, Go, AWS, GCP, Azure, LangChain, LangGraph, Google ADK, Claude Code, GitHub Copilot, Looker, Streamlit, Gradio, LLMs
7h
Save
Mark Applied
Hide
Senior Engineer, Data Engineering
Bengaluru or Bangalore
OnsiteFull Time
News Corp
News CorpNasdaq, ASX: NWS, NWSA, NWSLV: Global media and information services.
4+ YOERequires 4–6+ years of data engineering experience, deep GCP expertise, BigQuery, Python, SQL, Dataform, data modeling, CI/CD, infrastructure-as-code, and architectural leadership.
Google Cloud Platform (GCP), BigQuery, BigLake, PySpark, Dataproc, Python, SQL, Cloud Composer, Airflow, Terraform, CircleCI, GitHub, Dataform, Cloud Storage, Cloud Functions
10h
Save
Mark Applied
Hide
Data Engineer II
Hyderabad, Telangana, India
OnsiteFull Time
Electronic Arts
Electronic Arts: Global leader in interactive entertainment, publishing video games and online services.
Bachelor's degree in computer science or equivalent training; experience with data pipelines, databases, lakehouse architectures, cloud platforms, workflow tools, LLMs, agentic AI, and data visualization.
MySQL, MongoDB, Cassandra, Parquet, Iceberg, Avro, BigQuery, BigLake, DeltaLake, Snowflake, Redshift, Airflow, Python, Java, Go, AWS, GCP, Azure, LangChain, LangGraph, Google ADK, Claude Code, GitHub Copilot, Looker, Streamlit, Gradio
13h
Save
Mark Applied
Hide
Sr Data Engineer
Bangalore, Karnataka, India
OnsiteFull Time
Standard Chartered
Standard CharteredLondon Stock Exchange: STAN: International banking organization focused on commerce and prosperity.
10+ YOEBachelor’s or master’s degree in a technical field and 10+ years of data engineering experience. Requires Python, PySpark, shell scripting, SQL, ETL/ELT pipelines, and DevOps/CI/CD knowledge.
Python, SQL, Unix Scripting, Control-M, Airflow, DevOps, PySpark, Hive, Spark, HDFS, Iceberg, Docker, Kubernetes, Terraform
13h
Save
Mark Applied
Hide
Intern - Data Engineer (Aftersales)
Bengaluru, Karnataka, India
OnsiteInternship
Ultraviolette Automotive
Ultraviolette Automotive: Indian private manufacturer of high-performance electric motorcycles and EV platforms for global riders.
Recent B.E./B.Tech graduate in computer science, IT, data engineering, or related field; basic ETL, database, and SQL knowledge; intermediate Microsoft Excel and PowerPoint skills.
Microsoft Excel, Microsoft PowerPoint, SQL, ERP
17h
Save
Mark Applied
Hide
Data Engineer ( 2 to 5 Years)
Bangalore or Pune
OnsiteFull Time
Siemens
SiemensXetra: SIE: Global technology specializing in industry, infrastructure, transport, and healthcare.
3+ YOEBachelor's or master's degree in computer science, engineering, or related field; 3+ years of data-at-scale experience; Python, SQL, databases, CI/CD, data pipelines, and Apache Spark expertise.
AWS, Python, SQL, Apache Spark, PySpark, TypeScript, JavaScript, PostgreSQL, DynamoDB, CDK, CloudFormation, Terraform, AWS Lambda, Amazon Athena, AWS Glue, Amazon SageMaker, Amazon S3, EMR, Snowflake
17h
Save
Mark Applied
Hide
Sr. Data Engineer - Supply Chain Automation
Bengaluru, Karnataka, India
OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Multinational technology focused on e-commerce and cloud computing.
7+ YOERequires 7+ years of data engineering experience, 6+ years of professional software development, 3+ years developing supply chain solutions, programming expertise, SQL and NoSQL databases, and a relevant bachelor's degree or PhD.
Python, Java, Scala, NodeJS, SQL, EC2, Lambda, DynamoDB, Elastic Search, AWS Cloud, API, UI
This job has expired