Amgen
Posted 3mo ago

Data Engineer

Amgen
Hyderabad, Not Specified, India
OnsiteFull Time
Responsibilities
  • Build pipelines
  • Maintain platforms
  • Ensure data quality
Requirements
  • Proficient in building data pipelines and data lakes/warehouses on cloud platforms
  • Strong SQL, Python, PySpark, and Airflow
  • Experience with ETL, data modeling, and analytics
Technical tools mentioned
AWSDatabricksSQLPythonPySparkAirflowGitCI/CDSpark

Job description

Career Category

Information Systems

Job Description

Data Engineer 

  

Role Name: Data Engineer  

Department Name: XXX 

Role GCF: 4 

Hiring Manager Name: Chunhai He

 

ABOUT AMGEN 

Amgen harnesses the best of biology and technology to fight the world’s toughest diseases, and make people’s lives easier, fuller and longer. We discover, develop, manufacture and deliver innovative medicines to help millions of patients. Amgen helped establish the biotechnology industry more than 40 years ago and remains on the cutting-edge of innovation, using technology and human genetic data to push beyond what’s known today. 

  

ABOUT THE ROLE 

Role Description: 

We seek a skilled Data Engineer to build and optimize our data infrastructure. As a key contributor, you will collaborate closely with cross-functional teams to design and implement robust data pipelines that efficiently extract, transform, and load data into our AWS-based data lake and data warehouse. Your expertise will be instrumental in empowering data-driven decision making through advanced analytics and predictive modeling. 

Roles & Responsibilities:  

  • Building and optimizing data pipelines, data warehouses, and data lakes on the AWS and Databricks platforms. 

  • Managing and maintaining the AWS and Databricks environments. 

  • Ensuring data integrity, accuracy, and consistency through rigorous quality checks and monitoring. 

  • Maintain system uptime and optimal performance 

  • Working closely with cross-functional teams to understand business requirements and translate them into technical solutions. 

  • Exploring and implementing new tools and technologies to enhance ETL platform performance. 

Functional Skills: 

Must-Have Skills: 

  • Proficient in SQL for extracting, transforming, and analyzing complex datasets from both relational and columnar data stores. Proven ability to optimize query performance on big data platforms. 

  • Proficient in leveraging Python, PySpark, and Airflow to build scalable and efficient data ingestion, transformation, and loading processes. 

  • Ability to learn new technologies quickly. Strong problem-solving and analytical skills. Excellent communication and teamwork skills. 

Good-to-Have Skills: 

  • Experienced with SQL/NOSQL database, vector database for large language models 

  • Experienced with data modeling and performance tuning for both OLAP and OLTP databases  

  • Experienced with Apache Spark, Apache Airflow 

  • Experienced with software engineering best-practices, including but not limited to version control (Git, Subversion, etc.), CI/CD (Jenkins, Maven etc.), automated unit testing, and Dev Ops 

  • Experienced with AWS, GCP or Azure cloud services 

Education and Professional Certifications 

  • Bachelor’s degree in computer science and engineering preferred, other Engineering field is considered  

  • AWS Certified Data Engineer preferred 

  • Databricks Certificate preferred 

Soft Skills: 

  • Excellent analytical and troubleshooting skills. 

  • Strong verbal and written communication skills 

  • Ability to work effectively with global, virtual teams 

  • High degree of initiative and self-motivation. 

  • Ability to manage multiple priorities successfully.  

  • Team-oriented, with a focus on achieving team goals 

  • Strong presentation and public speaking skills. 

 

EQUAL OPPORTUNITY STATEMENT 

Amgen is an Equal Opportunity employer and will consider you without regard to your race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, or disability status. 

We will ensure that individuals with disabilities are provided with reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation. 

 

.

About Amgen

Develops and manufactures biotechnology medicines for serious diseases.

Similar jobs

Data Engineer roles near Hyderabad, Not Specified
1d
Save
Mark Applied
Hide
Senior Data Engineer (50378)
Hyderabad, Telangana, India
OnsiteFull Time
Incedo
Incedo: Providing digital transformation and AI technology services.
3+ YOERequires 3–5 years of relevant experience, a B.Tech, B.E., M.Tech, or MCA, AWS data platform expertise, data pipelines, Apache Spark, Python, Java, SQL, and strong communication skills.
Amazon Web Services (AWS), AWS Glue, AWS Redshift, AWS Lambda, Apache Spark, Python, Java, SQL
1d
Save
Mark Applied
Hide
GCP Data Engineer
Hyderabad, Telangana, India
OnsiteFull Time
Mattel
MattelNASDAQ: MAT: Designs, manufactures, and markets toys and family entertainment franchises.
3+ YOEBachelor's or master's degree in a technical field and 3–5 years of data engineering experience with BigQuery, Python, SQL, DBT, Airflow, cloud data warehousing, ETL/ELT, data quality, and Agile methods.
Google BigQuery, Python, SQL, DBT, Airflow, Cloud Composer, Ascend.io, Databricks, Dataflow, Fivetran, Collibra, Git, Pub/Sub, Kafka
1d
Save
Mark Applied
Hide
Google Data Engineer
Hyderabad, Telangana, India
OnsiteFull Time
TechVedika
TechVedika: Develops AI, IoT, and cloud-based engineering solutions for enterprises.
5+ YOERequires 5+ years engineering data pipelines, 5+ years DBT and IICS, advanced SQL, 4+ years cloud technologies, data modeling, ETL, and warehousing; Azure, Git, and Google certifications preferred.
DBT (Data Build Tool), Informatica Intelligent Cloud Services (IICS), Google Cloud Storage, Cloud Pub/Sub, Cloud Data Fusion, Dataflow, Dataproc, Dataform, BigQuery, Cloud SQL, Azure, Git, Azure DevOps, Google Sheets, Google Docs, Google Slides, Gmail
1d
Save
Mark Applied
Hide
Lead Assistant Manager
Gurugram or Hyderabad
HybridFull Time
EXL
EXLNASDAQ: EXLS: Provides data analytics and digital operations solutions to businesses.
4+ YOEBachelor's degree in a related technical field, 4–8 years of data engineering experience, and 2+ years of hands-on DBT development; requires SQL, cloud data platforms, Python, and ELT expertise.
DBT (Data Build Tool), SQL, Snowflake, Databricks, Google BigQuery, Azure Synapse Analytics, AWS Redshift, Python, Shell Scripting, Airflow, Azure Data Factory, Databricks Workflows, GitHub, GitLab, DBT Cloud, Jinja Macros, Terraform, Kafka, Spark, PySpark, GitHub Copilot, Microsoft Copilot, Microsoft Power BI, Tableau, Jira Align, Trello
1d
Save
Mark Applied
Hide
Data Engineer
Hyderabad, Telangana, India
OnsiteFull Time
Accenture
AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
3+ YOERequires at least 3 years of PySpark experience, data pipeline and workflow expertise, distributed computing knowledge, troubleshooting ability, and 15 years of full-time education.
PySpark
1d
Save
Mark Applied
Hide
Senior / Lead Data Engineer - Azure Databricks
Hyderabad, Telangana, India
OnsiteFull Time
Blend360
Blend360: Provides data science, AI, and marketing consulting services.
4+ YOERequires 4+ years in data engineering, SQL and Python, Azure Databricks, semantic and KPI modeling, data pipelines, Git, data quality, governance, and stakeholder collaboration.
Databricks Metric Views, SQL, Python, Azure Databricks, Git, Azure DevOps, RBAC, CI/CD
1d
Save
Mark Applied
Hide
Data Engineer, Specialist
Hyderabad, Telangana, India
HybridFull Time
Vanguard
Vanguard: Global investment management and financial services provider.
5+ YOE5+ years of relevant experience, including 2+ years designing data pipelines, ETL processes, and database architectures. Requires a bachelor's degree and expertise in Databricks, Python, SQL, cloud platforms, and Terraform.
Databricks, Python, SQL, Terraform, AWS, GCP
1d
Save
Mark Applied
Hide
Specialist , Data Engineering
Hyderabad or Pune
HybridFull Time
Merck & Co.
Merck & Co.NYSE: MRK: Produces prescription medicines, vaccines, and animal health products.
5+ YOEBachelor’s degree and 5+ years in enterprise data integration. Requires Informatica, REST/API, databases, cloud platforms, Python, Shell, SQL, ETL, and orchestration experience.
Informatica PowerCenter, Informatica Intelligent Data Management Cloud Services, CDI, CAI, Mass Ingest, Orchestration, REST, AWS, Google Cloud Platform, Microsoft Azure, AWS Well-Architected Framework, Terraform, GitHub Actions, Artifactory, Python, Shell, Autosys, Airflow, SQL, Informatica Cloud, Informatica Master Data Management, SQL Script