Comcast
Posted 1w ago

Platform Data Engineer - (DataBricks, PySpark, AWS)

Comcast
West Chester, Pennsylvania, United States
OnsiteFull Time
Responsibilities
  • developing pipelines
  • tuning performance
  • mentoring engineers
Requirements
  • Bachelor's degree or equivalent experience
  • 5+ years in data engineering, and hands-on AWS
  • PySpark, and Databricks experience. Requires Python
  • ETL/ELT
  • Distributed systems
  • Airflow
  • Kubernetes, and production support expertise
Technical tools mentioned
Amazon Web Services (AWS)PySparkDatabricksKafkaKubernetesApache AirflowManaged Workflows for Apache Airflow (MWAA)PythonSnowflakeAmazon EKS (Elastic Kubernetes Service)

Job description

Make your mark at Comcast -- a Fortune 30 global media and technology company. From the connectivity and platforms we provide, to the content and experiences we create, we reach hundreds of millions of customers, viewers, and guests worldwide. Become part of our award-winning technology team that turns big ideas into cutting-edge products, platforms, and solutions that our customers love. We create space to innovate, and we recognize, reward, and invest in your ideas, while ensuring you can proudly bring your authentic self to the workplace. Join us. You’ll do the best work of your career right here at Comcast. (In most cases, Comcast prefers to have employees on-site collaborating unless the team has been designated as virtual due to the nature of their work. If a position is listed with both office locations and virtual offerings, Comcast may be willing to consider candidates who live greater than 100 miles from the office for the remote option.)

Job Summary

We are seeking a Data Engineer(Engineer 3) to join our Data Product Engineering Team team responsible for managing and evolving the enterprise Data Lake that supports critical datasets across the GTO organization. This team owns large-scale workforce, billing, and interaction datasets and is focused on building scalable, reliable, and high-performance data solutions that enable analytics, reporting, and business decision-making.

The ideal candidate will have strong experience building and optimizing distributed data pipelines in cloud environments, working with high-volume datasets, and partnering with cross-functional teams to deliver impactful data products. This role offers the opportunity to work with environments processing over 50TB of interaction data, leveraging modern technologies including AWS, PySpark, Databricks, Kafka, Kubernetes, and Airflow.

Job Description

Key Responsibilities

  • Design, develop, maintain, and optimize scalable data pipelines supporting workforce, billing, interaction, and other enterprise datasets.
  • Build and enhance cloud-native data solutions using AWS, Databricks, and PySpark.
  • Develop and support batch and streaming data processing frameworks, integrating source systems and interfaces through modern data architectures.
  • Leverage technologies such as Kafka and Databricks streaming solutions to ingest and process high-volume data in near real-time.
  • Drive data pipeline performance tuning, automation initiatives, and operational improvements across the platform.
  • Provide production support, troubleshooting, and root-cause analysis for critical data workflows.
  • Work with large-scale distributed systems and high-concurrency environments processing tens of terabytes of data.
  • Utilize MWAA (Managed Workflows for Apache Airflow) to orchestrate and manage data workflows.
  • Collaborate closely with Product, Data Governance, Analytics, and Engineering teams across both onshore and offshore delivery models.
  • Support data warehousing initiatives and help establish best practices for data quality, scalability, and reliability.
  • Mentor junior engineers, provide technical guidance, and contribute to the growth and development of Engineering I team members.
  • Participate in architectural discussions and contribute to the long-term evolution of the enterprise data platform.

Required Qualifications

  • Bachelor's degree in Computer Science, Engineering, Information Systems, or a related field, or equivalent practical experience.
  • 5+ years of experience in Data Engineering, Data Platform Engineering, or related disciplines.
  • Strong hands-on experience with:
    • AWS
    • PySpark
    • Databricks
  • Experience building and maintaining large-scale ETL/ELT pipelines.
  • Strong understanding of distributed systems and large-volume data processing.
  • Experience with data warehousing concepts and modern data architectures.
  • Experience orchestrating workflows using Apache Airflow/MWAA.
  • Knowledge of Kubernetes fundamentals, including pod lifecycle, job orchestration, and workload configuration.
  • Proficiency in Python development within data engineering environments.
  • Experience supporting production data platforms and driving operational excellence.
  • Strong communication and collaboration skills with the ability to work effectively across multiple teams.

Preferred Qualifications

  • Experience with Snowflake.
  • Experience with Kafka and streaming data architectures.
  • Experience with Amazon EKS (Elastic Kubernetes Service).
  • Background working with large-scale interaction, advertising, marketing, or customer engagement datasets.
  • Experience implementing data platform automation, observability, and monitoring solutions.
  • Prior experience mentoring junior engineers and leading technical initiatives.

Disclaimer: This information has been designed to indicate the general nature and level of work performed by employees in this role. It is not designed to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and qualifications.

Skills

Amazon Web Services (AWS), Apache Airflow, Communication, Databricks Platform, PySpark

We believe that benefits should connect you to the support you need when it matters most, and should help you care for those who matter most. That's why we provide an array of options, expert guidance and always-on tools that are personalized to meet the needs of your reality—to help support you physically, financially and emotionally through the big milestones and in your everyday life.


Please visit the benefits summary on our careers site for more details.

Education

Bachelor's Degree

While possessing the stated degree is preferred, Comcast also may consider applicants who hold some combination of coursework and experience, or who have extensive related professional experience.

Certifications (if applicable)

Relevant Work Experience

5-7 Years

Comcast is an equal opportunity workplace. We will consider all qualified applicants for employment without regard to race, color, religion, age, sex, sexual orientation, gender identity, national origin, disability, veteran status, genetic information, or any other basis protected by applicable law.

About Comcast

A global media and technology.

Similar jobs

Data Engineer roles near West Chester, Pennsylvania
1d
Save
Mark Applied
Hide
Sr. Data Engineer - Flink
Reading, Pennsylvania, United States
OnsiteFull Time
Hitachi Rail
Hitachi Rail: Global provider of integrated railway technology and services.
10+ YOERequires 10+ years of software, data engineering, or distributed-systems experience, including 7+ years with Apache Flink, AWS Managed Service for Apache Flink, Kafka, Java, Python/PyFlink, SQL, and production streaming platforms.
Apache Flink, AWS Managed Service for Apache Flink, Kafka, Java, Python, PyFlink, Flink SQL, AWS, Amazon CloudWatch, Amazon S3, IAM, CI/CD
3d
Save
Mark Applied
Hide
Data Engineer, Alternative Data | Delta One Trading | Experienced Hire
Bala Cynwyd or Philadelphia
OnsiteFull Time
Susquehanna International Group
Susquehanna International Group: Global quantitative trading and technology firm.
5+ YOERequires 5+ years building Python data applications and pipelines, large-scale vendor data ingestion, strong SQL, analytical tooling, data modeling, production pipeline operations, and collaboration with researchers.
Python, S3, SFTP, Snowflake, SQL, Parquet, Arrow, DuckDB, NumPy, Pandas, Polars, LLMs, AWS S3, C++
3d
Save
Mark Applied
Hide
Data Engineer, Alternative Data | Delta One Trading | Experienced Hire
Bala Cynwyd, Pennsylvania, United States
OnsiteFull Time
Susquehanna International Group
Susquehanna International Group: Global quantitative trading and technology firm.
5+ YOE5+ years building Python data applications and large-scale pipelines; strong SQL, data modeling, production operations, analytical tooling, and third-party data ingestion experience. Quantitative finance experience is beneficial.
Python, Amazon S3 (S3), SFTP, Snowflake, SQL, Parquet, Arrow, DuckDB, NumPy, Pandas, Polars, LLMs, C++
4d
Save
Mark Applied
Hide
Data Engineer, Alternative Data | Delta One Trading | Experienced Hire
Bala Cynwyd, Pennsylvania, United States
OnsiteFull Time
Susquehanna International Group
Susquehanna International Group: Global quantitative trading and technology firm.
5+ YOERequires 5+ years building Python data applications and large-scale pipelines, strong SQL, analytical tooling, data modeling, production pipeline operations, and collaboration with researchers. Advanced degree and C++ are preferred.
Python, Amazon S3, SFTP, Snowflake, SQL, Parquet, Apache Arrow, DuckDB, NumPy, Pandas, Polars, LLMs, C++, AWS S3
4d
Save
Mark Applied
Hide
Lead Data Engineer - Tax Product Development
New York City or Washington or Houston or Chicago or Grand Rapids or Richmond or Raleigh or Boston or Denver or Atlanta or Orlando or Dallas or Milwaukee or Miami or Philadelphia or Charlotte or Seattle or United States
$140k-$210k/yr HybridFull Time
BDO USA
BDO USA: Professional services firm providing assurance, tax, and advisory services.
6+ YOEBachelor's degree required and 6+ years of data engineering experience. Requires programming experience and expertise in databases, SQL, ETL/ELT pipelines, data modeling, and distributed data environments.
SQL Server, PostgreSQL, Oracle, CosmosDB, MongoDB, Cassandra, DynamoDB, SQL, Apache Spark, Databricks, Airflow, Talend, Informatica, Python, Java, C#, C++, Scala, Microsoft Office Suite, Microsoft Azure DevOps, GitHub, Microsoft SQL Server, Azure SQL DB
4d
Save
Mark Applied
Hide
Data Engineer
New Holland or Pennsylvania or Ohio or South Dakota
HybridFull Time
Goodville Mutual Casualty Company
Goodville Mutual Casualty Company: Mennonite-founded mutual property and casualty insurer serving households, farms, churches, and small businesses through independent agents.
5+ YOEBachelor's degree in computer science, information systems, or related field; 60 months' data engineering and ETL development experience in property and casualty insurance; expertise with listed data tools and technologies.
SSIS, PySpark, PySQL, SQL, ADF, Azure Service Bus, Azure Functions, SQL Server, DB2, Azure SQL, Fabric Lakehouse, REST, SOAP, XML, JSON, Python, Shell
1w
Save
Mark Applied
Hide
Data Engineer Job
Lancaster, Pennsylvania, United States
$125k-$145k/yr OnsiteFull Time
Armstrong World Industries
Armstrong World IndustriesNYSE: AWI: Public American manufacturer of ceiling, wall, and exterior metal systems serving architects, designers, contractors, and building owners.
5+ YOEBachelor's degree in a related field and 5+ years as a Data Engineer or similar role. Requires SQL, programming experience, modern data platforms, cloud services, and strong analytical and collaboration skills.
SQL, Python, Scala, Java, Databricks, Snowflake, AWS, Microsoft Azure, Google Cloud Platform (GCP), Git, SAP BW, SAP Datasphere, SAP Business Data Cloud
1w
Save
Mark Applied
Hide
Digital Platforms, Senior Data Engineer
Philadelphia or Rockford
OnsiteFull Time
PCI Pharma Services
PCI Pharma Services: Global contract development and manufacturing organization for biopharma.
5+ YOEHands-on expertise in Databricks, Spark, Delta Lake, PySpark, data modeling, ETL/ELT, AWS cloud services, CI/CD, data governance, testing support, and secure production operations.
Databricks, Apache Spark (Spark), Delta Lake, PySpark, Databricks SQL, AWS Glue, Amazon S3, AWS Lambda, CI/CD, DevOps, Data Catalog, ETL/ELT