37 pyspark data engineer jobs at 29 companies in Anderson Mill, TX
2w
Save
Mark Applied
Hide
2w
Data Engineer
Munich or Porto or Austin
HybridFull Time
CELUS: AI-powered platform for automating complex electronics design processes.
5+ YOE5+ years data engineering experience building ETL/ingestion pipelines with Databricks/PySpark, SQL, dbt and Snowplow; Bachelor's in CS or related; strong communication and mentoring skills.
United States or San Francisco or New York City or Washington or Austin
$118k-$179k/yrRemoteFull Time
SamsaraNYSE: IOT: Connected Operations Cloud platform for physical operations and IoT.
8+ YOEBachelor's in CS or equivalent, 8+ years as a software/data engineer, 5+ years building production data pipelines and Spark/PySpark experience, strong Python and SQL, cloud data warehouse and ETL tooling experience.
United States or Arlington or Tysons or Washington or New York City or Chicago or Austin or Atlanta or Boston or Boulder
$113k-$188k/yrRemoteFull Time
Guidehouse: Provides management and technology consulting services to diverse organizations.
3+ YOEBachelor's degree,3+ years data engineering experience,proficiency with Python/PySpark/SQL,Databricks/Spark/Delta Lake experience,knowledge of data pipelines,governance,and cloud platforms.
Data Engineer, Data Platform Management, Grocery Tech Foundations
Austin, Texas, United States
$132k-$179k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
5+ YOE5+ years data engineering with distributed systems, data modeling, ETL/ELT, SQL, AWS (Redshift,S3,Glue,EMR,Kinesis,FireHose,Lambda), BI and ETL tools, Python/PySpark/Java; ability to modernize data infrastructure and support reporting apps.
Delan Associates: Provides engineering and professional services to government agencies.
12+ YOE12+ years experience, Databricks Certified Data Engineer, Azure Databricks, Lakehouse architecture, ETL/ELT pipelines, Delta Live Tables (DLT), PySpark; local Austin candidates for in-person interview.
Azure Databricks, Delta Live Tables (DLT), PySpark
Enverus: Software and analytics for the global energy industry.
Hands-on experience building scalable ETL pipelines with Databricks and PySpark, cloud (AWS S3/IAM/Lambda), relational databases, CI/CD with GitHub, and integrating LLM/GenAI into data workflows.
Plum: Provides AI-driven software and lending solutions for financial institutions.
3+ YOE3+ years data engineering experience, strong Python, Databricks/PySpark/SQL/Delta Lake, API integrations, data quality and entity resolution, Git, and relevant undergraduate degree.
San Francisco or Los Angeles or Denver or Austin or Chicago or New York City or Seattle or Toronto or Santa Barbara or San Diego
$127k-$190k/yrRemoteFull Time
Invoca: AI platform for conversation intelligence and revenue execution.
5+ YOE5+ years in data or software engineering; advanced Python, SQL, Databricks, Spark, and data modeling; experience with orchestration, streaming, cloud infrastructure, data quality, ML datasets, and privacy; bachelor's degree or equivalent.
University of Texas at Austin: Provides public higher education and conducts academic research.
4+ YOEBachelor's degree in CS/IS/Data Science or related; 4+ years data engineering; Python, SQL, data pipelines; data governance; NoSQL; healthcare data experience preferred.
Python, PySpark, SQL, NoSQL, Apache Spark, Airflow, Git, Microsoft Fabric, Azure Data Factory, Azure Synapse, Google BigQuery, AWS Redshift
VisaNYSE: V: Global payment technology facilitating electronic funds transfers.
2+ YOE2+ years experience with a bachelor's degree or 3+ years with advanced degree; strong data engineering, distributed systems, SQL, Spark, Python/Scala; experience in UNIX/Linux; ability to work hybrid and design scalable pipelines.
Cleveland or Cincinnati or Columbus or Philadelphia or Washington or Orlando or Atlanta or Austin or Dallas or Houston or Los Angeles or Seattle
$120k-$165k/yrOnsiteFull Time
BakerHostetler: Provides legal counsel and litigation services for corporate clients.
5+ YOEBachelor's in CS/IT,5+ years building data solutions; expertise with Azure data platform, Microsoft Fabric, Power Platform, data modeling, medallion architecture, MDM, DevOps, and strong communication skills.
Azure Data Factory, Azure Data Lake, Azure Data Lake Storage, Azure Synapse Analytics, Azure SQL Database, PostgreSQL, Azure Blob Storage, Azure Databricks, Apache Spark, Apache Airflow, Azure Logic Apps, Azure DevOps, Power Automate, REST API, Azure OpenAI, Document Intelligence, Delta Lake, Microsoft Fabric, Lakehouse, Power Apps, Power BI, Git
Senior Data Engineer – Agentic AI, Automation, and Data Platforms
Warren or Austin
$139k-$174k/yrHybridFull Time
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
5+ YOEBachelor's degree or equivalent experience, 5+ years in data engineering, and experience with Python or Scala, SQL, Spark, Databricks, cloud platforms, batch and streaming pipelines, and AI technologies.
Cursor, large language models, Vector Search, Databricks agents, RAG, Azure Databricks, Python, Scala, SQL, Apache Spark, Azure, AWS, GCP, Databricks, Delta Lake, Claude, GitHub Copilot, Genie, Glean, CI/CD, APIs
United States or New York City or Austin or Miami or Mountain View
$215k/yrRemoteFull Time
YipitData: Providing alternative data and research for institutional investors.
8+ YOE3+ Mgmt8+ years data engineering experience, 3+ years people management, hands-on with SQL, PySpark, Databricks, Airflow, strong data modeling and observability skills, experience scaling production data systems.
Claude Code, Codex, Cursor, Databricks, Airflow, SQL, PySpark
Bachelor’s or master’s degree in a related field; data engineering fundamentals, analytics, and collaboration skills. Cloud, SQL, Python, ETL/ELT, orchestration, Git, and visualization experience preferred.
AWS, Microsoft Azure, Google Cloud Platform (GCP), Snowflake, Databricks, Redshift, BigQuery, Apache Spark, Apache Kafka, SQL, Python, Airflow, dbt, Git, Power BI, Tableau, Looker
Acrisure: Provides AI-powered insurance, financial, and business risk solutions.
2+ YOE2+ years engineering experience with data warehouses/lakes; strong SQL; proficient in Scala, Python, or Java; experience with Databricks/BigQuery/Palantir; familiarity with Airflow/Dagster/Fivetran and cloud (GCP, Azure); DevOps and API experience.
Self Financial: Platform for building credit history and personal savings.
7+ YOE4+ MgmtRequires 7+ years of hands-on data engineering, 4+ years managing engineers, modern warehouse and transformation stack experience, architecture expertise, streaming knowledge, and strong communication.
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years software engineering experience delivering data platforms, expertise with distributed processing, data modeling, cloud data services, and applied programming in Python/Java/SQL.
R1 RCM: Revenue cycle management solutions for healthcare providers.
5+ YOEAt least 5 years of software engineering experience and 2 years with high-throughput data pipelines; strong Scala and SQL skills plus experience with production data systems, microservices, ETL, infrastructure, observability, and on-call operations.