67 pyspark data engineer jobs at 34 companies in Griffin, GA

1w
Save
Mark Applied
Hide
Sr. Associate, Date Engineer - PySpark
Atlanta or New York City or Philadelphia or Dallas
OnsiteFull Time
KPMG
KPMG: Global professional services network providing audit, tax, and advisory.
5+ YOE5+ Mgmt5+ years leading data science/engineering teams, Master’s degree in CS/DS/Statistics/Engineering, hands-on Python, PySpark, SQL, Databricks, Snowflake, experience with cloud platforms, ability to travel, U.S. work authorization required.
Python, PySpark, SQL, Databricks, Snowflake, Linux, AWS, GCP, Azure, R
6d
Save
Mark Applied
Hide
Data Engineer
Atlanta or Washington
$87k-$198k/yr OnsiteFull Time
Booz Allen Hamilton
Booz Allen HamiltonNYSE: BAH: Provides technology and management consulting services to diverse organizations.
3+ YOEBachelor’s degree and 3+ years with PySpark, SQL, and Azure Databricks; data engineering, analytics, technical collaboration, public health data, and Public Trust eligibility required.
PySpark, SQL, Azure Databricks, R, Scala, Java, AWS, Azure, GCP, Spark, Databricks, Hadoop, Hive, EMR, Kafka, Redshift, MySQL, Snowflake, FHIR, EHR, ETL, ELT
5d
Save
Mark Applied
Hide
Data Engineer
Atlanta, Georgia, United States
$87k-$198k/yr HybridFull Time
Booz Allen Hamilton
Booz Allen HamiltonNYSE: BAH: Consulting and technology services for government and commercial clients
3+ YOEBachelor's degree and 3+ years of PySpark, SQL, and Azure Databricks experience; requires data engineering, analytics, cross-functional collaboration, and ability to obtain Public Trust eligibility.
PySpark, SQL, Azure Databricks, R, Scala, Java, AWS, Azure, GCP, Spark, Databricks, Hadoop, Hive, EMR, Kafka, Redshift, MySQL, Snowflake, FHIR, EHR, ETL, ELT
5d
Save
Mark Applied
Hide
Data Engineer - Databricks
Chicago or Indianapolis or Fort Wayne or Columbus or Lansing or Denver or Dallas or Atlanta
RemoteFull Time
Resultant
Resultant: Provides data analytics and technology consulting services to organizations.
2+ YOEBachelor's degree or equivalent, 2+ years of production data engineering experience, Databricks, PySpark, Spark SQL, Delta Lake, SQL, cloud data services, data modeling, and strong client communication skills.
Databricks, PySpark, Spark SQL, Delta Lake, Delta Live Tables (DLT), Auto Loader, Structured Streaming, Databricks Workflows, Airflow, Azure Data Factory, Unity Catalog, MLflow, Photon, Git, Azure DevOps, GitHub Actions, GitLab, Terraform, dbt, Kafka, Event Hubs, Power BI, Tableau, Docker, Kubernetes, SQL Server, Postgres, Oracle, Snowflake, Azure, AWS, GCP
5d
Save
Mark Applied
Hide
Data Engineer Lead
San Antonio or Beavercreek or Houston or Charlotte or Huntsville or Atlanta
$92k-$153k/yr OnsiteFull Time
Guidehouse
Guidehouse: Provides management and technology consulting services to diverse organizations.
7+ YOEBachelor's degree in a technical field, 7+ years of data or software engineering experience, technical leadership, Python, SQL, Spark/PySpark, cloud platforms, data pipelines, ETL/ELT, CI/CD, Docker, and Kubernetes.
Python, SQL, Spark, PySpark, AWS, Azure, Databricks, Snowflake, Docker, Kubernetes, Terraform, CloudFormation, Bicep, ARM, Kafka, Event Hub, Kinesis, Splunk, Datadog, CloudWatch, Kibana, Elasticsearch
6d
Save
Mark Applied
Hide
Online Senior Data Engineer
Atlanta, Georgia, United States
OnsiteFull Time
The Home Depot
The Home DepotNYSE: HD: Retailer of home improvement products, building materials, and tools.
4+ YOERequires 4+ years of relevant experience, bachelor's degree or equivalent, data engineering, predictive modeling, ETL pipelines, Python, SQL, PySpark, AirFlow, DataProc, and production pipeline support.
Python, JavaScript, React, Nucleus, Retina, KPI Shield, Alert Goose, Google BigQuery, PySpark, AirFlow, DataProc, SQL
2mo
Save
Mark Applied
Hide
Lead Data Engineer
Atlanta, Georgia, United States
HybridFull Time
Honeywell
HoneywellNASDAQ: HON: Manufactures aerospace products, building technologies, and industrial control systems.
8+ YOE8+ years data engineering; 2+ years in lead role; medallion lakehouse; Spark/PySpark on Azure Databricks; real-time IoT pipelines; Kafka/Event Hub; cloud data lakes/warehouses; GenAI pipelines; MLOps; CI/CD with GitHub Actions.
Azure Databricks, PySpark, Apache Spark Streaming, Apache Kafka, Azure Event Hub, LangChain, LangGraph, GitHub Actions, CI/CD, Docker, Kubernetes
1mo
Save
Mark Applied
Hide
Data Engineer
Atlanta or Baltimore or Boston or Charlotte or Columbus or Detroit or Hartford or Miami or New York or Philadelphia or Raleigh or Tampa or Washington
$91k-$122k/yr OnsiteFull Time
Slalom
Slalom: Provides business and technology consulting and software engineering services.
4+ YOE4+ years building production data engineering solutions; experience with cloud data platforms, streaming, orchestration, Python/Java; strong communication and critical thinking.
Amazon Redshift, Snowflake, Google BigQuery, Azure Synapse Analytics, Amazon EMR, Apache Spark, Presto, Databricks, Amazon Kinesis, Apache Kafka, Apache Airflow, dbt, Dagster, Azure Data Factory, DynamoDB, Cosmos DB, MongoDB, Kubernetes, Amazon ECS, Amazon SageMaker, Azure ML Studio, Python, Java
1mo
Save
Mark Applied
Hide
Data and Automation Engineer - Python
Atlanta or United States
RemoteFull Time
Plum
Plum: Provides AI-driven software and lending solutions for financial institutions.
3+ YOE3–7+ years in data engineering or related field, strong Python, Databricks, PySpark, SQL, Delta Lake, API integrations, Git; undergraduate degree in CS/Engineering/Physics.
Python, Databricks, PySpark, SQL, Delta Lake, REST APIs, JSON, Git, Cursor, Claude Code, GitHub Copilot, ChatGPT, Apollo, Clay, LinkedIn Sales Navigator, ZoomInfo, Salesforce, HubSpot, LLM APIs
1mo
Save
Mark Applied
Hide
AI Data Engineer
Palo Alto or Atlanta or Washington
$160k-$200k/yr HybridFull Time
Quantifind
Quantifind: AI-driven platform for financial risk and fraud detection.
4+ YOEUS citizen with 4+ years' experience, technical education, strong Python and AWS skills, experience with large-scale ETL, Spark/PySpark, PostgreSQL/RDS, Neo4j, and hands-on AI coding tools; must be able to meet in person.
Claude Code, GPT/Codex, Cursor, Windsurf, PostgreSQL, RDS, Python, Spark/PySpark, Scala, AWS, Neo4j
3w
Save
Mark Applied
Hide
Online Sr Data Engineer (SEO) - Onsite
Atlanta, Georgia, United States
OnsiteFull Time
The Home Depot
The Home DepotNYSE: HD: Retailer of home improvement products and construction materials.
4+ YOERequires 4+ years of relevant experience, bachelor's degree or equivalent, Python, SQL, ETL pipelines, predictive modeling, data analysis, CI/CD, PySpark, AirFlow, DataProc, BigQuery, and strong communication skills.
Python, JavaScript, React, Nucleus, Retina, KPI Shield, Alert Goose, Google BigQuery, PySpark, AirFlow, DataProc, SQL, CI/CD, ETL, API
2d
Save
Mark Applied
Hide
Microsoft Fabric Data Engineer
Athens or Atlanta
OnsiteFull Time
Landmark Properties
Landmark Properties: Develops, builds, and manages student and multifamily residential communities.
3+ YOEBachelor's degree in a relevant technical field and 3+ years in data engineering, business intelligence, analytics development, or related technical work; Microsoft Fabric, Power BI, SQL, data pipelines, and modeling experience.
Microsoft Fabric, Power BI, Microsoft Power Platform, SQL, Python, PySpark, Spark, Microsoft Purview, GitHub, CI/CD, DevOps, DAX, Data Factory, Dataflows, Notebooks, Lakehouse, Warehouse
3mo
Save
Mark Applied
Hide
MID-LEVEL DATA ENGINEER-Python, AWS, Spark (Hybrid)
Bloomington or Richardson or Tempe or Dunwoody
HybridFull Time
State Farm
State Farm: Provider of auto, home, life, and health insurance.
2+ YOE2-4 years as a Data Engineer; proficiency in Python, Spark SQL/PySpark, R, Java, Bash; hands-on AWS (ETL tools, Lambda, Step Functions, S3, DynamoDB, Kinesis, Redshift, SageMaker); Spark/Databricks; IaC (OpenTofu); CI/CD with Airflow.
Python, Apache Spark, PySpark, R, Java, Bash, AWS, Glue, EMR Serverless, Lambda, Step Functions, EventBridge, S3, DynamoDB, Kinesis Firehose, Redshift, Iceberg, SageMaker, Databricks, OpenTofu, Terraform, Airflow, SQL, Athena
1mo
Save
Mark Applied
Hide
Senior Data Engineer - Technology
Tallassee or Duluth
OnsiteFull Time
Neptune Technology Group
Neptune Technology Group: Neptune manufactures water measurement and automated data collection systems for utilities.
5+ YOE5+ years data engineering experience; Bachelor’s or Master’s in CS/Engineering; deep AWS Glue (PySpark/Python), S3, Redshift, SQL; experience migrating legacy ETL to cloud-native pipelines; stream processing interest.
AWS Glue, S3, Redshift, Apache Flink, ClickHouse, Kafka, Kinesis, PySpark, Python, SQL, Aurora MySQL, DynamoDB, Apache Druid, MQTT, dbt, Airflow, Step Functions
2w
Save
Mark Applied
Hide
Data Solutions Engineer
Cleveland or Cincinnati or Columbus or Philadelphia or Washington or Orlando or Atlanta or Austin or Dallas or Houston or Los Angeles or Seattle
$120k-$165k/yr OnsiteFull Time
BakerHostetler
BakerHostetler: Provides legal counsel and litigation services for corporate clients.
5+ YOEBachelor's in CS/IT,5+ years building data solutions; expertise with Azure data platform, Microsoft Fabric, Power Platform, data modeling, medallion architecture, MDM, DevOps, and strong communication skills.
Azure Data Factory, Azure Data Lake, Azure Data Lake Storage, Azure Synapse Analytics, Azure SQL Database, PostgreSQL, Azure Blob Storage, Azure Databricks, Apache Spark, Apache Airflow, Azure Logic Apps, Azure DevOps, Power Automate, REST API, Azure OpenAI, Document Intelligence, Delta Lake, Microsoft Fabric, Lakehouse, Power Apps, Power BI, Git
1mo
Save
Mark Applied
Hide
Data Platform Engineer
Hyderabad or Ghent or Atlanta or London
HybridFull Time
Itineris
Itineris: Develops software solutions for energy and water utility companies.
5+ YOE5+ years in data infrastructure/platform engineering. Deep Apache Spark, lakehouse (Iceberg/Delta), storage design, CDC/replication, SQL. Proficient Python, containerized apps (Docker/Kubernetes), observability and IaC/GitOps experience.
Apache Iceberg, Polaris, Apache Spark, Delta, Kafka, Event Hubs, IoT Hub, SQL, Python, Docker, Kubernetes, Argo CD, Helm, Bicep, Cosmos DB, MongoDB, GitOps, Microsoft Dynamics 365
5d
Save
Mark Applied
Hide
Lead Data Engineer 2026- US
Atlanta, Georgia, United States
RemoteFull Time
Aimpoint Digital
Aimpoint Digital: Provides data engineering and analytics consulting services to businesses.
5+ YOEDegree in a quantitative or technical field or equivalent experience; 5+ years in databases, pipelines, data modeling, and production code; DevOps, cloud warehouses, stakeholder management, and small-team leadership experience.
Snowflake, Databricks, dbt, Fivetran, Snowflake Cortex, Databricks Genie, Git, SQL, Python, Spark, Scala, Java, Codex, Claude, Copilot, Snowflake Cortex Code, Databricks Genie Code, Matillion, Informatica, Talend, AWS, Azure, GCP, Docker, Kubernetes, Apache Spark
1mo
Save
Mark Applied
Hide
Databricks Platform & Data Engineer (Plano, TX, US)
Plano or Minneapolis or Chicago or Atlanta
$135k-$247k/yr OnsiteFull Time
NTT DATA
NTT DATA: Global provider of IT and business consulting services.
12+ YOE12+ years experience with 5+ years on Databricks; deep expertise in Apache Spark, Delta Lake, Databricks platform setup, cloud (AWS/Azure/GCP), data governance, and client-facing delivery.
Databricks, Apache Spark, Scala, Python, Delta Lake, Unity Catalog, Dataflow, Genie, Kafka, Structured Streaming, MLOps, AWS, Azure, GCP, PowerPoint
1mo
Save
Mark Applied
Hide
Database Engineer
Conyers, Georgia, United States
OnsiteFull Time
Batchelor & Kimball
Batchelor & KimballNYSE: EME: Mechanical and plumbing construction for complex commercial facilities.
3+ YOE3+ years data engineering experience; Azure/Microsoft Fabric and lakehouse experience; strong SQL, Python, Spark/PySpark, Delta Lake skills; bachelor’s degree or equivalent experience preferred.
Microsoft Fabric, OneLake, Microsoft Data Factory, Microsoft Power BI, Microsoft Azure Data Lake Storage, Microsoft Azure Data Factory, Microsoft Azure SQL, Microsoft Synapse, Microsoft Azure Databricks, Microsoft Key Vault, RBAC, Apache Spark, PySpark, SQL, Python, Delta Lake, CI/CD, version control, infrastructure as code
1w
Save
Mark Applied
Hide
Data Scientist
Atlanta or United States
RemoteFull Time
Sonatype
Sonatype: Secures software supply chains and automates open source governance.
5+ YOERequires 5+ years in applied data science, machine learning, AI engineering, or research; strong Python, ML/GenAI application development, LLM ecosystems, evaluation, Git, and software development practices.
Python, Databricks, LLM APIs, scikit-learn, OpenAI, Anthropic, Claude, Hugging Face, LangGraph, LangChain, Semantic Kernel, Git, MLflow, AWS SageMaker, Azure ML, MCP, Copilot, Claude Code, Codex, PySpark