28 pyspark developer jobs at 20 companies in Spring, TX
1mo
Save
Mark Applied
Hide
1mo
Palantir Forward Deployed Engineer
Chicago or Milwaukee or Dallas or Columbus or Kirkland or Cincinnati or New York or Cleveland or Oklahoma City or Austin or Albany or St. Petersburg or Hartford or Pittsburgh or St. Louis or Miami or Sacramento or Raleigh or Minneapolis or Mountain View or Scottsdale or San Francisco or Morristown or Denver or Boston or Philadelphia or Des Moines or Overland Park or Los Angeles or Charlotte or Walnut Creek or Carmel or Seattle or Houston or Arlington or Atlanta or Redmond or Bentonville or Beaverton or Nashville or Detroit or San Diego
$54k-$235k/yrHybridFull Time
AccentureNYSE: ACN: Global provider of management consulting and technology services.
3+ YOE3+ years data engineering; 1+ year Palantir Foundry experience; 1+ year prompt engineering/AIP; 3+ years cloud (AWS/Azure/GCP); 3+ years Python/PySpark/Java; Bachelor's in CS/Engineering or equivalent; strong communication.
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
5+ YOEBachelor's in CS/Engineering or equivalent and 5 years as a software developer; experience with multi-region cloud object-storage, Kubernetes stateful clusters, Sysdig, Activity Tracker, advanced Go, Spark with Scala, and Fluentd with Ruby required.
Boston or Atlanta or Charlotte or Chicago or Dallas or Houston or Los Angeles or New York
$135k-$195k/yrRemoteFull Time
Paul Hastings: International law firm providing legal and regulatory consulting services.
7+ YOE7+ years data engineering or BI development; 3+ years Azure cloud; 2+ years Azure DevOps; Bachelor's in CS/Engineering; SQL/ETL; PySpark; familiarity with Power BI; professional services experience preferred.
Azure Data Factory, Azure Data Lake Storage, Azure Databricks, Azure DevOps, SQL, PySpark, Git, Power BI, Microsoft Purview
Tailored Brands: Omnichannel retailer of menswear, formalwear, and tailored clothing.
8+ YOEBachelor's/Master's in CS or related,8+ years data engineering experience,proficiency with cloud data warehouses,SQL,Python/PySpark,Bash,ETL tools,and data governance;strong communication and troubleshooting skills.
Rysun Labs: Provides AI, data, and digital innovation services for enterprises.
4+ YOE4+ years data engineering experience; strong SQL, ETL/ELT, data warehousing, data lake, and cloud (AWS/Azure/GCP) skills; Python/PySpark/Scala and Big Data (Spark, Databricks, Hadoop, Kafka); CI/CD and Git experience.
SQL, Python, PySpark, Scala, Spark, Databricks, Hadoop, Kafka, AWS, Microsoft Azure, GCP, CI/CD, Git, Snowflake, Redshift, Microsoft Synapse, BigQuery
Fracht: Global freight forwarder providing comprehensive international logistics and transport.
5+ YOEBachelor's in a technical field required; 5+ years data engineering experience (or Master's + 3 years); expertise with Microsoft Fabric, Spark/PySpark, Power BI, dimensional modeling, SQL/T-SQL, dbt, and CI/CD; strong documentation and stakeholder skills.
Microsoft Fabric, Fabric Data Pipelines, OneLake, Lakehouse, Data Warehouse, Notebooks, Spark, PySpark, SQL, T-SQL, Power BI, DAX, Tabular Editor, ALM Toolkit, DAX Studio, dbt, Power Query (M), Microsoft Excel, Dataflows Gen2, Azure Data Factory, Git, Microsoft Entra ID, CargoWise
Visual Comfort & Co.: Designer and manufacturer of premium decorative and architectural lighting products.
3+ YOE3+ years data engineering experience; bachelor's in CS/IS or equivalent; proficiency in SQL/T-SQL, Python, PySpark; experience with Azure Data Factory, Azure Synapse, ADLS, Microsoft Fabric, Power BI, D365/JDE, and Git; strong data modelling and governance skills.
SQL, T-SQL, Python, PySpark, Spark, Azure Data Factory, Azure Synapse Analytics, Azure Data Lake Storage, Microsoft Fabric, Delta Lake, Microsoft Dynamics 365 (D365), JD Edwards (JDE), Power BI, Git
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
7+ YOE7+ years software engineering experience, enterprise system design, production AI/ML and LLM systems, strong Python/Java skills, cloud-native experience, data engineering and observability expertise.
CactusNYSE: WHD: Designs and manufactures wellhead and pressure control equipment.
6+ YOE6+ years data engineering/architecture experience; hands-on Azure Databricks/Spark, Python/PySpark, advanced SQL; experience ingesting ERP/CRM data, CI/CD with GitHub, data modeling, governance, and lakehouse patterns.
Azure Databricks, Apache Spark, Python, PySpark, SQL, Azure Data Lake Storage (ADLS), GitHub, Delta Lake, Power BI, SSRS, Tableau, C#, .NET, JavaScript
Cactus WellheadNYSE: WHD: Manufactures and rents wellheads and pressure control equipment.
6+ YOE6+ years data engineering/architecture experience building Databricks/Spark pipelines, strong Python/PySpark and SQL skills, Databricks/Unity Catalog governance, CI/CD with GitHub, ADLS integration, and data modeling for analytics/AI.
Azure Databricks, Apache Spark, Python, PySpark, SQL, Azure Data Lake Storage (ADLS), GitHub, Delta Lake, Power BI, SSRS, Tableau, C#, .NET, JavaScript
Imperative Chemical Partners: Provider of specialty oilfield chemicals and chemical management services.
5+ YOEMinimum 5 years BI experience, bachelor’s degree or equivalent, expert in Microsoft Fabric and Power BI, strong SQL and PySpark skills, Databricks/Delta Lake/Apache Spark experience, semantic modeling and data governance.
Power BI, Microsoft Fabric, Databricks, Snowflake, Azure SQL, Delta Lake, Apache Spark, PySpark, DAX, SQL
Microsoft Fabric Engineer | Camden Corporate Office
Houston, Texas, United States
OnsiteFull Time
Camden Property TrustNYSE: CPT: Owns and manages luxury apartment communities across the US.
5+ YOEBachelor's in CS/related or equivalent experience, 5+ years data engineering with 1+ year Microsoft Fabric experience, strong SQL and Python (PySpark), Delta Lake/Parquet, Power BI semantic modeling, Git/Azure DevOps, CI/CD, backup/DR, Purview and platform governance.
Microsoft Fabric, Fabric Data Factory, Dataflows Gen2, Spark notebooks, OneLake, Power BI, Copilot Studio, Azure, Purview, Git, Azure DevOps, SQL Server, Azure OpenAI, Delta Lake, Parquet, Python, PySpark, Azure Synapse, Databricks, RealPage
Persona AI: Developing rugged humanoid robots for industrial labor automation.
M.S. or Ph.D. in a related field; deep Python and PyTorch expertise; experience with force/time-series and terabyte-scale video processing; strong 3D geometry and data augmentation skills.
Python, PyTorch, OpenCV, FFmpeg, Decord, URDF, Ray, Apache Spark, Omniverse, MuJoCo, MANO, SMPL, Open X-Embodiment, DROID, AgiBot World, EgoDex
Conduit Power: Builds and operates on-site natural gas and battery power solutions.
3+ YOEBachelor's degree in a technical field, 3+ years data engineering experience, strong Python and SQL skills, experience with cloud data platforms, ETL/ELT, streaming and time-series data, and collaboration with cross-functional teams.
Atlanta or Austin or Boston or Charlotte or Chicago or Dallas or Denver or Detroit or Hartford or Houston or Los Angeles or Miami or Nashville or New Brunswick or New York or Raleigh or Seattle or Tampa or Washington
$90k-$147k/yrOnsiteFull Time
Slalom: Provides business and technology consulting and software engineering services.
5+ YOE5+ years data engineering experience, 2+ years building Databricks pipelines and platform features; proficiency with Spark, SQL, Python/Java, EMR/Redshift/Snowflake/BigQuery, orchestration (Airflow/dbt), and data quality tools.
Houston or Golden or Reno or Oakland or Salt Lake City
HybridFull Time
Fervo Energy: Generates clean electricity using advanced geothermal drilling technology.
2+ YOE2+ years building and operating production data pipelines with Apache Spark, Python/SQL, Databricks, Azure, and Snowflake; experience with streaming, data modeling, data quality, governance, and CI/CD.
Databricks, Delta Lake, Delta Live Tables, Unity Catalog, Databricks Workflows, Databricks SQL, Apache Spark, PySpark, Spark SQL, Structured Streaming, Kafka, Microsoft Event Hubs, Microsoft Azure Data Factory, Microsoft Azure Data Lake Storage (ADLS), Snowflake, Microsoft Power BI, Spotfire, Python, SQL, Git, Azure DevOps, GitHub Actions, dbt, Terraform, Docker, Airflow, Microsoft Entra ID, Microsoft Key Vault, MQTT, OPC UA, SparkplugB, Canary, Ignition, Snowflake Semantic Views, Databricks Metric Views
HPNYSE: HPQ: Manufacturer of personal computers, printers, and imaging devices.
7+ YOE7+ years experience in data engineering with expertise in data collection, cleansing, validation, ETL, data modeling, pipelines, and big data technologies; leadership, mentoring, and strong analytical skills; programming proficiency in SQL/Python.
Amazon Web Services, Apache Hadoop, Apache Kafka, Apache Spark, Java, Microsoft Azure, Python, Scala, SQL
Boardwalk Pipelines: Transporting and storing natural gas and liquids via pipelines.
7+ YOE7+ years cloud data architecture experience (5+ years AWS), proficiency with AWS services and Databricks, strong SQL/Python skills, experience with data governance and mentoring engineers; Bachelor’s required, Master’s preferred.
Glue, Redshift, Athena, Lake Formation, SageMaker, Bedrock, Step Functions, Databricks, Alation, SQL, T-SQL, Python, PySpark, Pandas, DAX, Microsoft Power Query (M), PL/SQL, Microsoft SQL Server, Oracle, PostgreSQL, Microsoft Azure SQL, Teradata, Microsoft Power BI, Git, Microsoft Visual Studio Code, SSMS
OxyChem: Manufacturer of essential chemicals for industry and water treatment.
15+ YOESenior, hands-on experience designing and building Azure/SAP-based enterprise data and analytics platforms with governance, IaC, CI/CD, and data integration expertise.
Microsoft Fabric, OneLake, SAP BDC, SAP Datasphere, CPI, Globalscape MFT, Microsoft BizTalk, Terraform, Bicep, Power BI, Azure Functions, Service Bus, Event Grid, API Management, Container Instances, Git, SQL, Python, PySpark, PI historian, OData