74 spark engineer jobs at 48 companies in Luling, TX
2w
Save
Mark Applied
Hide
2w
Senior Data Engineer
Austin, Texas, United States
$144k-$187k/yrHybridFull Time
SpyCloud: Prevents account takeover and ransomware using darknet data.
8+ YOE2+ Mgmt8+ years data engineering experience, 2+ years technical leadership, Databricks/Spark, AWS data services, Python, ETL, ML/LLM data preparation, Elasticsearch/OpenSearch and DynamoDB experience.
Austin or San Francisco or New York City or Seattle or Los Angeles or Chicago or Gurugram or Bengaluru
$170k-$230k/yrRemoteFull Time
SentiLink: Provides identity verification and fraud prevention for financial institutions.
5+ YOE5+ years engineering experience; strong Python or Golang skills; building ETL/ELT pipelines at scale with Spark/Hadoop/Kafka; cloud (AWS/Azure/GCP) and database expertise; containerization and IaC experience.
Site Reliability Engineer, Apple Data Platform - AI/ML Platform
Austin, Texas, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience operating large multi-cloud data platforms, incident response, supporting internal engineering teams, and running services such as Spark, Flink, Airflow, Ray and notebook/LLM agent platforms.
Databricks: A unified platform for data analytics and artificial intelligence.
6+ YOE6+ years in data engineering or software engineering, strong Spark and distributed systems experience, cloud expertise (AWS/Azure/GCP), coding in Python/Scala/JavaScript/TypeScript, CI/CD and MLOps familiarity, customer-facing delivery.
University of Texas at Austin: Provides public higher education and conducts academic research.
2+ YOEBachelor's in CS/IS/Engineering/Statistics or related, 2+ years data engineering/ETL experience, proficiency with Hadoop/Spark/Kafka, SQL and NoSQL, cloud (AWS/Azure/GCP) and Python/Java/C++/Scala.
Overhaul: Provides real-time supply chain visibility and risk management software.
Experienced data engineer with mastery of Python and SQL, Spark/Databricks, cloud data lake tools, strong data intuition, automation and CI/CD focus, and ownership of streaming data pipelines.
Databricks, Spark, SparkSQL, Python, SQL, Azure Data Lake, Synapse, Fabric, Power BI
United States or San Francisco or New York City or Washington or Austin
$118k-$179k/yrRemoteFull Time
SamsaraNYSE: IOT: Connected Operations Cloud platform for physical operations and IoT.
8+ YOEBachelor's in CS or equivalent, 8+ years as a software/data engineer, 5+ years building production data pipelines and Spark/PySpark experience, strong Python and SQL, cloud data warehouse and ETL tooling experience.
RokuNASDAQ: ROKU: Operates a TV streaming platform and sells streaming hardware.
8+ YOE8+ years data engineering experience, extensive SQL, Python, big data tech (Hadoop, Hive, Kafka, Spark, Airflow, Presto/Trino), data modeling, and cloud experience (AWS/GCP). BS required, MS preferred.
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
5+ YOE5+ years experience, bachelor’s in CS/Data Science/Engineering or equivalent, strong proficiency in Scala, Python, SQL, Spark, Databricks, experience with cloud ETL/ELT, ML lifecycle (MLflow), and CI/CD for data/ML products.
Quincy or Toronto or Princeton or Boston or Clifton or Austin
$120k-$203k/yrHybridFull Time
State StreetNYSE: STT: Provides investment servicing and management to institutional investors.
7+ YOE7+ years software engineering, 3+ years building AI/ML/LLM solutions; expertise in cloud-native architectures, RAG, AI evaluation/observability, CI/CD, security, and mentoring engineers.
Washington or Arlington or Austin or Atlanta or United States
RemoteFull Time
VetsEZ: Digital transformation and healthcare IT services for government agencies.
4+ YOEBachelor's degree, 4+ years data engineering experience; proficiency with Python, Spark, SQL, ETL, Azure Synapse, Azure Data Factory, Power BI, GitHub, and Jira; experience with federal healthcare/VA programs and ability to obtain Public Trust.
Azure Synapse, Azure Data Factory, Power BI, Python, Spark, SQL, GitHub, Microsoft Word, Microsoft Excel, Microsoft PowerPoint, Microsoft Visio, Jira, CX Insights, Lighthouse APIs, Medallia, Azure
webAI: Secure on-device AI infrastructure for distributed enterprise applications
6+ YOEPh.D. preferred; 6+ years ML experience with LLMs and MoE, strong Python and TensorFlow/PyTorch skills, leadership and publication record, cloud and production ML experience.
Senior Software Engineer-Bigdata & Hadoop Engineer with Development experience
Austin, Texas, United States
$111k-$172k/yrHybridFull Time
VisaNYSE: V: Global payment technology facilitating electronic funds transfers.
2+ YOE2+ years of relevant experience with a Bachelor's or 3+ years with higher degree; hands-on Hadoop/Spark development; Python/Java; front-end and back-end skills; GenAI/LLMs experience.
United States or Canada or Columbus or Austin or San Francisco or New York City
$167k-$231k/yrRemoteFull Time
UpstartNasdaq: UPST: AI-powered lending marketplace for consumer and automotive loans.
5+ YOE5+ years software development with full-stack, distributed systems, and API experience; proficiency in Ruby on Rails, Kotlin, PostgreSQL, React/Next.js, Python; cloud (AWS/GCP/Azure); microservices and real-time data pipelines (Kafka/Spark).
Staff Software Engineer, Data Engineering – Identity & AI Agent Governance
Austin or Denver or United States
$132k-$170k/yrOnsiteFull Time
Ping Identity: Provides identity and access management software for enterprise security.
8+ YOERequires a bachelor's degree or equivalent experience and 8+ years in data engineering software roles, with scalable pipelines, Spark, Beam or Flink, BigQuery, SQL, NoSQL, and Elasticsearch.
Amazon Bedrock, Google Vertex AI, Microsoft Copilot, Azure AI Foundry, Apache Spark, Apache Beam, Apache Flink, BigQuery, SQL, Elasticsearch, LangGraph, LlamaIndex, AWS, Azure, Google Cloud, Docker, Kubernetes, CI/CD, MCP
Unified: A social network platform for community organizing and activism.
Expertise in Python and SQL, experience with high-volume data pipelines, PostgreSQL and data warehouse systems, workflow orchestration (Airflow/Kafka), ML Ops and containerized deployments; strong engineering and collaboration skills.