344 data engineer kafka jobs at 176 companies in San Francisco, CA

2mo
Save
Mark Applied
Hide
Data Engineer
Santa Clara, California, United States
$135k-$180k/yr HybridFull Time
PlusAI
PlusAI: AI-based virtual driver software for factory-built autonomous trucks.
1+ YOEMS in CS/Electrical or related, 1+ years software engineering, expertise in Python and SQL, experience with large-scale data processing (Spark, MapReduce, Kafka), AWS (S3, EC2, RDS), and building/maintaining data pipelines.
Python, SQL, microservices, MapReduce, Spark, Kafka, AWS, S3, EC2, RDS, TensorFlow, PyTorch
1mo
Save
Mark Applied
Hide
Senior Data Engineer
San Jose, California, United States
$149k-$361k/yr HybridFull Time
Roku
RokuNASDAQ: ROKU: Operates a TV streaming platform and sells streaming hardware.
8+ YOE8+ years data engineering experience, strong SQL, Python preferred, expertise with big-data tech (Hadoop, Spark, Kafka, Hive, Airflow), data modeling, and cloud platforms (AWS/GCP).
SQL, Python, Hadoop, HDFS, YARN, MapReduce, Hive, Kafka, Spark, Airflow, Presto, Trino, AWS, GCP, Looker, Claude Code, Cursor, MCP servers
1mo
Save
Mark Applied
Hide
Data Engineer - Data Platform
San Jose, California, United States
$156k-$388k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
BS/MS in CS or equivalent; experience with Hadoop, Hive, Spark, Presto, Kafka, ClickHouse, Flink; ETL, data ingestion, schema design and SQL; big data system architecture experience.
Hadoop, M/R, Hive, Spark, Metastore, Presto, Flume, Kafka, ClickHouse, Flink, ETL, SQL
5d
Save
Mark Applied
Hide
Data Engineer - DataOps
Cupertino, California, United States
$130k-$150k/yr HybridContract
Monks
MonksLondon Stock Exchange: SFOR: A global digital marketing and technology services agency
5+ YOERequires 5+ years of data engineering experience, expert SQL, Python, and experience with Airflow, Spark, Trino/Dremio, Apache Iceberg, Kafka, Docker, ETL/ELT, Git, CI/CD, lakehouse architectures, and production pipelines.
SQL, Python, Java, Scala, Airflow, Spark, Trino, Dremio, Apache Iceberg, Kafka, Docker, Git, CI/CD
2mo
Save
Mark Applied
Hide
Senior Data Engineer
San Francisco, California, United States
$138k-$221k/yr OnsiteFull Time
Mastercard
MastercardNYSE: MA: Global payment network facilitating electronic transactions and financial services.
Senior-level Big Data engineering experience with Java or Scala, Apache Spark, Hive/Impala, streaming tools (Kafka, NiFi), relational and NoSQL databases, Linux/Shell, strong CS fundamentals, and experience designing high-scale distributed systems.
Java, Scala, Apache Spark, MySQL, PostgreSQL, Hive, Impala, Oozie, Airflow, NiFi, Kafka, Linux/Unix, Shell, Azure, AWS, GCP
3w
Save
Mark Applied
Hide
Staff Data Engineer
Denver or San Francisco
$190k-$264k/yr HybridFull Time
Checkr
Checkr: AI-powered platform for background checks and identity verification.
10+ YOE10+ years building scalable data platforms; expert PySpark, Python, SQL; experience with Kafka, Spark, Iceberg, data lakes, AWS; strong data modeling and security awareness.
PySpark, Python, SQL, MongoDB, Kafka, Spark, Iceberg, EKS, EMR, Serverless, Glue, Athena, S3, Databricks, Snowflake
3mo
Save
Mark Applied
Hide
Senior Data Engineer - Data Lead
San Francisco, California, United States
$150k-$250k/yr OnsiteFull Time
Foundry Robotics
Foundry Robotics: An AI-native robotics manufacturer focused on deploying advanced assembly and production capability for robotics companies and national-security-critical hardware.
5+ YOE5-7 years data engineering experience; TB–PB-scale pipelines; Spark/Flink/Kafka/Airflow/dbt; Python/SQL; Go/Java/TypeScript a plus; cloud data services; architecture and governance; independent, end-to-end ownership.
Spark, Flink, Kafka, Airflow, dbt, Python, SQL, Go, Java, TypeScript, AWS, GCP, Azure, Kubernetes
2mo
Save
Mark Applied
Hide
Senior Data Engineer
Foster City, California, United States
$123k-$191k/yr HybridFull Time
Visa
VisaNYSE: V: Global payment technology facilitating electronic funds transfers.
2+ YOE2+ years experience with a Bachelor's (or Advanced) degree; strong programming in Java/Python/Go/Scala; experience building scalable data pipelines, Spark/Kafka/Hadoop ecosystem, cloud platforms, GenAI and data modeling.
Java, Python, Go, Scala, Gen AI, GenAI, Apache Spark, Spark, Kafka, Hadoop, Hive, Trino, Presto, Apache Airflow, Hbase, AWS, Azure, DataBricks, HDFS, S3, LLMs, NLP, RAG, Vector DBs, MCP, A2A, J2EE, Microservices, REST APIs, CI/CD, Jira, Jira Align
2w
Save
Mark Applied
Hide
(USA) Staff, Data Engineer
Sunnyvale, California, United States
$143k-$286k/yr OnsiteFull Time
Walmart
WalmartNYSE: WMT: Multinational retail operating discount stores and supermarkets.
3+ YOERequires a bachelor's degree and 4 years' software engineering experience, or 6 years' experience, or a master's degree and 2 years' experience, plus 3 years' data engineering experience. Python, SQL, ETL, cloud, Kafka, REST APIs, and Git required.
Python, SQL, ETL, Apache Kafka, REST APIs, Git, CLI, CI/CD, Web Content Accessibility Guidelines (WCAG) 2.2 AA
2mo
Save
Mark Applied
Hide
Data Science Engineer
San Jose or California or United States
$109k-$201k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
8+ YOESenior data engineer with 8+ years experience, distributed data and cloud technologies (Hadoop, Spark, Databricks, S3, EMR), streaming (Kafka/Kinesis), strong SQL and Python/PySpark/Scala skills, CI/CD and orchestration experience, and strong communication.
Hadoop, Hive, Presto, Spark, Databricks, S3, Azure Blob Storage, Notebooks, AWS EMR, Athena, Glue, Delta, Parquet, ORC, Kafka, Kinesis, SQL, Python, PySpark, Scala, Pandas, NumPy, Koalas, APIs, GitHub, Jenkins, Apache Air Flow, Azkaban, AEP, AJO, CJA, Adobe Experience Platform, Adobe Analytics, Customer Journey Analytics, Adobe Journey Optimizer, Collibra, JIRA, Confluence, Copilot, Claude, LLAMA, Databricks Genie, n8n, RAG, MCPs, GenStudio, Adobe Acrobat Studio, Adobe Express, Adobe Firefly, Creative Cloud
2mo
Save
Mark Applied
Hide
Data Engineer
Cupertino, California, United States
$180k-$235k/yr OnsiteFull Time
Gridmatic
Gridmatic: AI-powered platform for optimizing energy trading and battery storage.
Experience building large-scale production data pipelines, designing storage and schema for timeseries and warehouse data, proficiency with DBT and data processing tools, strong software engineering skills, and startup experience.
DBT, spark, kafka, flink, beam, dataflow, Python, GCP, Kubernetes, Terraform, Flyte, Temporal, React, NextJS, Postgres, BigQuery
3w
Save
Mark Applied
Hide
Staff Data Engineer
Sunnyvale, California, United States
$200k-$260k/yr OnsiteFull Time
RADAR
RADAR: Real-time inventory tracking using RFID and computer vision technology.
8+ YOE8+ years data/analytics engineering experience, strong SQL and Python, experience with batch and streaming pipelines, data modeling, data quality, testing, monitoring, and mentoring engineers.
Airflow, Apache Beam, Python, SQL, Apache Spark, Dagster, dbt, Kafka Streams, Flink, Looker, Tableau, Git, Snowflake, Databricks, BigQuery, Docker
2mo
Save
Mark Applied
Hide
Senior Big Data Engineer
San Francisco or New York City or Seattle or Little Rock
$130k-$197k/yr OnsiteFull Time
LiveRamp
LiveRampNew York Stock Exchange: RAMP: Provides a platform for secure data collaboration and identity.
7+ YOE7+ years in software/data engineering; hands-on with Apache Spark, Airflow or Temporal, Dataproc, Kubernetes, Terraform, streaming (Redpanda/Kafka); on-call experience; bachelor’s in CS/Engineering/Math or equivalent.
Apache Spark, Airflow, Temporal, Dataproc, Redpanda, Kafka, Kubernetes, Terraform, Grafana, Linux, Python, Go, Java, Scala, GIT, Subversion, AWS, GCP, Azure
4w
Save
Mark Applied
Hide
Staff Data Engineer - Data Engineering
Livingston or New York City or Sunnyvale or San Francisco or Bellevue
$207k-$275k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
10+ YOE10+ years in data or software engineering with enterprise data architecture experience; expertise in lakehouse architectures, data modeling, governance, SQL and production programming (Python/Scala/Java/Rust); strong cross-team leadership.
Apache Iceberg, Delta Lake, Apache Hudi, Apache Paimon, Apache Fluss, StarRocks, ClickHouse, Trino, Spark, Flink, Kafka, AutoMQ, Pulsar, SQL, Python, Scala, Java, Rust, Kubernetes
3mo
Save
Mark Applied
Hide
AI Data Engineer
Cupertino or San Francisco Bay Area
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
4+ YOE4+ years of data engineering experience for AI applications; strong SQL and Python; cloud data platforms; REST APIs; CI/CD; data modeling.
dbt, Fivetran, Airflow, Snowflake, Databricks, BigQuery, REST APIs, Python, SQL, Git, CI/CD, Pinecone, Weaviate, Chroma, MongoDB, Neo4j, Kafka
2mo
Save
Mark Applied
Hide
Senior Staff Data Engineer - Data & ML Platform
San Francisco, California, United States
$240k-$360k/yr HybridFull Time
Hinge Health
Hinge HealthNYSE: HNGE: Digital provider of musculoskeletal care and physical therapy
10+ YOE10+ years data engineering experience; platform/infrastructure focus; Kafka, Flink, Spark, Python, SQL; ML platform infrastructure; mentoring; growth-stage experience.
Python, SQL, Spark, dbt, Kafka, Flink, Databricks, AWS
6d
Save
Mark Applied
Hide
Senior Data Engineer
San Francisco or Los Angeles or Denver or Austin or Chicago or New York City or Seattle or Toronto or Santa Barbara or San Diego
$127k-$190k/yr RemoteFull Time
Invoca
Invoca: AI platform for conversation intelligence and revenue execution.
5+ YOE5+ years in data or software engineering; advanced Python, SQL, Databricks, Spark, and data modeling; experience with orchestration, streaming, cloud infrastructure, data quality, ML datasets, and privacy; bachelor's degree or equivalent.
Python, PySpark, Pandas, Polars, PyArrow, SQL, MySQL, PostgreSQL, Databricks, Delta Lake, Unity Catalog, Workflows, Jobs/Compute, Snowflake, BigQuery, Apache Spark, Airflow, Dagster, Kafka, Kinesis, Spark Structured Streaming, AWS, S3, IAM, FastAPI, Databricks Model Serving, SageMaker endpoints
3w
Save
Mark Applied
Hide
GCP Data Engineer
San Jose, California, United States
RemoteFull Time
Tredence
Tredence: Providing AI-driven data science and business analytics solutions.
4+ YOERequires 4+ years in data engineering, big data, or cloud data platforms; Scala, Apache Spark, GCP, SQL, ETL/ELT, Git, CI/CD, and DevOps expertise; bachelor's or master's degree preferred.
Scala, Apache Spark, Spark SQL, Spark Streaming, DataFrames, Datasets, Google Cloud Platform (GCP), BigQuery, Cloud Storage, Dataproc, Dataflow, Pub/Sub, Composer (Airflow), Cloud Functions, Cloud Run, SQL, Git, Kafka, Python, Docker, Kubernetes, GKE, Terraform, AWS, Azure
1mo
Save
Mark Applied
Hide
Data Engineer
Houston or Golden or Reno or Oakland or Salt Lake City
HybridFull Time
Fervo Energy
Fervo Energy: Generates clean electricity using advanced geothermal drilling technology.
2+ YOE2+ years building and operating production data pipelines with Apache Spark, Python/SQL, Databricks, Azure, and Snowflake; experience with streaming, data modeling, data quality, governance, and CI/CD.
Databricks, Delta Lake, Delta Live Tables, Unity Catalog, Databricks Workflows, Databricks SQL, Apache Spark, PySpark, Spark SQL, Structured Streaming, Kafka, Microsoft Event Hubs, Microsoft Azure Data Factory, Microsoft Azure Data Lake Storage (ADLS), Snowflake, Microsoft Power BI, Spotfire, Python, SQL, Git, Azure DevOps, GitHub Actions, dbt, Terraform, Docker, Airflow, Microsoft Entra ID, Microsoft Key Vault, MQTT, OPC UA, SparkplugB, Canary, Ignition, Snowflake Semantic Views, Databricks Metric Views
1mo
Save
Mark Applied
Hide
Sr. Data Engineer
Palo Alto, California, United States
$145k-$220k/yr RemoteFull Time
Allocate
Allocate: A fintech providing a data-rich platform for discovering, modeling, and managing private market investments.
5+ YOE5+ years data engineering experience; strong AWS (S3, EC2, EKS, Redshift, Glue, Athena) and cloud skills; SQL, Python, data modeling, ETL/ELT, and pipeline production experience; familiarity with vector DBs and ML data workflows.
S3, EC2, ECS, EKS, Athena, Redshift, Glue, Step Functions, Terraform, CloudFormation, SQL, Snowflake, Databricks Delta Lake, Neo4j, AWS Neptune, Python, pandas, PySpark, TypeScript, Node.js, C#, Postgres pgvector, Chroma, Pinecone, LangChain, LlamaIndex, Docker, Kubernetes, Kafka, Kinesis, Airflow, Prefect, dbt