436 data pipeline engineer jobs at 273 companies in Kentfield, CA
4w
Save
Mark Applied
Hide
4w
Principal Scientist - Data Pipeline Engineer
San Jose or Seattle or San Francisco
$206k-$388k/yrOnsiteFull Time
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years in data engineering/ML infrastructure, distributed systems expertise, Python and a systems language, experience with Ray or Spark, GPU inference optimization, large-scale databases and data curation for model training.
AVEVA: Industrial software for engineering and operational performance management.
3+ YOEBachelor's in a STEM field,minimum 3 years data engineering experience,proficiency with data pipelines,ETL/ELT,SQL,and cloud data platforms (Azure preferred).
Microsoft Azure, Azure Synapse Analytics, Azure Data Factory, Azure Data Lake, Python, Spark, SQL
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
3+ YOEDesign and manage Snowflake/Databricks data architecture, build and optimize ETL pipelines, deploy data governance, collaborate cross-functionally; 3+ years data engineering experience.
ParamountNASDAQ: PSKY: Global media producing films, television, and streaming content.
2+ YOEBachelor's or master's degree in computer science, engineering, or related field; 2–4+ years of data engineering experience; advanced SQL, Python, distributed systems, cloud architectures, and production data pipelines.
Bedrock Robotics: Automates heavy construction machinery with AI retrofit kits.
5+ YOE5+ years data engineering experience in large-scale data lake/warehouse environments; strong SQL and distributed query engine experience; pipeline orchestration (Airflow/Prefect) and Spark/Databricks experience; data quality and observability expertise.
Premier Nutrition CompanyNYSE: BRBR: Produces and distributes protein-based nutritional beverages and supplements.
5+ YOEDesign and maintain cloud-native ELT/ETL pipelines and Snowflake data platforms; 5+ years data engineering experience; strong SQL and Python; experience with dbt, Fivetran, data integration and observability.
Fivetran, dbt, Python, Snowflake, Oracle, SQL Server, GitHub, GitLab, Airflow, Azure Data Factory, Dell Boomi, NetSuite, Oracle EPM, O9, Salesforce, Snowflake Cortex, MCP, Tableau, SAP BO, Oracle BI, Microsoft Power BI, Azure, AWS, GCP, SFTP
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building and operating production data pipelines, modeling messy domains into stable schemas, data quality engineering, extracting structure from unstructured sources, and shipping datasets with SLAs.
Mill: Provides smart kitchen bins that recycle food waste.
5+ YOE5+ years operating production data engineering systems; Python, dbt/Airflow/Fivetran experience; strong SQL and cloud data warehouse skills; recommendation systems and CI/CD for data pipelines; collaborative communication.
11x: Builds autonomous AI agents for sales and revenue teams.
4+ YOE4+ years software or data engineering experience; building production data systems and pipelines; proficiency with Python and TypeScript; strong backend and systems-thinking skills; comfortable with ambiguity and AI tooling.
Python, TypeScript, ClickHouse, Airbyte, Claude Code, Codex, Cursor, CRM, Customer Data Platform (CDP)
RADAR: Real-time inventory tracking using RFID and computer vision technology.
8+ YOE8+ years data/analytics engineering experience, strong SQL and Python, experience with batch and streaming pipelines, data modeling, data quality, testing, monitoring, and mentoring engineers.
CrewAI: Platform for orchestrating collaborative multi-agent AI systems.
Experienced data engineer to own data platform, build pipelines, define trusted metrics, improve instrumentation, and make data self-serve for product and go-to-market teams.
Braintrust: Platform for evaluating and monitoring AI model performance.
10+ YOE10+ years in data engineering or related roles; deep SQL, data modeling, pipelines, orchestration, and production-grade code; experience across telemetry, CRM, billing, and operational systems; strong communication and AI experience.
Waystation: AI-native procurement platform for consumer packaged goods.
8+ YOE8+ years building production data systems with hands-on early-stage startup experience; deep Python and SQL skills; experience with extraction, ML/NLP pipelines, and owning data models and pipelines.
SS&C TechnologiesNasdaq: SSNC: Provides software and technology for financial and healthcare sectors.
Experienced in data modeling, large-scale data pipelines, ETL/ELT, Python and PySpark, machine learning/NLP, data warehousing, and collaboration across engineering and product teams.
Bellevue or Chicago or San Francisco or Washington
$197k-$271k/yrHybridFull Time
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
10+ YOE10+ years data engineering experience; expert SQL, ETL/ELT, MPP databases, AWS services; HR/People data experience; experience with batch and realtime pipelines, lakehouse architectures, and data quality frameworks.
San Francisco or Denver or New York or Culver City
$133k-$160k/yrHybridFull Time
FastlyNYSE: FSLY: Provides edge cloud platform for content delivery and cybersecurity.
3+ YOE3+ years in BI/analytics; experience with cloud data warehouses (BigQuery a plus), dbt, data visualization, strong SQL, and ability to build data pipelines and data products for technical and non-technical stakeholders.
Google Cloud Platform (GCP), BigQuery, dbt, Looker, Tableau, Mode, SQL
United States or San Francisco or New York City or Washington or Austin
$118k-$179k/yrRemoteFull Time
SamsaraNYSE: IOT: Connected Operations Cloud platform for physical operations and IoT.
8+ YOEBachelor's in CS or equivalent, 8+ years as a software/data engineer, 5+ years building production data pipelines and Spark/PySpark experience, strong Python and SQL, cloud data warehouse and ETL tooling experience.
Houston or Golden or Reno or Oakland or Salt Lake City
HybridFull Time
Fervo Energy: Generates clean electricity using advanced geothermal drilling technology.
2+ YOE2+ years building and operating production data pipelines with Apache Spark, Python/SQL, Databricks, Azure, and Snowflake; experience with streaming, data modeling, data quality, governance, and CI/CD.
Databricks, Delta Lake, Delta Live Tables, Unity Catalog, Databricks Workflows, Databricks SQL, Apache Spark, PySpark, Spark SQL, Structured Streaming, Kafka, Microsoft Event Hubs, Microsoft Azure Data Factory, Microsoft Azure Data Lake Storage (ADLS), Snowflake, Microsoft Power BI, Spotfire, Python, SQL, Git, Azure DevOps, GitHub Actions, dbt, Terraform, Docker, Airflow, Microsoft Entra ID, Microsoft Key Vault, MQTT, OPC UA, SparkplugB, Canary, Ignition, Snowflake Semantic Views, Databricks Metric Views
5+ YOE5+ years data engineering experience, strong dbt and Databricks skills, Python/PySpark, streaming and batch ingestion, CDC pipelines, data quality and observability, and experience integrating with operational systems.
dbt, Databricks, Delta Lake, Delta Live Tables, Python, PySpark, Salesforce, NetSuite, Stripe, Mulesoft, Great Expectations, Unity Catalog, Airflow, Cloud Run, Cloud Functions, BigQuery, AWS, GCP, Terraform, Spacelift, Boomi