279 data pipeline engineer jobs at 192 companies in Cotati, CA
3d
Save
Mark Applied
Hide
3d
Data Pipeline Engineer
Redmond or Las Colinas or Fargo or Charlotte or San Francisco or New York City
$86k-$170k/yrHybridFull Time
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
3+ YOEBachelor's degree in a technical field and 3+ years of relevant experience, or 5+ years of relevant experience. Requires SQL, Python or Scala, ETL/ELT, data modeling, cloud platforms, Microsoft Fabric, Spark, Kafka, and security screening.
SQL, Python, Scala, ETL, ELT, Airflow, Azure Data Factory, Microsoft Fabric, Spark, Kafka, Synapse, OneLake, Lakehouse, CI/CD, Azure Event Hubs, Azure SQL Managed Instances, Power BI Gateways, DataFlows
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years in data engineering/ML infrastructure, distributed systems expertise, Python and a systems language, experience with Ray or Spark, GPU inference optimization, large-scale databases and data curation for model training.
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
3+ YOEDesign and manage Snowflake/Databricks data architecture, build and optimize ETL pipelines, deploy data governance, collaborate cross-functionally; 3+ years data engineering experience.
Bedrock Robotics: Automates heavy construction machinery with AI retrofit kits.
5+ YOE5+ years data engineering experience in large-scale data lake/warehouse environments; strong SQL and distributed query engine experience; pipeline orchestration (Airflow/Prefect) and Spark/Databricks experience; data quality and observability expertise.
Premier Nutrition CompanyNYSE: BRBR: Produces and distributes protein-based nutritional beverages and supplements.
5+ YOEDesign and maintain cloud-native ELT/ETL pipelines and Snowflake data platforms; 5+ years data engineering experience; strong SQL and Python; experience with dbt, Fivetran, data integration and observability.
Fivetran, dbt, Python, Snowflake, Oracle, SQL Server, GitHub, GitLab, Airflow, Azure Data Factory, Dell Boomi, NetSuite, Oracle EPM, O9, Salesforce, Snowflake Cortex, MCP, Tableau, SAP BO, Oracle BI, Microsoft Power BI, Azure, AWS, GCP, SFTP
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building and operating production data pipelines, modeling messy domains into stable schemas, data quality engineering, extracting structure from unstructured sources, and shipping datasets with SLAs.
11x: Builds autonomous AI agents for sales and revenue teams.
4+ YOE4+ years software or data engineering experience; building production data systems and pipelines; proficiency with Python and TypeScript; strong backend and systems-thinking skills; comfortable with ambiguity and AI tooling.
Python, TypeScript, ClickHouse, Airbyte, Claude Code, Codex, Cursor, CRM, Customer Data Platform (CDP)
CrewAI: Platform for orchestrating collaborative multi-agent AI systems.
Experienced data engineer to own data platform, build pipelines, define trusted metrics, improve instrumentation, and make data self-serve for product and go-to-market teams.
Braintrust: Platform for evaluating and monitoring AI model performance.
10+ YOE10+ years in data engineering or related roles; deep SQL, data modeling, pipelines, orchestration, and production-grade code; experience across telemetry, CRM, billing, and operational systems; strong communication and AI experience.
SS&C TechnologiesNasdaq: SSNC: Provides software and technology for financial and healthcare sectors.
Experienced in data modeling, large-scale data pipelines, ETL/ELT, Python and PySpark, machine learning/NLP, data warehousing, and collaboration across engineering and product teams.
Bellevue or Chicago or San Francisco or Washington
$197k-$271k/yrHybridFull Time
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
10+ YOE10+ years data engineering experience; expert SQL, ETL/ELT, MPP databases, AWS services; HR/People data experience; experience with batch and realtime pipelines, lakehouse architectures, and data quality frameworks.
United States or San Francisco or New York City or Washington or Austin
$118k-$179k/yrRemoteFull Time
SamsaraNYSE: IOT: Connected Operations Cloud platform for physical operations and IoT.
8+ YOEBachelor's in CS or equivalent, 8+ years as a software/data engineer, 5+ years building production data pipelines and Spark/PySpark experience, strong Python and SQL, cloud data warehouse and ETL tooling experience.
Gallup: Provides global analytics and advice for leaders and organizations.
5+ YOE5+ years data engineering or backend data systems; strong SQL and data modeling; production pipelines; cloud data platforms; Python; Airflow/Dagster/dbt; on-site at San Francisco.
5+ YOE5+ years data engineering experience, strong dbt and Databricks skills, Python/PySpark, streaming and batch ingestion, CDC pipelines, data quality and observability, and experience integrating with operational systems.
dbt, Databricks, Delta Lake, Delta Live Tables, Python, PySpark, Salesforce, NetSuite, Stripe, Mulesoft, Great Expectations, Unity Catalog, Airflow, Cloud Run, Cloud Functions, BigQuery, AWS, GCP, Terraform, Spacelift, Boomi
Houston or Golden or Reno or Oakland or Salt Lake City
HybridFull Time
Fervo Energy: Generates clean electricity using advanced geothermal drilling technology.
2+ YOE2+ years building and operating production data pipelines with Apache Spark, Python/SQL, Databricks, Azure, and Snowflake; experience with streaming, data modeling, data quality, governance, and CI/CD.
Databricks, Delta Lake, Delta Live Tables, Unity Catalog, Databricks Workflows, Databricks SQL, Apache Spark, PySpark, Spark SQL, Structured Streaming, Kafka, Microsoft Event Hubs, Microsoft Azure Data Factory, Microsoft Azure Data Lake Storage (ADLS), Snowflake, Microsoft Power BI, Spotfire, Python, SQL, Git, Azure DevOps, GitHub Actions, dbt, Terraform, Docker, Airflow, Microsoft Entra ID, Microsoft Key Vault, MQTT, OPC UA, SparkplugB, Canary, Ignition, Snowflake Semantic Views, Databricks Metric Views
Candid Health: Automates medical billing and revenue cycle management for healthcare providers.
4+ YOEBachelor's in CS/CE/Math or similar, 4+ years building data pipelines/products, experience with modern data warehouse patterns, and strong communication and technical design skills.
Mariana Minerals: Building software-first infrastructure to produce and refine critical minerals.
4+ YOE4+ years in data engineering, strong Python and SQL, schema and time-series design, building and operating orchestrated pipelines in cloud with containers and CI/CD, data quality/observability/lineage, mentoring experience.
Regard: A healthcare-focused software building AI-powered documentation tools to improve care delivery and clinician workflow.
5+ YOE5+ years data engineering experience, 3+ years PySpark and public cloud (AWS S3/EMR), strong Python and SQL, BS in CS/related or equivalent, experience with data modeling, pipelines, and on-call support.
Clera: AI talent agent matching professionals with high-growth startup roles
Requires SQL, PostgreSQL, end-to-end data pipeline design, AWS, and TypeScript experience; strong communication and customer collaboration skills preferred.