353 data pipeline engineer jobs at 226 companies in American Canyon, CA
1mo
Save
Mark Applied
Hide
1mo
Principal Scientist - Data Pipeline Engineer
San Jose or Seattle or San Francisco
$206k-$388k/yrOnsiteFull Time
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years in data engineering/ML infrastructure, distributed systems expertise, Python and a systems language, experience with Ray or Spark, GPU inference optimization, large-scale databases and data curation for model training.
AVEVA: Industrial software for engineering and operational performance management.
3+ YOEBachelor's in a STEM field,minimum 3 years data engineering experience,proficiency with data pipelines,ETL/ELT,SQL,and cloud data platforms (Azure preferred).
Microsoft Azure, Azure Synapse Analytics, Azure Data Factory, Azure Data Lake, Python, Spark, SQL
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
3+ YOEDesign and manage Snowflake/Databricks data architecture, build and optimize ETL pipelines, deploy data governance, collaborate cross-functionally; 3+ years data engineering experience.
Bedrock Robotics: Automates heavy construction machinery with AI retrofit kits.
5+ YOE5+ years data engineering experience in large-scale data lake/warehouse environments; strong SQL and distributed query engine experience; pipeline orchestration (Airflow/Prefect) and Spark/Databricks experience; data quality and observability expertise.
Premier Nutrition CompanyNYSE: BRBR: Produces and distributes protein-based nutritional beverages and supplements.
5+ YOEDesign and maintain cloud-native ELT/ETL pipelines and Snowflake data platforms; 5+ years data engineering experience; strong SQL and Python; experience with dbt, Fivetran, data integration and observability.
Fivetran, dbt, Python, Snowflake, Oracle, SQL Server, GitHub, GitLab, Airflow, Azure Data Factory, Dell Boomi, NetSuite, Oracle EPM, O9, Salesforce, Snowflake Cortex, MCP, Tableau, SAP BO, Oracle BI, Microsoft Power BI, Azure, AWS, GCP, SFTP
Sakata Seed AmericaTokyo Stock Exchange: 1377: Breeding and distributing high-quality vegetable and flower seeds.
5+ YOEDesign and maintain scalable ETL/ELT pipelines, data models, and Microsoft Fabric solutions; strong SQL and Python skills; 5+ years data engineering experience; knowledge of data governance and monitoring.
Microsoft Fabric, Data Factory, Lakehouse, Warehouse, Dataflow Gen2, notebooks, semantic models, SQL Server, PostgreSQL, MySQL, SQL, Python, REST APIs, Spark
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building and operating production data pipelines, modeling messy domains into stable schemas, data quality engineering, extracting structure from unstructured sources, and shipping datasets with SLAs.
Mill: Provides smart kitchen bins that recycle food waste.
5+ YOE5+ years operating production data engineering systems; Python, dbt/Airflow/Fivetran experience; strong SQL and cloud data warehouse skills; recommendation systems and CI/CD for data pipelines; collaborative communication.
11x: Builds autonomous AI agents for sales and revenue teams.
4+ YOE4+ years software or data engineering experience; building production data systems and pipelines; proficiency with Python and TypeScript; strong backend and systems-thinking skills; comfortable with ambiguity and AI tooling.
Python, TypeScript, ClickHouse, Airbyte, Claude Code, Codex, Cursor, CRM, Customer Data Platform (CDP)
CrewAI: Platform for orchestrating collaborative multi-agent AI systems.
Experienced data engineer to own data platform, build pipelines, define trusted metrics, improve instrumentation, and make data self-serve for product and go-to-market teams.
Braintrust: Platform for evaluating and monitoring AI model performance.
10+ YOE10+ years in data engineering or related roles; deep SQL, data modeling, pipelines, orchestration, and production-grade code; experience across telemetry, CRM, billing, and operational systems; strong communication and AI experience.
Waystation: AI-native procurement platform for consumer packaged goods.
8+ YOE8+ years building production data systems with hands-on early-stage startup experience; deep Python and SQL skills; experience with extraction, ML/NLP pipelines, and owning data models and pipelines.
SS&C TechnologiesNasdaq: SSNC: Provides software and technology for financial and healthcare sectors.
Experienced in data modeling, large-scale data pipelines, ETL/ELT, Python and PySpark, machine learning/NLP, data warehousing, and collaboration across engineering and product teams.
United States or San Francisco or New York City or Washington or Austin
$118k-$179k/yrRemoteFull Time
SamsaraNYSE: IOT: Connected Operations Cloud platform for physical operations and IoT.
8+ YOEBachelor's in CS or equivalent, 8+ years as a software/data engineer, 5+ years building production data pipelines and Spark/PySpark experience, strong Python and SQL, cloud data warehouse and ETL tooling experience.
Houston or Golden or Reno or Oakland or Salt Lake City
HybridFull Time
Fervo Energy: Generates clean electricity using advanced geothermal drilling technology.
2+ YOE2+ years building and operating production data pipelines with Apache Spark, Python/SQL, Databricks, Azure, and Snowflake; experience with streaming, data modeling, data quality, governance, and CI/CD.
Databricks, Delta Lake, Delta Live Tables, Unity Catalog, Databricks Workflows, Databricks SQL, Apache Spark, PySpark, Spark SQL, Structured Streaming, Kafka, Microsoft Event Hubs, Microsoft Azure Data Factory, Microsoft Azure Data Lake Storage (ADLS), Snowflake, Microsoft Power BI, Spotfire, Python, SQL, Git, Azure DevOps, GitHub Actions, dbt, Terraform, Docker, Airflow, Microsoft Entra ID, Microsoft Key Vault, MQTT, OPC UA, SparkplugB, Canary, Ignition, Snowflake Semantic Views, Databricks Metric Views
5+ YOE5+ years data engineering experience, strong dbt and Databricks skills, Python/PySpark, streaming and batch ingestion, CDC pipelines, data quality and observability, and experience integrating with operational systems.
dbt, Databricks, Delta Lake, Delta Live Tables, Python, PySpark, Salesforce, NetSuite, Stripe, Mulesoft, Great Expectations, Unity Catalog, Airflow, Cloud Run, Cloud Functions, BigQuery, AWS, GCP, Terraform, Spacelift, Boomi
Denver or New York City or San Francisco or Los Angeles or Seattle or San Jose or Scottsdale
$155k-$220k/yrHybridFull Time
Gusto: Cloud-based payroll and HR software for small businesses.
8+ YOE8–10+ years in data engineering; SQL and Python, Scala, or Java; scalable pipelines, ETL, dbt, data modeling, cloud platforms, CI/CD, testing, observability, monitoring, and incident response.
Mariana Minerals: Building software-first infrastructure to produce and refine critical minerals.
4+ YOE4+ years in data engineering, strong Python and SQL, schema and time-series design, building and operating orchestrated pipelines in cloud with containers and CI/CD, data quality/observability/lineage, mentoring experience.
Robert HalfNYSE: RHI: Provides specialized staffing and global business consulting services.
5+ YOE5+ years Python/SQL data engineering; 3+ years cloud AWS; ETL/ELT pipelines; Docker/Kubernetes; Terraform/CloudFormation; Spark; CI/CD; collaboration with data science and MLOps; mentoring.