25 batch data engineer jobs at 21 companies in Los Altos, CA
1w
Save
Mark Applied
Hide
1w
Staff Data Engineer
Sunnyvale, California, United States
$200k-$260k/yrOnsiteFull Time
RADAR: Real-time inventory tracking using RFID and computer vision technology.
8+ YOE8+ years data/analytics engineering experience, strong SQL and Python, experience with batch and streaming pipelines, data modeling, data quality, testing, monitoring, and mentoring engineers.
Bellevue or Chicago or San Francisco or Washington
$197k-$271k/yrHybridFull Time
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
10+ YOE10+ years data engineering experience; expert SQL, ETL/ELT, MPP databases, AWS services; HR/People data experience; experience with batch and realtime pipelines, lakehouse architectures, and data quality frameworks.
5+ YOE5+ years data engineering experience, strong dbt and Databricks skills, Python/PySpark, streaming and batch ingestion, CDC pipelines, data quality and observability, and experience integrating with operational systems.
dbt, Databricks, Delta Lake, Delta Live Tables, Python, PySpark, Salesforce, NetSuite, Stripe, Mulesoft, Great Expectations, Unity Catalog, Airflow, Cloud Run, Cloud Functions, BigQuery, AWS, GCP, Terraform, Spacelift, Boomi
Beast Industries: Produces digital media and consumer goods for MrBeast brands.
8+ YOE8+ years building production data platforms, expertise in batch and streaming pipelines, data modeling, ETL/ELT, reliability, observability, and privacy-safe foundations for AI/ML.
Metriport: Provides an open-source API for healthcare data interoperability.
6+ YOE6+ years data engineering experience building and scaling pipelines; experience across ingestion, storage, processing, warehousing; production coding in TypeScript/Python; mentoring experience; located in or willing to relocate to San Francisco Bay Area.
BILLNYSE: BILL: Automated financial operations software for small and midsize businesses.
8+ YOELead architecture and delivery of data platform capabilities (ingest, lake, streaming, feature store, query, graph, search). Requires distributed systems, streaming and batch experience, SQL and Python, and technical leadership.
xAI: Develops advanced artificial intelligence systems to understand the universe.
3+ YOE3+ years software engineering experience; expertise in Python,Rust,Scala,Go or Java; experience with data pipelines, realtime and batch processing, and distributed systems.
Physical Intelligence: Creating foundation models for general-purpose robot intelligence.
Strong software engineering fundamentals with experience building distributed systems, large-scale data pipelines, object storage and batch/streaming systems; ownership mindset and performance focus.
Experienced data engineer with production-grade pipeline experience for large-scale batch and streaming workloads; strong debugging, cost/performance optimization, and data quality/governance skills.
Spark, Flink, Beam, Airflow, Dagster, Kafka, PubSub, Parquet, Iceberg, Delta Lake, BigQuery, Snowflake, Great Expectations
Member of Technical Staff — Data Ingestion & Quality
San Francisco, California, United States
OnsiteFull Time
Causal Labs: Building physics-based causal AI models for predictive weather intelligence.
Experience building large-scale data pipelines and QA systems, familiarity with streaming and batch ingestion, vendor collaboration, and strong problem-solving and domain learning skills.
Hinge HealthNYSE: HNGE: Digital provider of musculoskeletal care and physical therapy
5+ YOE2+ Mgmt5+ years data engineering; 2+ years managing engineering teams; 2+ years building ML platform capabilities; experience with batch & streaming systems and tools such as Kafka, Flink, Spark; proficiency with Python, SQL, dbt, Databricks, and AWS.
IntuitiveNASDAQ: ISRG: Robotic-assisted systems for minimally invasive surgery.
8+ YOEBachelor's or Master's degree, minimum 8 years software engineering experience building data-intensive services, expertise in Java/Go/Python, stream and batch processing, data pipelines, APIs, CI/CD, containers, and AWS.
Senior Manager, Content Promotion & Distribution Data Engineering
Los Angeles or Los Gatos
$525k-$950k/yrOnsiteFull Time
NetflixNASDAQ: NFLX: Provider of global streaming entertainment and video content.
7+ Mgmt7+ years leading data engineering teams; expertise in data modeling, batch/streaming pipelines, data quality, and multi-modal media pipelines for ML/GenAI; strong stakeholder communication and people management.
WalmartNYSE: WMT: Multinational retail operating discount stores and supermarkets.
15+ YOE4+ Mgmt15+ years building data integration and platform solutions; experience with RDBMS and NoSQL, Hadoop/Spark, data modeling, real-time and batch pipelines; leadership of distributed platform teams; strong solution architecture and DevOps knowledge.
Oracle, My SQL, Cassandra, HBase, Hadoop, Spark, DevOps, Web Content Accessibility Guidelines (WCAG) 2.2 AA
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
6+ YOE6+ years in AI/ML engineering or applied data science; strong Python in production; ML model development and deployment; data pipelines (ETL/ELT, batch or streaming); APIs and backend systems; experience with LLM-powered systems and agent workflows; familiarity with Spark, Airflow/Dagster, Snowflake/BigQuery.
Duckbill: SaaS platform for enterprise cloud financial planning and analysis.
Proven experience building data systems and ETL for batch and streaming; strong Python and SQL; experience with data warehouses/lakehouses/OLAP, columnar databases, cloud storage, and data quality practices.
Member of Technical Staff (Software Engineer, Data Platform)
San Francisco or New York City or Palo Alto
$220k-$405k/yrOnsiteFull Time
Perplexity: AI-powered search engine providing conversational answers with citations.
5+ YOE5+ years software engineering experience (8+ for staff), production data infrastructure and batch/streaming experience, Airflow/Dagster, Python plus another backend language (Go/TypeScript), ML/AI workflow support, data quality and observability knowledge.
Verily: Developing data-driven technologies for clinical research and precision health.
6+ YOEBA/BS in CS or equivalent, 6+ years experience with Python and/or Java, SQL proficiency, experience with life science/biomedical data and batch workflows, strong communication and project management skills.
Python, Java, SQL, Google Cloud Platform, Amazon Web Services, Microsoft Azure, Docker, Linux
Beacon AI: Developing an AI-powered-pilot for safer flight operations.
Experience designing and operating AWS cloud and LLM/ML infrastructure, building data pipelines, and implementing security and observability for production systems.
Parallel: Build web infrastructure and search APIs for AI agents.
Deep experience in distributed data processing, data modeling, system reliability, batch and streaming pipelines, and building data quality, lineage, and observability systems.