DataSpring
Posted 2mo ago

Data Engineer

DataSpring
United States
$100k-$115k/yrRemoteFull Time
Responsibilities
  • building pipelines
  • ensuring quality
  • supporting vendors
Requirements
  • 1–3 years data or analytics engineering experience
  • Strong SQL
  • Familiarity with Databricks, Delta Lake, and Azure SQL
  • Experience with Git, DevOps/CI/CD
  • Bachelor's degree preferred
  • Azure Data Engineer Associate certification preferred
Technical tools mentioned
DatabricksAzure SQLDelta LakeSQLGitDevOpsCI/CDConfluence

Job description

Position Summary
The Data Engineer will support the design, development, and maintenance of data pipelines and data models that power DataSpring's analytics and data platforms. This role focuses on implementing scalable data solutions, ensuring data quality, and collaborating with senior engineers and stakeholders to deliver reliable data for reporting and analysis. This position provides an opportunity to build hands-on experience with Databricks, Azure SQL, and modern data engineering practices while contributing to enterprise data initiatives and discussions.

 

The Data Engineer is a full-time, remote, exempt position and reports to the Sr. Director, Data Engineering & Architecture.


Specific Responsibilities


Data Engineering & Architecture


  • Build and maintain ETL/ELT pipelines across Databricks, Azure SQL, and downstream gold-layer models supporting priority projects
  • Support development and enhancement of enriched data models, including field-level enrichment logic, recency rules, and provider-level enrichment flags.
  • Assist in maintaining data logic, including reconciliation between source and target data sources and resolution of duplication and data discrepancies.
  • Assist in implementing medallion architecture patterns (bronze → gold), ensuring data quality, traceability, and performance at scale.


Data Quality, Governance & Reliability

  • Support identification and resolution of systemic data quality issues, including null handling, soft deletes, authorization flags, and incorrect organizational mappings.
  • Support implementation of rules for data in collaboration with product, governance, and engineering stakeholders.
  • Assist in documenting (Confluence, mapping workbooks) to serve as a single source of truth for enrichment logic and data behavior.


Cross‑Functional & Vendor Collaboration

  • Support collaboration with vendors and partners for vendors providing detailed queries, validation logic, and corrective guidance on upstream data issues.
  • Collaborate with product owners and engineering teams to ensure data models align with product defined use cases.
  • Support UAT and release readiness by preparing data, validating counts, and resolving last‑mile data defects under tight timelines.


Skills

  • Strong foundational SQL skills (complex joins, reconciliation, performance tuning).
  • Familiarity with Databricks, Delta Lake, and Azure SQL.
  • Basic understanding of data modeling for analytical, operational, and API‑driven use cases.
  • Ability to support troubleshooting of messy, evolving enterprise data domains.
  • Excellent written and verbal communication, especially for explaining complex data behavior to non‑technical stakeholders.
  • Experience using Git, DevOps tools, and CI/CD pipelines for data engineering workflows.


Experience

  • 1–3 years of experience in a data engineering or analytics engineering role, including internships or academic projects.
  • Demonstrated success contributing to data modernization or migration initiatives in cloud environments.
  • Prior experience working with healthcare or other regulated data environments is highly desirable.
  • Bachelor’s degree in Computer Science, Information Systems, Data Engineering, or a related field.
  • Azure Data Engineer Associate or related certification (preferred).
  • Coursework or certification in AI/ML (preferred but not required).


Who We Are

DataSpring is the trusted data connector at the core of healthcare. For more than 25 years, we have powered the industry with the largest and most complete healthcare data foundation in the U.S., including more than 4.8 million provider data records sourced directly from providers and member data representing 75% of covered lives supplied by health plans. By improving how essential information flows across the system, DataSpring helps healthcare operate more efficiently, accurately, and with greater confidence.

What You Get

At DataSpring, you will do meaningful work at the intersection of healthcare, data, and technology, helping solve complex problems that make the healthcare system work better. You will collaborate with experienced professionals who care deeply about accuracy, trust, and meaningful impact in a fully remote environment.

DataSpring offers competitive compensation and a comprehensive benefits package for full-time employees, including medical, dental, and vision coverage, a 401(k) with company contributions and matching, paid parental leave, tuition assistance, and generous paid time off. We are committed to investing in our people and supporting professional growth over time.

Equal Opportunity Employer

DataSpring is proud to be an equal opportunity employer and is committed to fostering a workplace where all individuals are valued, respected, and empowered.

Employment decisions at DataSpring are made without regard to race, color, religion, sex, national origin or ancestry, age, marital status, disability, protected veteran status, personal appearance, sexual orientation, gender identity or expression, familial status, family responsibilities, matriculation, political affiliation, genetic information, source of income, place of residence, or any other characteristic protected by law. DataSpring does not tolerate unlawful discrimination or harassment of any kind.

Applicants have rights under the Family and Medical Leave Act (FMLA), Equal Employment Opportunity (EEO), and the Employee Polygraph Protection Act (EPPA). If you need a reasonable accommodation to apply for a posted position, please contact the DataSpring People & Culture team at [email protected] or 202-517-0436.

About DataSpring

Provider data management and healthcare administrative automation services.

Similar jobs

Data Engineer roles
3h
Save
Mark Applied
Hide
Data Engineer
Miami, Florida, United States
OnsiteFull Time
Base-2 Solutions
Base-2 Solutions: Delivers engineering and technology services for national security.
5+ YOEBachelor's degree in computer science, data engineering, or related field, or 5 years equivalent experience; full-stack data applications, APIs, Kubernetes, CI/CD, databases, and geospatial platforms required.
Databricks, Kubernetes, ESRI ArcGIS, Unity Catalog, RESTful, GraphQL, AWS Certified Data Analytics - Specialty, Microsoft Azure Data Engineer Associate, Microsoft Azure, CI/CD
1d
Save
Mark Applied
Hide
Senior Data Engineer, Databricks
Denver, Colorado, United States
$119k-$146k/yr OnsiteFull Time
Bouygues
BouyguesEuronext Paris: EN: Diversified industrial group focused on construction, energy, and telecommunications.
8+ YOEBachelor's degree or equivalent experience; 8+ years in data engineering, warehousing, and BI; 5+ years building Databricks/Spark platforms; expertise in Python, SQL, modeling, governance, Azure, migration, and CI/CD.
Databricks, Databricks SQL, Apache Spark, PySpark, Spark SQL, Delta Live Tables, Unity Catalog, Databricks Workflows, Delta Lake, Databricks Asset Bundles, Power BI, Microsoft SQL Server, T-SQL, Azure Synapse, Azure Data Factory, Qlik Replicate, Git, Azure DevOps, GitHub Actions, Microsoft Fabric, OneLake, ADLS Gen2, Key Vault, Entra ID, Terraform, SSIS, SSRS, Databricks Genie Agents, Photon, Auto Loader, Structured Streaming, MERGE, OPTIMIZE, Z-ORDER, VACUUM, Direct Lake, DirectQuery, JD Edwards, BMS, CMS, Cority, HCSS, Intelex, Anaplan, Datavail, Birlasoft
1d
Save
Mark Applied
Hide
Data Engineer
Dearborn, Michigan, United States
OnsiteFull Time
Ford Motor Company
Ford Motor CompanyNYSE: F: Designs, manufactures, and sells cars, trucks, and SUVs.
5+ YOEBachelor's degree or equivalent; 5+ years Python, 4+ years automotive AI/ML, and experience with data pipelines, databases, APIs, cloud platforms, agentic AI, and machine learning operations.
Google Cloud Platform (GCP), Amazon Web Services (AWS), Vertex AI, JavaScript, TypeScript, React, HTML, Next.js, Node.js, Python, Go, SQL, NoSQL, REST, GraphQL, Flask, FastAPI, LangChain, LangGraph, Google Agent Development Kit (ADK), Model Context Protocol (MCP), Pandas, PyTorch, TensorFlow, Keras, Hugging Face, Docker, Kubernetes, Terraform, RAGEval
1d
Save
Mark Applied
Hide
Data Engineer
Dearborn, Michigan, United States
$100k-$163k/yr HybridFull Time
Ford Motor Company
Ford Motor CompanyNYSE: F: Ford manufactures, sells, and services vehicles and mobility solutions.
4+ YOEMaster's degree and 4 years of experience, or bachelor's degree with 6+ years; 4 years in data engineering, 3 years in cloud data engineering, and proficiency in at least three named programming languages and cloud data technologies.
Java, Python, Spark, Scala, SQL, Amazon Redshift, Microsoft Azure Synapse Analytics, Google BigQuery, Airflow, Cloud Run, MySQL, PostgreSQL, SQL Server, Apache Kafka, GCP Pub/Sub, Tekton, GitHub Actions, Git, GitHub, Terraform, Docker, Atlassian JIRA, Google Cloud Platform (GCP), Amazon Web Services (AWS), Vertex AI, SonarQube, Checkmarx, Fossa, Cycode, Test Driven Development (TDD), CI/CD, REST APIs, MLOps
2d
Save
Mark Applied
Hide
Business Intelligence - Data Engineer 133-2015
Tulsa, Oklahoma, United States
OnsiteFull Time
CommunityCare
CommunityCare: Provides community-based health insurance and managed care services in Oklahoma.
College degree or equivalent experience; data pipeline, ETL and SSIS experience; process improvement, project management, organization, independence and multitasking skills.
SSIS, ETL
2d
Save
Mark Applied
Hide
Senior Data Engineer, DX
Salt Lake City, Utah, United States
$139k-$219k/yr HybridFull Time
Atlassian
AtlassianNASDAQ: TEAM: Develops software for team collaboration and project management.
Requires strong SQL, Postgres or relational database experience, ETL/ELT pipeline development, large-scale data cleaning and validation, documentation skills, and ability to manage recurring deadlines independently.
SQL, Postgres, JSONB, ETL, ELT, Git, GitHub, GitLab, Bitbucket, CI/CD, Jira, DORA, SPACE, DevEx
2d
Save
Mark Applied
Hide
Senior Data Engineer, Databricks
Denver, Colorado, United States
$119k-$146k/yr OnsiteFull Time
Colas
Colas: Constructs and maintains transportation infrastructure and road systems globally.
5+ YOEBachelor’s degree or equivalent experience; 5+ years building Databricks/Spark platforms and 8+ years overall data engineering. Requires enterprise lakehouse ownership, legacy migration, and hands-on production coding.
Databricks, Apache Spark, Databricks SQL, Databricks Workflows, Delta Live Tables, Unity Catalog, Auto Loader, Photon, PySpark, Spark SQL, Delta Lake, Power BI, SQL, T-SQL, ADLS Gen2, Azure Data Factory, Azure Synapse, Key Vault, Entra ID, Qlik Replicate, Structured Streaming, Git, Azure DevOps, GitHub Actions, Databricks Asset Bundles, Databricks Repos, Terraform, Microsoft Fabric, OneLake, SQL Server, SSIS, SSRS, Direct Lake, DirectQuery, Import
2d
Save
Mark Applied
Hide
Engineer Data Integration
San Antonio, Texas, United States
$64k-$85k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
4+ YOERequires strong SQL, Snowflake data warehouse, data modeling, data analysis, validation, data quality, Tableau dashboard development, Agile/Scrum familiarity, and Jira experience.
SQL, Snowflake, Tableau, Jira, dbt