Steampunk
Posted 4d ago

Data Engineer - Databricks

Steampunk
McLean, Virginia, United States
$110k-$160k/yrOnsiteFull Time
Responsibilities
  • architecting migrations
  • optimizing queries
  • inspecting pipelines
Requirements
  • Requires 2–4 years of data engineering and commercial software coding experience
  • Databricks, SQL
  • PySpark/Python, AWS
  • Data architecture
  • Pipelines, Agile, and ability to hold US government Public Trust
Technical tools mentioned
DatabricksSQLPySparkPythonAWSApache SparkDelta LakeDatabricks WorkflowsAirflowStep FunctionsS3EC2RDSAzureGCPT-SQLpgSQLMySQLJavaC++ScalaApache IcebergCI/CDInformatica EDCUnity CatalogCollibraAlationPurviewDataZone

Job description

We are looking for seasoned Data Engineer to work with our team and our clients to develop enterprise grade data platforms, services, and pipelines in Databricks. We are looking for more than just a "Data Engineer", but a technologist with excellent communication and customer service skills and a passion for data and problem solving. 

 

 

  • Lead and architect migrations of data using Databricks with focus on performance, reliability, and scalability. 
  • Assess and understand ETL jobs, workflows, data marts, BI tools, and reports 
  • Address technical inquiries concerning customization, integration, enterprise architecture and general feature/functionality of data products 
  • Experience working with database/data warehouse/data mart solutions in cloud (Preferably AWS. Alternatively Azure, GCP). 
  • Key must have skill sets – Databricks, SQL, PySpark/Python, AWS 
  • Support an Agile software development lifecycle 
  • You will contribute to the growth of our AI & Data Exploitation Practice! 

Required:

 

  • Ability to hold a position of public trust with the US government. 
  • 2-4 years industry experience coding commercial software and a passion for solving complex problems.  
  • 2-4 years direct experience in Data Engineering with experience in tools such as: 
  • Big data tools: Databricks, Apache Spark, Delta Lake, etc. 
  • Relational SQL (Preferably T-SQL. Alternatively pgSQL, MySQL). 
  • Data pipeline and workflow management tools: Databricks Workflows, Airflow, Step Functions, etc. 
  • AWS cloud services: Databricks on AWS, S3, EC2, RDS (or Azure equivalents). 
  • Object-oriented/object function scripting languages: PySpark/Python, Java, C++, Scala, etc. 
  • Experience working with Data Lakehouse architecture and Delta Lake/Apache Iceberg 
  • Advanced working SQL knowledge and experience working with relational databases, query authoring and optimization (SQL) as well as working familiarity with a variety of databases. 
  • Experience manipulating, processing, and extracting value from large, disconnected datasets. 
  • Ability to inspect existing data pipelines, discern their purpose and functionality, and re-implement them efficiently in Databricks. 
  • Experience manipulating structured and unstructured data. 
  • Experience architecting data systems (transactional and warehouses). 
  • Experience the SDLC, CI/CD, and operating in dev/test/prod environments. 
  • Experience with data cataloging tools such as Informatica EDC, Unity Catalog, Collibra, Alation, Purview, or DataZone is a plus. 
  • Commitment to data governance. 
  • Experience working in an Agile environment. 
  • Experience supporting project teams of developers and data scientists who build web-based interfaces, dashboards, reports, and analytics/machine learning models 

 

Steampunk relies on several factors to determine salary, including but not limited to geographic location, contractual requirements, education, knowledge, skills, competencies, and experience. The projected compensation range for this position is $110,000 to $160,000.  The estimate displayed represents a typical annual salary range for this position. Annual salary is just one aspect of Steampunk’s total compensation package for employees. Learn more about additional Steampunk benefits here. 

 

Identity Statement

As part of the application process, you are expected to be on camera during interviews and assessments. We reserve the right to take your picture to verify your identity and prevent fraud.

 

Steampunk is a Change Agent in the Federal contracting industry, bringing new thinking to clients in the Homeland, Federal Civilian, Health and DoD sectors.  Through our Human-Centered delivery methodology, we are fundamentally changing the expectations our Federal clients have for true shared accountability in solving their toughest mission challenges.  As an employee owned company, we focus on investing in our employees to enable them to do the greatest work of their careers – and rewarding them for outstanding contributions to our growth. If you want to learn more about our story, visit http://www.steampunk.com. 

 

We are an equal opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, disability status, protected veteran status, or any other characteristic protected by law. Steampunk participates in the E-Verify program.

 

About Steampunk

Federal contractor providing human-centered IT and mission solutions.

Year founded
2003
Employees
750
Organization type
Private
Headquarters
US

Similar jobs

Data Engineer roles near McLean, Virginia
14h
Save
Mark Applied
Hide
Data Engineer, PXT Central Science
Seattle or Bellevue or Arlington or San Francisco
$132k-$206k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Multinational technology focused on e-commerce and cloud computing.
3+ YOE3+ years of data engineering experience, programming in Python, Java, Scala, or NodeJS, data modeling, warehousing, ETL, AWS, and non-relational database experience required.
AWS, AWS Glue, EMR, Lambda, Redshift, S3, Kinesis, Firehose, IAM, Python, Java, Scala, NodeJS, Hadoop, Hive, Spark
2d
Save
Mark Applied
Hide
Sr. Data Engineer - Reston, VA - Freewheel
Reston, Virginia, United States
$134k-$201k/yr OnsiteFull Time
FreeWheel
FreeWheelNASDAQ: CMCSA: A global media and technology.
7+ YOEBachelor's degree preferred and 7–10 years of relevant experience. Requires advanced SQL, programming, cloud data warehousing, orchestration, data modeling, observability, and distributed computing expertise.
Structured Query Language (SQL), Python, Apache Spark, AWS, Google Cloud Platform (GCP), Microsoft Azure, dbt, Airflow, Kubernetes, Teradata, Databricks, Snowflake
2d
Save
Mark Applied
Hide
Lead Data Engineer
United States or New York City or Washington or Oakland or Boulder or Basalt
$94k-$116k/yr RemoteFull Time
RMI
RMI: Independent nonprofit transforming global energy systems through market-driven clean-energy solutions for businesses, policymakers, communities, and funders.
3+ YOERequires 3–5 years in data engineering, data science, or a related field; Python, R, SQL, relational databases, version control, LLM APIs, cloud workflows, and strong stakeholder communication.
Python, R, Git, OpenAI, Anthropic, MySQL, PostgreSQL, SQL, Microsoft Azure, Azure SQL, Azure Functions, Azure Blob Storage, Workday, Salesforce, ChatGPT, Claude, Power BI, MCP
2d
Save
Mark Applied
Hide
Data Engineer 1, Operational Technology - Operations #4941
Durham or California or Washington or North Carolina or United Kingdom or New York
$86k-$106k/yr OnsiteFull Time
GRAIL
GRAILNasdaq Global Select Market: GRAL: Public healthcare biotechnology developing blood-based tests to detect multiple cancers early for patients and providers.
1+ YOEDegree in a relevant quantitative or technical field, 1+ year of relevant experience, SQL and programming proficiency, and understanding of ETL/ELT pipelines and relational databases. Strong analytical, communication, and data quality skills required.
SQL, Python, Rust, C++, Airflow, dbt, AWS S3, Redshift, Glue, Snowflake, Git, APIs
2d
Save
Mark Applied
Hide
Data Engineer
Arlington or United States
$95k-$157k/yr HybridFull Time
ManTech International
ManTech International: Advancing the future of defense through innovative technology.
2+ YOERequires 2–7+ years in data science, analytics, or a related technical field; a relevant bachelor's degree; Python and SQL pipeline development; and an IRS Public Trust clearance with full background investigation.
Python, SQL, Java, Databricks, Git, SVN, Mercurial, venv, conda, AWS, Unity Catalog, PySpark, Spark SQL, Power BI, Tableau, CRISP-DM, bash
2d
Save
Mark Applied
Hide
Data Engineer
Fort Gregg-Adams or Fort Lee or Tysons
$80k-$91k/yr OnsiteFull Time
LMI
LMI: Consulting and technology solutions for federal government agencies.
1+ YOEBachelor's degree or equivalent practical experience, 1+ year relevant engineering experience, Python and SQL, ETL/ELT, Git, and ability to obtain a DoD Secret clearance; U.S. citizenship required.
Python, SQL, Git, Palantir Foundry, Databricks, AWS, Azure
3d
Save
Mark Applied
Hide
Sr. Data Engineer, Enterprise - Slack
San Francisco or Seattle or McLean
$173k-$260k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: The #1 AI CRM driving customer success together.
5+ YOERequires 5+ years in data architecture, modeling, master or metadata management; Python and SQL; relational and dimensional modeling; ETL optimization; Airflow; governance; and U.S. citizenship or lawful permanent residency.
Airflow, Python, SQL, Spark SQL, AWS, Amazon S3, Amazon EMR, Apache Pinot, Hadoop, Flink, Scala, GCP, Azure
3d
Save
Mark Applied
Hide
Sr. Data Engineer, Enterprise - Slack
San Francisco or Seattle or McLean
$173k-$260k/yr OnsiteFull Time
Slack
Slack: Slack is a privately owned, Salesforce-controlled workplace communication platform serving businesses with messaging, automation, integrations, and AI tools.
5+ YOERequires 5+ years in data architecture, modeling, master data, or metadata management; strong Python and SQL; relational modeling, ETL optimization, Airflow, governance, SDLC, Agile, and U.S. citizenship or permanent residency.
Python, SQL, Airflow, Spark SQL, AWS S3, AWS EMR, Apache Pinot, Hadoop, Flink, Scala, AWS, GCP, Azure