Steampunk
Posted 1mo ago

Data Engineer - Databricks

Steampunk
McLean, Virginia, United States
$125k-$160k/yrRemoteFull Time
Responsibilities
  • designing solutions
  • developing pipelines
  • migrating data
Requirements
  • Ability to obtain Public Trust,2+ years data engineering experience,Databricks,SQL,PySpark/Python,AWS experience,ETL/pipeline migration,agile and CI/CD experience
Technical tools mentioned
DatabricksApache SparkDelta LakeT-SQLpgSQLMySQLDatabricks WorkflowsAirflowStep FunctionsAWSS3EC2RDSPySparkPythonJavaC++ScalaApache IcebergInformatica EDCUnity CatalogCollibraAlationPurviewDataZoneCI/CD

Job description

Overview:

In today’s rapidly evolving technology landscape, an organization’s data has never been a more important aspect in achieving mission and business goals. Our data exploitation experts work with our clients to support their mission and business goals by creating and executing a comprehensive data strategy using the best technology and techniques, given the challenge. 

 

At Steampunk, our goal is to build and execute a data strategy for our clients to coordinate data collection and generation, to align the organization and its data assets in support of the mission, and ultimately to realize mission goals with the strongest effectiveness possible. 

 

For our clients, data is a strategic asset. They are looking to become a facts-based, data-driven, customer-focused organization.  To help realize this goal, they are leveraging visual analytics platforms to analyze, visualize, and share information.  At Steampunk you will design and develop solutions to high-impact, complex data problems, working with the best and data practitioners around. Our data exploitation approach is tightly integrated with Human-Centered Design and DevSecOps. 



Contributions:

We are looking for seasoned Data Engineer to work with our team and our clients to develop enterprise grade data platforms, services, and pipelines in Databricks. We are looking for more than just a "Data Engineer", but a technologist with excellent communication and customer service skills and a passion for data and problem solving. 

 

  • Lead and architect migrations of data using Databricks with focus on performance, reliability, and scalability. 
  • Assess and understand ETL jobs, workflows, data marts, BI tools, and reports 
  • Address technical inquiries concerning customization, integration, enterprise architecture and general feature/functionality of data products 
  • Experience working with database/data warehouse/data mart solutions in cloud (Preferably AWS. Alternatively Azure, GCP). 
  • Key must have skill sets – Databricks, SQL, PySpark/Python, AWS 
  • Support an Agile software development lifecycle 
  • You will contribute to the growth of our AI & Data Exploitation Practice! 


Qualifications:

Required:

 

  • Ability to hold a position of public trust with the US government. 
  • 2-4 years industry experience coding commercial software and a passion for solving complex problems.  
  • 2-4 years direct experience in Data Engineering with experience in tools such as: 
  • Big data tools: Databricks, Apache Spark, Delta Lake, etc. 
  • Relational SQL (Preferably T-SQL. Alternatively pgSQL, MySQL). 
  • Data pipeline and workflow management tools: Databricks Workflows, Airflow, Step Functions, etc. 
  • AWS cloud services: Databricks on AWS, S3, EC2, RDS (or Azure equivalents). 
  • Object-oriented/object function scripting languages: PySpark/Python, Java, C++, Scala, etc. 
  • Experience working with Data Lakehouse architecture and Delta Lake/Apache Iceberg 
  • Advanced working SQL knowledge and experience working with relational databases, query authoring and optimization (SQL) as well as working familiarity with a variety of databases. 
  • Experience manipulating, processing, and extracting value from large, disconnected datasets. 
  • Ability to inspect existing data pipelines, discern their purpose and functionality, and re-implement them efficiently in Databricks. 
  • Experience manipulating structured and unstructured data. 
  • Experience architecting data systems (transactional and warehouses). 
  • Experience the SDLC, CI/CD, and operating in dev/test/prod environments. 
  • Experience with data cataloging tools such as Informatica EDC, Unity Catalog, Collibra, Alation, Purview, or DataZone is a plus. 
  • Commitment to data governance. 
  • Experience working in an Agile environment. 
  • Experience supporting project teams of developers and data scientists who build web-based interfaces, dashboards, reports, and analytics/machine learning models 

 



About steampunk:

Steampunk relies on several factors to determine salary, including but not limited to geographic location, contractual requirements, education, knowledge, skills, competencies, and experience. The projected compensation range for this position is $125,000 to $160,000.  The estimate displayed represents a typical annual salary range for this position. Annual salary is just one aspect of Steampunk’s total compensation package for employees. Learn more about additional Steampunk benefits here. 

 

Identity Statement

As part of the application process, you are expected to be on camera during interviews and assessments. We reserve the right to take your picture to verify your identity and prevent fraud.

 

Steampunk is a Change Agent in the Federal contracting industry, bringing new thinking to clients in the Homeland, Federal Civilian, Health and DoD sectors.  Through our Human-Centered delivery methodology, we are fundamentally changing the expectations our Federal clients have for true shared accountability in solving their toughest mission challenges.  As an employee owned company, we focus on investing in our employees to enable them to do the greatest work of their careers – and rewarding them for outstanding contributions to our growth. If you want to learn more about our story, visit http://www.steampunk.com. 

 

We are an equal opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, disability status, protected veteran status, or any other characteristic protected by law. Steampunk participates in the E-Verify program.  

About Steampunk

Provides digital transformation and IT consulting for federal agencies.

Year founded
2019
Employees
660
Organization type
Private
Latest investment
Non-Equity Assistance (2024) — led by AcceliCITY
Subsidiaries
Headquarters
US

Similar jobs

Data Engineer roles near McLean, Virginia
10h
Save
Mark Applied
Hide
Data Engineer
Washington, District of Columbia or Arlington
$93k-$176k/yr RemoteFull Time
Accenture Federal Services
Accenture Federal ServicesNYSE: ACN: Provides technology and consulting services to U.S. federal agencies.
Experience with cloud ETL, data warehousing, Python, SQL, Spark, PySpark, Databricks, Palantir, Snowflake, orchestration, and data quality; must be a US citizen.
Databricks, Spark, AWS, Amazon S3, AWS Glue, AWS Lambda, Amazon Redshift, Google Cloud Dataflow, Azure Data Factory, Google BigQuery, Snowflake, Python, SQL, PySpark, Palantir, ElasticSearch, NiFi, Docker, Kubernetes, Hadoop, ELK stack, Agile, Scrum
11h
Save
Mark Applied
Hide
Staff Data Engineer, Mission Data Products - Clearance Required
Fort Belvoir, Virginia, United States
$128k-$191k/yr OnsiteFull Time
LMI
LMI: Provides management consulting and digital solutions to federal government agencies.
8+ YOEBachelor’s degree or equivalent experience, 8+ years building production data systems, Python and SQL proficiency, enterprise data-platform experience, DoD customer experience, and active Secret clearance.
Python, SQL, Palantir Foundry, Army Vantage, CI/CD, PySpark, DevSecOps
11h
Save
Mark Applied
Hide
Staff Data Engineer, Mission Data Products - Clearance Required
Fort Belvoir, Virginia, United States
$128k-$191k/yr OnsiteFull Time
LMI
LMI: Provides management consulting and digital solutions to government agencies.
8+ YOEBachelor’s degree or equivalent experience, 8+ years building production data systems, Python and SQL expertise, enterprise data-platform experience, and active Secret clearance; U.S. government clearance eligibility required.
Python, SQL, Palantir Foundry, Army Vantage, CI/CD, PySpark, DevSecOps
12h
Save
Mark Applied
Hide
Data Engineer, Sr.
Reston, Virginia, United States
$145k-$225k/yr OnsiteFull Time
Chenega Corporation
Chenega Corporation: Provides professional government, defense, and facility support services.
8+ YOEBachelor’s degree or high school diploma/GED with additional experience; 8+ years relevant experience, including data migration; 5+ years RDBMS and SQL; 5+ years data analytics; active TS/SCI with CI Poly.
SQL, REST APIs, JSON, XML, AWS, Jenkins, Puppet, Chef, Ansible, AWS CloudFormation, Kubernetes, Docker, Agile, RDBMS
13h
Save
Mark Applied
Hide
Data Engineer
Washington, District of Columbia, United States
OnsiteFull Time
VIA
VIA: Provides secure decentralized identity and data privacy software solutions.
3+ YOEBachelor’s degree plus 3 years, associate degree plus 7 years, major certification plus 7 years, or 11 years’ specialized experience; Python, SQL, Spark, ML, data visualization, and secure cloud experience required.
Databricks, Docker, Terraform, Python, Spark, Amazon SageMaker, MLflow, Palantir MSS Workshop, Slate, SQL, PySpark, scikit-learn, TensorFlow, XGBoost, Palantir Foundry, MLOps, AWS, Microsoft Azure
1d
Save
Mark Applied
Hide
Data Engineer
Fairfax, Virginia, United States
$130k-$155k/yr OnsiteFull Time
Prometheus Federal Services
Prometheus Federal Services: Providing strategic consulting and program management for federal health agencies.
5+ YOEBachelor’s degree and 5+ years in data engineering, ETL/ELT, or database administration. Requires SQL, Python, PySpark, Spark, Databricks, Azure, relational databases, Git tools, U.S. work authorization, and ability to obtain public trust.
SQL, Python, PySpark, Spark, Databricks, Azure Data Factory, Azure Synapse, Azure Data Lake Storage, Azure SQL, SQL Server, PostgreSQL, GitHub, GitLab, GitHub Actions, Amazon, Azure
1d
Save
Mark Applied
Hide
Senior Data Engineer (Secret Cleared)
Arlington, Virginia, United States
$71k-$172k/yr OnsiteFull Time
CGI
CGINYSE: GIB: Provides information technology and business consulting services.
4+ YOEBachelor's degree in computer science, engineering, or related field; 4–8 years of data engineering experience; expertise in ETL/ELT, SQL, Python, data modeling, cloud platforms, distributed processing, orchestration, and CI/CD.
SQL, Python, Apache Spark, Databricks, Microsoft Azure Synapse, Microsoft Azure Data Factory, Tableau, Microsoft Power BI, CI/CD, ETL, ELT, Microsoft Azure
1d
Save
Mark Applied
Hide
Data Engineer / Databricks Engineer-Senior
Joint Base Andrews, Maryland, United States
$125k-$178k/yr HybridFull Time
Nationwide IT Services
Nationwide IT Services: Provides IT and management consulting services to federal government agencies.
5+ YOEHands-on Azure Databricks or Spark ETL/ELT experience, advanced SQL and Python, batch and streaming ingestion, Git-based CI/CD, data quality, orchestration, monitoring, security, and DoD training compliance.
Azure Databricks, SQL, Python, Git, Power BI, Palantir Foundry, Microsoft Azure, Advana/WDP, AskSage, MILPDS, DCPDS, AFRISS, CHRIS, CMS, CRIS, GFEBS, DEAMS, REMIS, AROWS, M4S, iEMS, Lockheed Martin Tableau, Spark