🏢 Offshore Consulting Shops

Omm IT Solutions is an IT consulting and staff augmentation firm that hires contractors to work on client-specific projects and task orders at client sites, such as government agencies.

This company was flagged and excluded from default search results. Proceed with caution.

OI
Posted 1mo ago

Sr Data Engineer

Omm IT Solutions
McLean, Virginia, United States
OnsiteFull Time
Responsibilities
  • designing pipelines
  • developing pipelines
  • optimizing performance
Requirements
  • 6+ years data engineering experience with PySpark
  • Snowflake, AWS
  • Advanced SQL and Python
  • Experience migrating ETL to cloud
  • Building scalable batch/stream pipelines, and optimizing Snowflake and Spark workloads
Technical tools mentioned
PySparkPythonSnowflakeAWSAmazon S3ParquetDeltaSnowSQLSnowpipeStreamsTasksIAMGlueEMRLambdaStep FunctionsCloudWatchRedshiftApache IcebergAirflowGitJenkinsGitHub ActionsTerraformGreat ExpectationsdbtApache KafkaAWS Kinesis

Job description

Job Information

Title: Sr Data Engineer
Job Type: Permanent W-2 Employee / Corp2Corp Contractor
Company: Omm IT Solutions
Country: United States
Location: Mc Lean
Published: 07/13/2026
Start Date: 07/15/2026 12:00 AM
Compensation: Negotiable
Industry: IT Services
Work Authorization: USC/GC/H1B/H4-EAD
Job Opening ID: Omm2971J
State/Province: Virginia
City: Mc Lean
Zip/Postal Code: 22101




PLEASE NOTE:
  • IT IS 100 % ON SITE POSITION IN Mc lean VA
KEY REQUIRED SKILLS:
  • PySpark & Python for data pipeline development, Snowflake & AWS
DESCRIPITION:

We are seeking a hands-on, delivery-focused Senior Data Engineer to help build and scale our cloud data platform. In this role, you will design and develop modern data pipelines using PySpark, Snowflake, and AWS to optimize cloud data workloads. The ideal candidate combines strong engineering fundamentals with cloud-native data expertise and is capable of translating complex business needs into robust, performant, and well-documented data solutions. Experience within Fannie Mae, Freddie Mac, or equivalent GSE/mortgage enterprise environments is highly valued.

RESPONSIBILITES:
  • Scalable Architecture: Design and build scalable batch and streaming data pipelines using PySpark for large-scale data processing.

  • Modernization: Migrate legacy, on-premises ETL workloads (e.g., IBM DataStage, Informatica) to high-performing PySpark and Snowflake cloud pipelines.
  • Data Transformation: Write production-grade PySpark code to read from Amazon S3 (Parquet/Delta files), execute complex transformations, and process massive datasets efficiently.

  • Deduplication: Design and implement robust deduplication strategies for high-volume datasets using PySpark.
    Platform Engineering: Build and manage Snowflake warehouses, schemas, and data models optimized for enterprise analytics and business intelligence reporting.

  • Iceberg Tables: Design and implement Apache Iceberg tables in Snowflake to support open lakehouse architectures and data interoperability.
    Incremental Processing: Build and maintain Snowflake Dynamic Tables and Materialized Views to enable near real-time analytics and query acceleration.

  • PySpark Tuning: Optimize distributed Spark jobs by leveraging partitioning, caching, broadcast joins, and shuffle optimization.
    Snowflake Optimization: Tune Snowflake workloads using clustering keys, micro-partition pruning, query profiling, precise warehouse sizing, and strategic result caching.

  • Cost Management: Continuously monitor and optimize Spark jobs, Snowflake queries, and AWS infrastructure to balance speed and cloud expenditure.

  • Data Quality: Implement data validation, lineage tracking, and monitoring solutions across all pipeline stages to ensure high data integrity.
    Cross-Functional Collaboration: Partner closely with data architects, business analysts, Technical Program Managers (TPMs), and corporate stakeholders to deliver dependable data products.

  • Technical Documentation: Author comprehensive technical designs, data schemas, and operational runbooks to ensure every pipeline is maintainable and audit-ready.




Requirements

REQUIRED QUALIFICATION:
  • Experience: 6+ years of hands-on data engineering experience in large-scale enterprise environments.

  • PySpark Expertise: Deep proficiency in building distributed data processing pipelines, handling S3 Parquet/Delta files, and implementing complex transformations and deduplication logic.

  • Snowflake Proficiency: Strong hands-on experience with SnowSQL, Snowpipe, Streams, Tasks, and Role-Based Access Control (RBAC). Proven track record establishing Iceberg tables, Dynamic Tables, and Materialized Views.

  • AWS Cloud Ecosystem: Robust working knowledge of AWS services, including S3, Glue, EMR, Lambda, IAM, Step Functions, CloudWatch, and Redshift.

  • Advanced SQL & Python: Mastery of advanced SQL techniques (window functions, CTEs, complex joins) alongside strong Python programming skills for automation, scripting, and orchestration utilities.

  • Orchestration & Architecture: Solid understanding of data warehousing, ELT/ETL patterns, data lakes, and lakehouse architectures using tools like Airflow or AWS Step Functions.

  • Communication: Strong verbal and written communication skills with the ability to articulate technical decisions clearly to both technical peers and business leaders.

PREFERRED QUALIFICATION:
  • Industry Experience: Prior experience working within heavily regulated environments such as financial services, mortgage banking, or GSE programs (Fannie Mae / Freddie Mac).

  • ETL Migration: Hands-on experience with legacy ETL frameworks (e.g., IBM DataStage) to support modernization initiatives.

  • DevOps & CI/CD: Familiarity with continuous integration and continuous deployment pipelines for data infrastructure (Git, Jenkins, GitHub Actions, Terraform).

  • Data Quality Frameworks: Exposure to automated data quality and validation frameworks (e.g., Great Expectations, dbt testing suites).

  • Streaming Analytics: Knowledge of real-time streaming platforms like Apache Kafka or AWS Kinesis.

  • Professional Certifications: AWS Certified Data Analytics, AWS Certified Solutions Architect, or SnowPro Core/Advanced certifications.




Similar jobs

Data Engineer roles near McLean, Virginia
2h
Save
Mark Applied
Hide
AL Lead Data Engineer
Herndon, Virginia, United States
$135k-$140k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
10+ YOERequires 10+ years of experience, a bachelor's degree in computer science, advanced Azure Databricks, Azure Data Factory, PySpark, Python, and Git expertise, plus identity pipeline and AI automation experience.
Python, Azure Databricks, Azure Data Factory, PySpark, Git, Databricks Genie, GitHub Copilot, Microsoft 365 Copilot
4h
Save
Mark Applied
Hide
Data Engineer
Alexandria, Virginia, United States
$110k-$115k/yr HybridFull Time
Five Guys
Five Guys: Fast-food restaurant chain specializing in fresh burgers and fries.
3+ YOEBachelor's degree and 3+ years of data engineering experience required, with Microsoft Azure, Python, PySpark, SQL, Databricks, and data pipeline expertise; Power BI and certifications preferred.
Microsoft Azure, Azure Data Factory, Azure Functions, Logic Apps, Databricks, Python, PySpark, Azure Data Lake Storage Gen2, Microsoft Office Suite, Power BI, DAX, Microsoft Power BI Service, Unity Catalog, SQL, REST APIs, CSV, XML
7h
Save
Mark Applied
Hide
Data Engineer
Reston, Virginia, United States
$138k-$212k/yr OnsiteFull Time
Federal Home Loan Banks Office of Finance
Federal Home Loan Banks Office of Finance: Sues consolidated debt securities for the Federal Home Loan Banks.
5+ YOEBachelor’s degree in a quantitative field and 5–7 years of data engineering experience, including production pipelines, data analysis, ETL/ELT, and analytical data platforms. U.S. work eligibility required.
Python, pandas, PySpark, SQLAlchemy, SQL, Bash, Azure Data Factory, Azure Synapse, Azure Data Lake Storage, PostgreSQL, SAP ASE, Azure Synapse Analytics, Delta Live Tables, Apache Spark, Power BI, SAP BusinessObjects, GitHub Enterprise, CI/CD, Microsoft Purview, Datadog, Grafana, Prometheus
8h
Save
Mark Applied
Hide
Data Engineer
United States or Reston
RemoteFull Time
Resonate
Resonate: AI-powered platform providing consumer intelligence and marketing data analytics.
5+ YOERequires 5+ years in software or data engineering, 3+ years with Spark and Scala, multi-terabyte or petabyte Spark tuning, relational databases, cloud big data stacks, testing, architecture, and production operations.
Apache Spark, Scala, Amazon Web Services (AWS), Amazon EMR, Amazon S3, Snowflake, Grafana, Apache Kafka, Hadoop, Elastic Stack, Docker, Amazon Lambda
13h
Save
Mark Applied
Hide
Data Engineer, Lead
Herndon, Virginia, United States
$99k-$225k/yr OnsiteFull Time
Booz Allen Hamilton
Booz Allen HamiltonNYSE: BAH: Consulting and technology services for government and commercial clients
7+ YOERequires 7+ years leading enterprise software and complex data applications, 5+ years with ETL and structured/unstructured data, TS/SCI with polygraph, bachelor's degree, and ability to obtain an approved security certification.
IoT, Python, Java, AWS, Docker, Kubernetes, Helm, CI/CD, Apache NiFi, Kafka, Linux, UNIX, MongoDB, Microsoft?
1d
Save
Mark Applied
Hide
ITECH-FM Data Engineer
Washington or North America
OnsiteFull Time
Global C2 Integration Technologies
Global C2 Integration Technologies: Provides AI and engineering solutions for defense missions.
10+ YOEMore than 10 years of relevant experience, a master's degree, senior data engineering expertise, knowledge management, AI/ML implementation, software development, and ability to work independently and oversee junior staff.
Python, Tableau, SQL, Java, C++, Microsoft SharePoint, Microsoft Visual Studio, Microsoft PowerPoint
1d
Save
Mark Applied
Hide
Data Engineer
McLean or Alexandria or Aurora or Orlando
OnsiteFull Time
Knight Federal Solutions
Knight Federal Solutions: Provides defense, intelligence, and IT services to government agencies.
3+ YOEBachelor's degree in computer science, information technology, or related field; 3+ years of data engineering experience; TS SCI with CI Polygraph; SQL, database, pipeline, ETL, cloud, and programming expertise.
AWS, Azure, Google Cloud, Python, Java, Scala, SQL, ETL
1d
Save
Mark Applied
Hide
Data Engineer, Lead
Herndon, Virginia, United States
$99k-$225k/yr OnsiteFull Time
Booz Allen Hamilton
Booz Allen HamiltonNYSE: BAH: Provides technology and management consulting services to diverse organizations.
7+ YOERequires 7+ years leading enterprise software and distributed data applications, 5+ years with ETL and structured/unstructured data, TS/SCI clearance with polygraph, bachelor's degree, and ability to obtain an approved security certification.
IoT, Python, Java, AWS, Docker, Kubernetes, Helm, Apache NiFi, Kafka, Linux, UNIX, MongoDB, Microsoft Internet Explorer, Google Chrome, Mozilla Firefox, Microsoft Edge, Apple Safari, Opera Browser