Comtech
Posted 9y ago

Data Engineer with Java & Scala

Comtech
San Jose, California, United States
OnsiteContract
Responsibilities
  • Gathering data
  • Processing data
  • Analyzing data
Requirements
  • 8+ years experience in Java, Python, and Scala
  • Big Data experience
  • Familiarity with Spark and Machine Learning
Technical tools mentioned
JavaPythonScalaSparkMachine Learning

Job description

Company Description:

Comtech is a woman-owned small business founded in 1998 and headquartered in Reston, VA. We offer IT solutions across the disciplines of program/project management, applications development, infrastructure, Cyber security, and enterprise content/data management services. We have developed our methodologies and processes based on the IT Infrastructure Library (ITIL) v.3 Framework across enterprise infrastructure operations. These methodologies and processes are reinforced through our organization’s externally accredited certifications, which include ISO 9001:2008 Quality Management System (QMS), ISO/IEC 20000-1:2011 IT Service Management Systems (SMS, corporate ITIL certification), ISO 27001:2005 Information Security Management System (ISMS), and CMMI-DEV Level 3"

Job Description:

Data Engineer
San Jose, CA
3 +(with possibility of extension)



• Big Data experience


• 8+ years exp in Java, Python and Scala


• With Spark and Machine Learning (3+)


• Data mining, Data analysis


• Data mining, Data analysis


Secondary Skills


• Data mining, Data analysis


Responsibilities


• Gather and process raw data at scale (including writing scripts, web scraping, calling APIs, write SQL queries, etc.).


• Work closely with our engineering team to integrate and build algorithms


• Process unstructured data into a form suitable for analysis – and then do the analysis.


• Support business decisions with ad hoc analysis as needed


• Extract data from a variety of relational databases, manipulate, explore data using quantitative, statistical and visualization tools


• Inform the selection of appropriate modeling techniques to ensure that predictive models are developed using rigorous statistical processes


• Establish and maintain effective processes for validating and updating predictive models


• Analyze, model, and forecast health service utilization patterns/ trends and create capability to model outcomes of what-if scenarios for novel health care delivery models


• Collaborate with internal business, analytics and data strategy partners to improve efficiency and increase applicability of predictive models into the core software products


• Perform statistical analysis to prioritize in order to maximize success


• Identify areas for improvement, communicating action plans


• Perform strategic data analysis and research to support business needs


• Identify opportunities to improve productivity via sophisticated statistical modeling


• Explore data to identify opportunities to improve business results


• Develop understanding of business processes, goals and strategy in order to provide analysis and interpretation to management


• Gain understanding of business needs and necessary analysis where appropriate through internal discussion



Additional Information:

Regards

Amit Singh

703-291-8188

About Comtech

Woman-owned IT services and consulting firm providing enterprise solutions.

Similar jobs

Data Engineer roles near San Jose, California
6h
Save
Mark Applied
Hide
Senior Data Engineer, Bioinformatics, Cheminformatics, Materials
San Francisco, California, United States
$144k-$240k/yr OnsiteFull Time
Lila Sciences
Lila Sciences: Builds AI and autonomous labs for scientific discovery.
2+ YOERequires 2–6 years in data engineering, bioinformatics, cheminformatics, or computational science; strong Python and SQL; ETL, data modeling, statistics, scientific data validation, and workflow orchestration experience.
Python, SQL, Postgres, pandas, NumPy, Flyte, Airflow, Prefect, Dagster, Nextflow, Parquet, Iceberg, DuckDB, Polars, Ibis, NATS, Kafka, LIMS, ELN, XRD, XRF, SEM, TGA, DSC
11h
Save
Mark Applied
Hide
Data Engineer
San Francisco or New York City
HybridFull Time
Beast Industries
Beast Industries: Privately held creator-led holding producing digital entertainment, snacks, software, and consumer brands for global audiences.
3+ YOE3+ years building and operating high-volume production data pipelines, with streaming and batch experience, event instrumentation, schema design, data quality, event-driven architectures, and major cloud data stacks.
AWS, GCP, Databricks, Snowflake
1d
Save
Mark Applied
Hide
Data Engineer, Foundations
San Francisco or Seattle
$157k-$245k/yr HybridFull Time
Superhuman
Superhuman: Private AI productivity platform for people and teams, combining writing, collaborative docs, email, and proactive AI agents.
3+ YOE3+ years operating production environments and data-intensive workflows; proficiency in SQL, Python, Spark, and modern data platforms; expertise in ETL/ELT, orchestration, data modeling, quality, observability, and CI/CD.
SQL, Python, Spark, Databricks, Delta Lake, dbt, Snowflake, Airflow, Dagster, Git, BigQuery, Redshift, Kafka, Flink, Spark Structured Streaming, Iceberg, Hudi, Terraform
1d
Save
Mark Applied
Hide
Data Engineer
Mountain View, California, United States
$220k-$226k/yr HybridFull Time
Google
GoogleNASDAQ: GOOG, GOOGL: Global technology specializing in internet-related services and products.
3+ YOEBachelor’s degree and 5 years of progressive experience, or master’s degree and 3 years. Requires data pipeline design, data modeling, programming, processing stacks, exploratory queries, and BI visualization.
Python, Java, Go, C++, Tableau, Power BI, DataStudio
2d
Save
Mark Applied
Hide
Data Engineer, Expert
Oakland, California, United States
$140k-$238k/yr HybridFull Time
Pacific Gas and Electric Company
Pacific Gas and Electric CompanyNYSE American: PCG-PA: Provides natural gas and electric service.
7+ YOEBA/BS or equivalent experience; 7 years in data engineering or ETL ecosystems; machine learning deployment experience. Cloud platforms, CI/CD, generative AI, and geospatial data experience preferred.
Palantir Foundry, Spark, Informatica, SAP BODS, OBIEE, Generative AI, LLMs, RAG
3d
Save
Mark Applied
Hide
Data Engineer, PXT Central Science
Seattle or Bellevue or Arlington or San Francisco
$132k-$206k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Multinational technology focused on e-commerce and cloud computing.
3+ YOE3+ years of data engineering experience, programming in Python, Java, Scala, or NodeJS, data modeling, warehousing, ETL, AWS, and non-relational database experience required.
AWS, AWS Glue, EMR, Lambda, Redshift, S3, Kinesis, Firehose, IAM, Python, Java, Scala, NodeJS, Hadoop, Hive, Spark
4d
Save
Mark Applied
Hide
Staff Data Engineer
Berkeley, California, United States
$148k-$185k/yr OnsiteFull Time
Form Energy
Form Energy: Developing cost-effective, multi-day iron-air battery storage technology.
7+ YOEBachelor's degree in a relevant technical field and 7+ years of data or software engineering experience. Requires production programming, pipelines, streaming systems, orchestration, AWS, and infrastructure-as-code expertise.
Python, Scala, Java, Go, Apache Spark, PySpark, dbt, Apache Iceberg, Delta Lake, Kafka, Kinesis, Dagster, Prefect, Airflow, AWS, Pulumi, Terraform, Databricks, Snowflake, MQTT, Modbus, OPC-UA, InfluxDB, TimescaleDB, ClickHouse, Druid, REST APIs, webhooks
4d
Save
Mark Applied
Hide
Data Engineer
San Francisco, California, United States
$150k-$250k/yr OnsiteFull Time
Serval
Serval: Private AI-native ITSM automating help desks, access, and workflows for enterprise teams.
5+ YOERequires 5+ years in data engineering and business analytics, expert SQL, PostgreSQL, Python, warehouse and pipeline development, AWS, Terraform, and independent end-to-end project ownership.
SQL, PostgreSQL, Python, Snowflake, Databricks, Fivetran, DBT, AWS, Terraform, Go, gRPC, React, TypeScript, Kubernetes