Smarty
Posted 2w ago

Data Engineer

Smarty
Orem, Utah, United States
$90k-$120k/yrOnsiteFull Time
Responsibilities
  • designing pipelines
  • maintaining models
  • managing infrastructure
Requirements
  • 3+ years data engineering experience
  • Hands-on dbt, Python, Airbyte
  • Strong AWS data services and SQL skills
  • MySQL and Git/Bitbucket experience
  • Pipeline monitoring and data quality focus
Technical tools mentioned
AirbytePythondbtAmazon S3Amazon AthenaAWS GlueAWS LambdaAWS Step FunctionsEventBridgeAWS ECS/ECRAWS CloudFormationAWS CloudWatchAWS IAMSecrets ManagerPower BIMySQLGit/Bitbucket

Job description

About Smarty


Since its founding in 2011, Smarty has become a recognized industry leader in address data intelligence (exciting, we know). We provide enterprise-grade SaaS solutions worldwide for address validation, geocoding, and data enrichment.


Our culture is all about creating, building, helping, and collaborating with energy and excitement. We want to enjoy our time together by being outward toward others, making appropriate efforts to master our respective crafts, and doing lots of fun things together.

We are looking for a skilled Data Engineer to join our Business Intelligence team at Smarty, a growing SaaS company. In this role, you will be responsible for designing, building, and maintaining the data infrastructure that powers our analytics and reporting capabilities. You will work closely with data analysts and business stakeholders to ensure that clean, reliable, and well-structured data is always available when and where it's needed.


This is an onsite, non-remote position in Orem, UT.

What You'll Do

  • Design, build, and maintain scalable data pipelines and ELT workflows using open-source/self-hosted Airbyte and custom Python
  • Develop and maintain dbt models, tests, and documentation to transform raw source data into reliable fact and dimension tables
  • Manage and optimize our AWS data lakehouse architecture, including S3 storage, Glue data catalog, and Athena query layer
  • Orchestrate automated data workflows using AWS Step Functions and EventBridge schedules
  • Deploy and manage Lambda functions for near-real-time pipelines, reverse ETL, and Slack-based data delivery
  • Manage infrastructure as code using AWS CloudFormation for Lambdas, Step Functions, and Glue resources
  • Maintain and administer Airbyte for data ingestion across sources such as HubSpot, MySQL databases, Google Analytics, and others
  • Write and maintain bespoke Python pipelines for sources and use cases not well-served by off-the-shelf connectors
  • Monitor pipeline health via AWS CloudWatch, respond to failures, and maintain data quality through dbt testing
  • Manage AWS services including ECS/ECR (containerized pipeline execution), Secrets Manager, SNS, and IAM
  • Collaborate with data analysts to support Power BI reporting, including scheduled semantic model refreshes
  • Document data architecture, pipelines, and operational runbooks to support team knowledge and continuity

Who You Are

  • 3+ years of experience in a Data Engineering or similar role
  • Strong hands-on experience with dbt (data build tool) — modeling, incremental strategies, testing, and documentation
  • Proficiency in Python for data pipeline development and automation
  • Strong working knowledge of AWS data services, including:
    • Amazon S3 — storage management, partitioning, lifecycle
    • Amazon Athena — query optimization, DDL management
    • AWS Glue — data catalog, crawlers, DDL management
    • AWS Lambda — event-driven processing and automation
    • AWS Step Functions & EventBridge — pipeline orchestration and scheduling
    • AWS ECS/ECR — running containerized workloads
    • AWS CloudFormation — infrastructure as code
    • AWS CloudWatch — logging and monitoring
    • AWS IAM & Secrets Manager — permissions and credential management
  • Experience managing an Airbyte (or similar ELT tool such as Fivetran) deployment and its source connections
  • Solid SQL skills and comfort with data modeling concepts (staging, intermediate, fact/dimension layers)
  • Familiarity with MySQL or similar relational databases as data sources
  • Experience with version control (Git/Bitbucket or similar) and code review workflows
  • Strong problem-solving skills and a proactive approach to pipeline monitoring and data quality
  • Nice to have:
    • Experience integrating with SaaS APIs such as Salesforce, HubSpot, Google Analytics, or Help Scout
    • Familiarity with Power BI or other BI tools (Tableau, Looker, etc.)
    • Experience with reverse ETL patterns (pushing transformed data back to source systems)
    • Experience working in a SaaS company environment or with product usage/billing data
    • Familiarity with Apache Iceberg or similar table formats (VACUUM/OPTIMIZE operations)
    • AWS certifications (e.g., AWS Certified Data Engineer, Solutions Architect)


Benefits and Perks

  • Competitive compensation (DOE)
    • Range for this role is $90k - $120k
  • 100% paid health, dental, and basic life insurance premiums (including family coverage)
  • Long-term disability insurance
  • Generous PTO that increases with tenure
  • 401(k) with company matching
  • Ongoing training and professional development
  • Adjustable standing desk and modern tools
  • Drinks, snacks, team lunches, and activities
  • In-office chiropractic services
  • Company retreats and trips to genuinely fun places


To apply:

Sounds too good to be true? Apply and find out for yourself.

  • For more information about the company, please visit us at www.smarty.com.


We are an Equal Opportunity Employer and we require all candidates (that receive and accept employment offers) to complete a background check.

About Smarty

Provides location data intelligence and address validation software APIs.

Year founded
2011
Employees
129
Organization type
Private
Headquarters
US

Similar jobs

Data Engineer roles near Orem, Utah
17h
Save
Mark Applied
Hide
Senior Data Engineer, DX
Salt Lake City, Utah, United States
$139k-$219k/yr HybridFull Time
Atlassian
AtlassianNASDAQ: TEAM: Develops software for team collaboration and project management.
Requires strong SQL, Postgres or relational database experience, ETL/ELT pipeline development, large-scale data cleaning and validation, documentation skills, and ability to manage recurring deadlines independently.
SQL, Postgres, JSONB, ETL, ELT, Git, GitHub, GitLab, Bitbucket, CI/CD, Jira, DORA, SPACE, DevEx
21h
Save
Mark Applied
Hide
Senior Data Engineer
Salt Lake City or Marietta or Carol Stream or Phoenix or Cypress or DFW Airport
$118k/yr OnsiteFull Time
R.S. Hughes
R.S. Hughes: Distributes industrial supplies and provides custom material converting services.
3+ YOERequires 3+ years in data or analytics engineering, bachelor's degree in computer science or computer engineering, Azure Synapse and Azure SQL experience, SQL, Python or PySpark, ETL/ELT, dimensional modeling, and Power BI.
Azure Synapse Analytics, Azure Logic Apps, Microsoft Graph, Python, PySpark, SQL, Azure SQL Database, Power BI, Microsoft SQL Server
2w
Save
Mark Applied
Hide
Data Engineer
Woods Cross, Utah, United States
OnsiteFull Time
AutoSavvy: Retailer specializing in branded title used vehicles and inspections.
3+ YOE3+ years data engineering experience, strong SQL and Python, hands-on with Azure data services, Git, ETL/ELT pipelines, able to work independently, valid driver\u0002s license, background check, US work authorization.
Azure SQL, Python, Azure Functions, Container Apps, Azure Blob Storage, Git, Airflow, Azure DevOps, Power BI, Excel, ChatGPT, Claude, Codex, Grok
3w
Save
Mark Applied
Hide
Data Engineer
Draper, Utah, United States
OnsiteFull Time
Sunwest Bank
Sunwest Bank: Entrepreneurial business bank serving businesses and entrepreneurs nationwide.
5+ YOE5+ years data engineering experience building scalable Azure-based data pipelines; strong Python, PySpark, SQL, ETL/ELT, orchestration, and document-processing skills.
Microsoft Azure, Blob Storage, Microsoft Fabric, OneLake, Python, PySpark, SQL, Apache Airflow, Azure Data Factory, Databricks, Apache Spark, Azure Cognitive Services, Azure Synapse, Snowflake, OCR, NLP
3w
Save
Mark Applied
Hide
Senior Data Engineer (Databricks & Cloud Analytics)
Salt Lake City, Utah, United States
$80k-$139k/yr OnsiteFull Time
CGI
CGINYSE: GIB: Provides information technology and business consulting services.
6+ YOE6+ years building enterprise data platforms using Databricks, Spark, Python, SQL and cloud data services; bachelor's degree or equivalent; strong SQL, CI/CD, and streaming experience.
Databricks, Apache Spark (PySpark), Delta Lake, Unity Catalog, Delta Live Tables (DLT), Databricks SQL, MLflow, Azure Data Factory, Azure Data Lake Storage (ADLS Gen2), Apache Kafka, Python, SQL, Scala, Git, Linux, Terraform, Azure DevOps, GitHub Actions, Power BI, Microsoft Fabric, Azure Synapse Analytics
1mo
Save
Mark Applied
Hide
Data Engineer
Salt Lake City, Utah, United States
HybridFull Time
Packsize
Packsize: Provides on-demand automated box manufacturing and packaging systems.
7+ YOEBachelor's or equivalent experience,7+ years data engineering,advanced SQL,intermediate Python,experience building scalable data pipelines,data architecture and dimensional modeling experience.
Python, SQL, DOMO, Databricks, Snowflake, Fabric
1mo
Save
Mark Applied
Hide
Engineering - Client Data Engineering - Data Engineer - Analyst - Bengaluru
Bengaluru or Salt Lake City or Singapore or Sydney
OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Provides investment banking, securities, and wealth management services globally.
1+ YOEBachelor's degree or equivalent, 1+ years in data stewardship/management, strong SQL and data governance knowledge, ability to translate business requirements, communication and analytical skills.
Legend, GLEIF, FactSet, Bloomberg, SQL
1mo
Save
Mark Applied
Hide
Data Engineer
Houston or Golden or Reno or Oakland or Salt Lake City
HybridFull Time
Fervo Energy
Fervo Energy: Generates clean electricity using advanced geothermal drilling technology.
2+ YOE2+ years building and operating production data pipelines with Apache Spark, Python/SQL, Databricks, Azure, and Snowflake; experience with streaming, data modeling, data quality, governance, and CI/CD.
Databricks, Delta Lake, Delta Live Tables, Unity Catalog, Databricks Workflows, Databricks SQL, Apache Spark, PySpark, Spark SQL, Structured Streaming, Kafka, Microsoft Event Hubs, Microsoft Azure Data Factory, Microsoft Azure Data Lake Storage (ADLS), Snowflake, Microsoft Power BI, Spotfire, Python, SQL, Git, Azure DevOps, GitHub Actions, dbt, Terraform, Docker, Airflow, Microsoft Entra ID, Microsoft Key Vault, MQTT, OPC UA, SparkplugB, Canary, Ignition, Snowflake Semantic Views, Databricks Metric Views