Cargill
Posted 1w ago

Data Engineer

Cargill
Atlanta, Georgia, United States
OnsiteFull Time
Responsibilities
  • designing systems
  • building pipelines
  • maintaining systems
Requirements
  • Minimum 2 years relevant experience
  • Design
  • Build and maintain streaming and batch data pipelines
  • Familiarity with cloud platforms
  • Data architectures
  • Ingestion
  • Governance
  • Modeling, and DevOps
  • Strong SQL and programming skills
Technical tools mentioned
KafkaAWS GlueIcebergParquetFlinkSparkSpark UIdbtAirflowPythonJavaScalaSQLAWSGCPAzure

Job description

Title: 
Data Engineer

Job Requisition ID: 
331574
Location: 

Atlanta, Georgia, US United States, 30340




Category: 
Data


Description: 

Cargill is committed to providing food and agricultural solutions to nourish the world in a safe, responsible, and sustainable way. Sitting at the heart of the supply chain, we partner with farmers and customers to source, make and deliver products that are vital for living. 
Our 155,000 team members innovate with purpose, providing customers with life’s essentials so businesses can grow, communities prosper, and consumers live well. With over 160 years of experience as a family company, we look ahead while remaining true to our values. We put people first. We reach higher. We do the right thing—today and for generations to come.

Job Purpose and Impact

The Professional, Data Engineering job designs, builds and maintains moderately complex data systems that enable data analysis and reporting. With limited supervision, this job collaborates to ensure that large sets of data are efficiently processed and made accessible for decision making.

Key Accountabilities

  • DATA & ANALYTICAL SOLUTIONS: Develops moderately complex data products and solutions using advanced data engineering and cloud based technologies, ensuring they are designed and built to be scalable, sustainable and robust.
  • DATA PIPELINES: Maintains and supports the development of streaming and batch data pipelines that facilitate the seamless ingestion of data from various data sources, transform the data into information and move to data stores like data lake, data warehouse and others.
  • DATA SYSTEMS: Reviews existing data systems and architectures to implement the identified areas for improvement and optimization.
  • DATA INFRASTRUCTURE: Helps prepare data infrastructure to support the efficient storage and retrieval of data.
  • DATA FORMATS: Implements appropriate data formats to improve data usability and accessibility across the organization.
  • STAKEHOLDER MANAGEMENT: Partners with multi-functional data and advanced analytic teams to collect requirements and ensure that data solutions meet the functional and non-functional needs of various partners.
  • DATA FRAMEWORKS: Builds moderately complex prototypes to test new concepts and implements data engineering frameworks and architectures to support the improvement of data processing capabilities and advanced analytics initiatives.
  • AUTOMATED DEPLOYMENT PIPELINES: Implements automated deployment pipelines to support improving efficiency of code deployments with fit for purpose governance.
  • DATA MODELING: Performs moderately complex data modeling aligned with the datastore technology to ensure sustainable performance and accessibility.

Qualifications

Minimum requirement of 2 years of relevant work experience. Typically reflects 3 years or more of relevant experience.

 

Preferred Qualifications

  • CLOUD ENVIRONMENTS: Familiarity with major cloud platforms (AWS, GCP, Azure). 
  • DATA ARCHITECTURE: Experience with modern data architectures, including data lakes, data lakehouses, and data hubs, along with related capabilities such as ingestion, governance, modeling, and observability. 
  • DATA INGESTION: Proficiency in data collection, ingestion tools (Kafka, AWS Glue), and storage formats (Iceberg, Parquet). 
  • DATA STREAMING: Knowledge of streaming architectures and tools (Kafka, Flink). 
  • DATA MODELING: Strong background in data transformation and modeling using SQL-based frameworks and orchestration tools (dbt, AWS Glue, Airflow). Experience with modeling concepts like SCD and schema evolution. 
  • DATA TRANSFORMATION: Familiarity with using Spark for data transformation, including streaming, performance tuning, and debugging with Spark UI. 
  • PROGRAMMING: Proficient with programming in Python, Java, Scala, or similar languages. Expert-level proficiency in SQL for data manipulation and optimization. 
  • DEVOPS: Demonstrated experience in DevOps practices, including code management, CI/CD, and deployment strategies. 
  • DATA GOVERNANCE: Understanding of data governance principles, including data quality, privacy, and security considerations for data product development and consumption. 


The business will not sponsor applicants for work visa for this position.

Equal Opportunity Employer, including Disability/Vet.


Nearest Major Market: Atlanta

About Cargill

Produces and distributes food, agricultural, and industrial products globally.

Similar jobs

Data Engineer roles near Atlanta, Georgia
23h
Save
Mark Applied
Hide
Senior Data Engineer
Salt Lake City or Marietta or Carol Stream or Phoenix or Cypress or DFW Airport
$118k/yr OnsiteFull Time
R.S. Hughes
R.S. Hughes: Distributes industrial supplies and provides custom material converting services.
3+ YOERequires a bachelor's degree in computer science or computer engineering, 3+ years in data or analytics engineering, Azure Synapse, Azure SQL, SQL, ELT/ETL, REST APIs, Python or PySpark, dimensional modeling, and Power BI.
Azure Synapse Analytics, Azure Logic Apps, REST, Microsoft Graph APIs, Python, PySpark, SQL, Azure SQL Database, Power BI, Microsoft SQL Server
1d
Save
Mark Applied
Hide
Senior Data Engineer (AWS, Azure, GCP)
Atlanta, Georgia, United States
HybridFull Time
CapTech
CapTech: Provides technology and management consulting services to large enterprises.
5+ YOERequires 5+ years in cloud data engineering, ETL/orchestration, databases, SQL, and programming; experience with distributed systems, cloud platforms, data warehousing, DevOps, and technical leadership.
Amazon Web Services (AWS), Microsoft Azure, Google Cloud Platform (GCP), Azure Data Factory, SSIS, Informatica, Alteryx, Ab Initio, Pentaho, Talend, Matillion, Snowflake, Amazon Redshift, Databricks, PostgreSQL, MySQL, SQL Server, Oracle, Amazon Aurora, Presto, BigQuery, SQL, Python, Java, R, C, C#, C++, Shell, Git, Jenkins, CI/CD, Jira
1d
Save
Mark Applied
Hide
Microsoft Fabric Data Engineer
Athens or Atlanta
OnsiteFull Time
Landmark Properties
Landmark Properties: Develops, builds, and manages student and multifamily residential communities.
3+ YOEBachelor's degree in a relevant technical field and 3+ years in data engineering, business intelligence, analytics development, or related technical work; Microsoft Fabric, Power BI, SQL, data pipelines, and modeling experience.
Microsoft Fabric, Power BI, Microsoft Power Platform, SQL, Python, PySpark, Spark, Microsoft Purview, GitHub, CI/CD, DevOps, DAX, Data Factory, Dataflows, Notebooks, Lakehouse, Warehouse
2d
Save
Mark Applied
Hide
Enterprise Data Engineer
Ridgeland or Alabama or Houston or Memphis or Florida or Atlanta
RemoteFull Time
Trustmark
TrustmarkNASDAQ: TRMK: Provides retail and commercial banking, wealth, and insurance services.
4+ YOEBachelor's degree in data or computer science or equivalent certification, 4 years with modern ETL platforms, database, data warehousing, modeling, SQL, Python, and advanced analytical skills.
IBM Datastage, Informatica, Snowpipe, Azure, AWS, SQL, Python
2d
Save
Mark Applied
Hide
Associate Data Engineer, AWS Hybrid
Atlanta, Georgia, United States
$91k-$119k/yr HybridFull Time
Publicis Groupe
Publicis GroupeEuronext Paris: PUB: Global communications, advertising, and digital transformation holding.
5+ YOERequires 5+ years in Big Data engineering, processing, and analytics; proficiency with Hadoop, Spark, Kafka, Flink, Hive, cloud platforms, programming, ETL, data modeling, SQL, and distributed computing.
Hadoop, Spark, Kafka, Flink, Hive, AWS, GCP, Azure, Java, Scala, Python, SQL, Docker, Kubernetes
2d
Save
Mark Applied
Hide
Associate Data Engineer, AWS Hybrid
Atlanta, Georgia, United States
$91k-$119k/yr HybridFull Time
Publicis Groupe
Publicis GroupeEuronext Paris: PUB: Global advertising and digital transformation agency holding.
5+ YOERequires 5+ years in Big Data engineering, proficiency with Hadoop, Spark, Kafka, Flink, Hive, cloud platforms, Java, Scala, Python, SQL, data modeling, ETL, and distributed computing.
Hadoop, Spark, Kafka, Flink, Hive, AWS, GCP, Azure, Java, Scala, Python, SQL, NoSQL, Docker, Kubernetes
3d
Save
Mark Applied
Hide
Google Senior Data Engineer
Albany or Arlington or Atlanta or Austin or Beaverton or Bentonville or Boston or Carmel or Charlotte or Chicago or Cincinnati or Cleveland or Columbus or Culver City or Denver or Des Moines or Detroit or Hartford or Houston or Irving or Kirkland or Miami or Milwaukee or Minneapolis or Morristown or Mountain View or Nashville or New York City or Oklahoma City or Overland Park or Philadelphia or Pittsburgh or Raleigh or Redmond or Sacramento or San Diego or San Francisco or Scottsdale or Seattle or St. Louis or St. Petersburg or Walnut Creek or United States
$80k-$266k/yr HybridFull Time
Accenture
AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
5+ YOERequires 5+ years in data engineering, analytics, or ML; 4+ years with GCP; 5+ years with SQL and pipelines; 3+ years with Python or AI tools; and a bachelor's degree or equivalent.
Google Cloud Platform (GCP), BigQuery, Looker, Vertex AI, Gemini Foundation Models, Gemini Enterprise, Dataflow, Dataproc, Pub/Sub, Cloud Storage, Looker Studio, Model APIs, Embeddings, Dataplex, IAM, SQL, Python, Git
3d
Save
Mark Applied
Hide
Microsoft Fabric Data Engineer
Atlanta or Dallas
OnsiteFull Time
2020 Companies
2020 Companies: Provides outsourced sales and marketing services for global brands.
2+ YOERequires 2–5 years of data engineering experience, strong SQL, data modeling, Delta Lake, Microsoft Fabric, Power BI semantic model optimization, scalable ingestion, collaboration, communication, and agile experience.
Microsoft Fabric, Power BI, Delta Lake, SQL, GitHub, dbt, Git, CI/CD