EPAM Systems
Posted 2w ago

Data Software Engineer

EPAM Systems
Colombia
RemoteFull Time
Responsibilities
  • constructing applications
  • optimizing workflows
  • reviewing code
Requirements
  • 2+ years in data software engineering with Big Data tools
  • Proficiency in Python/Java/Scala/Kotlin
  • SQL and NoSQL
  • Spark/PySpark
  • AWS (Glue
  • Redshift), CI/CD, and strong English communication
Technical tools mentioned
PythonJavaScalaKotlinSQLNoSQLApache SparkPySparkAmazon Web Services (AWS)GlueRedshiftCI/CDDockerKubernetesYarnDatabricksHadoopHiveFlink

Job description

We are seeking a highly skilled and experienced Data Software Engineer to join our team and contribute to building data-centric applications.

You will utilize advanced Big Data tools, cloud technologies, and collaborate across teams to develop innovative solutions addressing complex business needs.

Responsibilities

  • Construct and extend data software applications leveraged by Data Integration Engineers
  • Deliver analytical solutions with Spark, PySpark, NoSQL, and other Big Data technologies
  • Optimize data workflows by integrating AWS cloud services and features
  • Work with product and engineering stakeholders to collect insights and enable decision-making
  • Coordinate with architects, technical leads, and other teams to keep solutions aligned
  • Analyze business needs and existing technical environments to produce suitable implementations
  • Review code to uphold best practices and ensure consistently high-quality output
  • Verify solutions through functional, technical, and performance testing
  • Create clear project documentation for future development and reference
  • Consult with clients to understand objectives and provide technical recommendations

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Software Engineering, or a comparable field
  • 2+ years of experience in Data Software Engineering using Big Data technologies
  • Strong foundation in data engineering topics including data management, storage, visualization, operations, and security
  • Thorough understanding of ingestion pipelines, Data Warehousing, and Data Lakes
  • Proficiency in at least one of Python, Java, Scala, or Kotlin
  • Demonstrated expertise with SQL plus NoSQL databases
  • Hands-on skills with Spark and PySpark within modern Big Data ecosystems
  • Background delivering solutions on AWS, including Glue and RedShift
  • Understanding of CI/CD workflows for integration and deployment
  • Working familiarity with Docker, Kubernetes, and Yarn in containerized environments
  • Capability with Databricks for advanced analytics and engineering tasks
  • Strong communication skills in English, meeting at least a B2 proficiency level

Nice to have

  • Experience with Hadoop, Hive, and Flink or other Big Data tools
  • Familiarity with SDLC methodologies, particularly Agile
  • Skills in executing and overseeing SDLC implementation end to end

We offer/Benefits

  • International projects with top brands
  • Work with global teams of highly skilled, diverse peers
  • Healthcare benefits
  • Employee financial programs
  • Paid time off and sick leave
  • Upskilling, reskilling and certification courses
  • Unlimited access to the LinkedIn Learning library and 22,000+ courses
  • Global career opportunities
  • Volunteer and community involvement opportunities
  • EPAM Employee Groups
  • Award-winning culture recognized by Glassdoor, Newsweek and LinkedIn

EPAM is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, age, sexual orientation, gender identity or expression, disability, protected veteran status, or any other characteristic protected by applicable law.

About EPAM Systems

Provides global digital platform engineering and software development services.

Similar jobs

Data Software Engineer roles
1mo
Save
Mark Applied
Hide
Senior Software Developer, AI Data Engineer
Bogotá or Medellín
RemoteFull Time
Caseware
Caseware: AI-powered audit and financial reporting software platform.
Experience building scalable AI data pipelines, vector/graph retrieval systems, distributed systems and observability; production AWS experience; infrastructure-as-code (CDK/CloudFormation/Terraform); strong software engineering and communication skills.
TypeScript, NestJS, Python, AWS EKS, AWS Lambda, AWS Bedrock, AWS AgentCore, AWS OpenSearch Serverless, AWS S3 Vectors, AWS Knowledge Bases, AWS Textract, DynamoDB, S3, LangFuse, LangSmith, GitHub Copilot, GitHub, GitHub Actions, Nx, Jira, Confluence, Microsoft Teams, Outlook, CDK, CloudFormation, Terraform