DRW
Posted 4mo ago

Data Developer

DRW
Montreal, Quebec, Canada
OnsiteFull Time
Responsibilities
  • design pipelines
  • build pipelines
  • optimize data
Requirements
  • Bachelor’s or Master’s in CS/Data Eng
  • 2-5 years building data systems
  • Experience with RAG architectures
  • Vector DBs
  • Python
  • DAG tools
  • Embedding models
  • Spark/Ray/Dask
  • Docker
  • Data modeling and ETL/ELT
Technical tools mentioned
PythonAirflowDagsterPrefectMilvusChromaDBPineconeWeaviateQdrantDockerApache SparkRayDask

Job description

DRW is a diversified trading firm with over 3 decades of experience bringing sophisticated technology and exceptional people together to operate in markets around the world. We value autonomy and the ability to quickly pivot to capture opportunities, so we operate using our own capital and trading at our own risk.

Headquartered in Chicago with offices throughout the U.S., Canada, Europe, and Asia, we trade a variety of asset classes including Fixed Income, ETFs, Equities, FX, Commodities and Energy across all major global markets. We have also leveraged our expertise and technology to expand into three non-traditional strategies: real estate, venture capital and cryptoassets.

We operate with respect, curiosity and open minds. The people who thrive here share our belief that it’s not just what we do that matters–it's how we do it. DRW is a place of high expectations, integrity, innovation and a willingness to challenge consensus.

We are looking for a Data Developer to join our AI and Multi Asset Systematic Strategies team. This team builds AI and ML-powered tools and solutions that enable teams across the firm and support AI researchers. You'll build data pipelines for RAG systems, optimize embedding workflows, and architect scalable solutions for managing analytical, relational, structured, and unstructured data.

Responsibilities:

  • Design and build data pipelines for RAG systems, including document ingestion, chunking, embedding generation, and vector storage.
  • Build ingestion pipelines for structured and unstructured data sources into a centralized data lake, ensuring data is clean, normalized, and accessible for analytics, research, and AI workloads. 
  • Develop data processing workflows to prepare and optimize datasets for fine-tuning and inference workloads. 
  • Build monitoring and evaluation frameworks to measure retrieval quality, latency, and system performance. 
  • Collaborate with ML engineers to optimize data formats and storage patterns for GPU-accelerated inference. 
  • Implement caching strategies and data versioning systems to support efficient model serving. 
  • Deploy and manage vector databases, embedding services, and data processing pipelines. 
  • Drive initiatives to improve data quality, reduce latency, and enhance the accuracy of retrieval systems. 
  • Continuously learn and stay up-to-date with emerging technologies and best practices in data engineering and AI. 
  • Proactively contribute ideas for new tools, process improvements, and technology adoption that move the team forward. 

Requirements:

  • Bachelor's or Master's degree in Computer Science, Data Engineering, or related field. 
  • 2-5 years building data systems and pipelines in production environments. 
  • Strong experience with RAG architectures, including vector databases (Milvus, ChromaDB, Pinecone, Weaviate, or Qdrant). 
  • Proficiency in Python with experience using DAG-based orchestration platforms (Airflow, Dagster, Prefect, or similar). 
  • Hands-on experience with embedding models and semantic search systems. 
  • Experience with distributed data processing frameworks (Apache Spark, Ray, or Dask). 
  • Understanding of LLM inference optimization techniques and prompt engineering. 
  • Familiarity with Docker, containerization, and orchestration platforms. 
  • Strong grasp of data engineering best practices including data modeling, ETL/ELT patterns, and data quality.

For more information about DRW's processing activities and our use of job applicants' data, please view our Privacy Notice at https://drw.com/privacy-notice.

California residents, please review the California Privacy Notice for information about certain legal rights at https://drw.com/california-privacy-notice.

[#LI-KS1] 

About DRW

Technology-driven principal trading firm operating in global financial markets

Similar jobs

Data Developer roles near Montreal, Quebec
21h
Save
Mark Applied
Hide
Sr Data Engineer
Montreal, Quebec, Canada
OnsiteFull Time
Kruger
Kruger: Manufactures paper products and generates renewable energy from natural resources.
7+ YOEBachelor’s or master’s degree in a technical field and 7+ years of data engineering experience. Requires Python, SQL, Spark/PySpark, scalable pipelines, modern data platforms, Git, CI/CD, DataOps, and French and English fluency.
Microsoft Fabric, SQL, Python, Spark, PySpark, Git, CI/CD, DataOps, Azure Data Factory, Synapse Analytics, Databricks, Snowflake, Delta Lake, Azure Data Explorer (KQL), Azure ML, Great Expectations, MCP servers, SAP ECC/S4, SCADA, PI Historian
23h
Save
Mark Applied
Hide
Développeur(euse), ingénierie de données
Montreal, Quebec, Canada
HybridFull Time
Osedea
Osedea: Developing custom software, AI, and robotics for business clients.
5+ YOE5+ years in a similar role; strong Python and SQL skills; Spark, modern analytics architectures, Azure, communication in French and English, and client-service orientation. Snowflake, BigQuery, dbt, and Microsoft certifications preferred.
Microsoft Fabric, Microsoft Synapse, Microsoft Data Factory, Microsoft Azure Functions, Microsoft Azure Storage, Python, SQL, Spark, Snowflake, BigQuery, dbt, Power BI, CI/CD, APIs
2d
Save
Mark Applied
Hide
Data Engineer
Toronto or Vancouver or Ottawa or Calgary or Halifax or St. John's or Kelowna or Prince George or Montreal or Quebec City
$83k-$93k/yr HybridFull Time
Canadian Cancer Society
Canadian Cancer Society: A national cancer organization funding research, supporting people affected by cancer, and advocating for healthier communities.
3+ YOEBachelor's or master's degree in a relevant computing or data field, 3-5 years of related experience, and proficiency with Microsoft Fabric, Apache Spark, SQL, Python, cloud data platforms, integrations, DevOps, and data governance.
Microsoft Fabric, Apache Spark, SQL, Python, Microsoft Azure, Azure DevOps, Microsoft Purview, Salesforce, Delta Lake, OneLake
6d
Save
Mark Applied
Hide
Conseiller(ère) en ingénierie de données
Montreal, Quebec, Canada
HybridFull Time
onepoint
onepoint: Architect of digital transformation for businesses and public organizations.
15+ YOEBachelor's degree in information technology, business administration, management, or related field; 15 years in IT, 7 years in data engineering, 5 years in BI or analytics, and advanced French proficiency.
Microsoft Power BI, Tableau, Microsoft Power Platform, DAX, Power Query, Microsoft Fabric, Azure DevOps, Power Automate, Power Apps, Dataverse, HL7 FHIR, NDJSON, Lakehouses
6d
Save
Mark Applied
Hide
Conseiller(ère) en ingénierie de données
Montreal, Quebec, Canada
HybridFull Time
Onepoint
Onepoint: International consulting firm specializing in digital transformation.
15+ YOEBachelor's degree in information technology, business administration, management, or related field; 15 years IT experience, including 7 years data engineering and 5 years recent BI, analytics, or data visualization experience.
Microsoft Power BI, Tableau, DAX, Power Query, Microsoft Fabric, Azure DevOps, Power Automate, Power Apps, Dataverse, HL7 FHIR, NDJSON, Scrum, Scrumban, Kanban, DAD, SAFe, Waterfall
2w
Save
Mark Applied
Hide
Ingénieur de données
Montreal, Quebec, Canada
OnsiteFull Time
LGS
LGSNYSE: IBM: Provides IT consulting and digital transformation solutions for businesses.
5+ YOERequires 5–10 years of data engineering experience, a bachelor's degree or equivalent, advanced SQL, Python, ETL/ELT, Azure Data Factory, Databricks, Data Lake Storage, Synapse, and French-English bilingualism.
SQL, Python, ETL, ELT, Microsoft Azure, Azure Data Factory, Azure Databricks, Azure Data Lake Storage, Azure Synapse Analytics, PySpark, Apache Spark, Delta Lake, Microsoft Fabric, Snowflake, Azure DevOps, Git, CI/CD, Infrastructure as Code, Power BI
2w
Save
Mark Applied
Hide
Développeur(se) de données (Expérience client)
Montreal or Toronto
HybridFull Time
Dialogue
Dialogue: Provides virtual healthcare and mental health services to employers.
Experience with Snowflake, dbt, Airflow, Git, CI/CD, Terraform, AWS data services, data modeling, governance, anonymization, and regulated health data.
Snowflake, dbt, Airflow, Snowpipe, EventBridge, Terraform, Git, CI/CD, AWS, S3, Secrets Manager, Pinpoint
2w
Save
Mark Applied
Hide
Data Developer (Client Experience)
Montréal or Toronto
HybridFull Time
Dialogue
Dialogue: Provide virtual healthcare and wellness programs for Canadian employees.
Experience with Snowflake, dbt, Airflow, data modeling, Git, CI/CD, testing, observability, Terraform, AWS data services, data governance, and anonymization in regulated health data environments.
Snowflake, dbt, Airflow, Snowpipe, EventBridge, Terraform, Git, CI/CD, AWS, S3, Secrets Manager, Pinpoint, Coursera