Home / Companies / Qdrant / Blog / Post Details
Content Deep Dive

Fine-Tuning Sparse Embeddings for E-Commerce Search | Part 2: Training SPLADE on Modal

Blog post from Qdrant

Post Details
Company
Date Published
Author
Thierry Damiba
Word Count
2,047
Company Posts That Month
12
Language
English
Hacker News Points
-
Post removed?
No
Summary

This article, part of a series on fine-tuning sparse embeddings for e-commerce search, focuses on training the SPLADE model using Amazon's ESCI dataset on Modal's serverless GPUs. The ESCI dataset, notable for its graded relevance labels, allows the model to learn nuanced product matches by treating both exact and substitute products as relevant during training. The SPLADE model's performance depends on careful product text formatting, using specific tokens to maintain lexical signals. The training process involves a SparseEncoder built from a DistilBERT base model, utilizing a contrastive loss combined with sparsity regularization to optimize query and product embeddings. The article emphasizes the use of Modal's infrastructure for efficient training, highlighting persistent storage to manage checkpoints and the advantages of detached runs to prevent data loss. It also warns against the pitfalls of replacing transformers with static embeddings, which led to poor results due to the loss of contextual understanding essential for e-commerce queries.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Vector Search 12 3,215 679 175 +33%
AI Model Fine-tuning 4 1,167 231 79 +5%
Serverless 3 1,341 270 110 +29%
LLM 2 7,531 1,250 268 +26%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.