Home / Companies / Mixedbread / Blog / October 2025

October 2025 Summaries

4 posts from Mixedbread

Filter
Month: Year:
Post Summaries Back to Blog
Mixedbread Search is now available through the Vercel Marketplace, enabling Vercel users to integrate and manage its multimodal semantic search service within their projects. The platform is intended to improve search systems and AI agents by retrieving relevant information from knowledge bases with low latency and support for varied data formats, helping reduce failures caused by missing context. Users can install the service from the Marketplace, create a Mixedbread account, configure a store and plan, and receive API credentials and store identifiers for connection through Vercel’s interface. Mixedbread also provides documentation and an ecommerce starter template, while inviting feedback during the product’s beta phase to guide ongoing development.
Oct 23, 2025 414 words in the original blog post.
MXBAI introduced the open-source mxbai-edge-colbert-v0 late-interaction retrieval models in 17-million- and 32-million-parameter versions, designed as compact, reproducible baselines for retrieval research and edge deployment. Built on small Ettin/ModernBERT-style encoders, the models were trained through weakly supervised contrastive pretraining, supervised retrieval fine-tuning with hard negatives, and Stella-inspired knowledge distillation before ColBERT-specific optimization and ablation studies. Benchmark results indicate that the 17M version outperforms ColBERTv2 despite using a 48-dimensional projection, while both variants perform strongly on BEIR and LongEmbed, including long-context retrieval. Their low parameter counts, small projection dimensions, Flash Attention 2 support, and built-in unpadding reduce memory and CPU requirements, making them suitable for embedding or reranking documents on modest hardware. Both checkpoints are available through Hugging Face and supported by PyLate, with the developers planning future updates to their edge-focused retrieval models.
Oct 16, 2025 1,762 words in the original blog post.
ColBERT-style late-interaction retrieval models represent queries and documents as token-level vectors and rank documents using MaxSim, which retains only each query token’s highest similarity to a document token. The authors argue that this scoring mechanism creates sparse gradient flow during training, potentially updating only a small fraction of document-token representations, and that its preference for highly discriminative similarity peaks makes the conventional single-layer projection head suboptimal. They propose deeper projection heads with options including intermediate upscaling, residual connections, nonlinear activations, and GLU gating, aiming to preserve and sharpen diverse token representations more effectively. Experiments using modified PyLate training tools across hundreds of controlled model runs found that most projection variants significantly improved retrieval performance, with the strongest configurations averaging roughly two NDCG@10 points, or nearly 4% relative improvement, while adding only a small number of parameters and little practical cost.
Oct 15, 2025 1,855 words in the original blog post.
Mixedbread has launched the public beta of Mixedbread Search, an end-to-end AI-era search API designed to retrieve information from diverse and unstructured sources, including text documents, PDFs, images, audio, video, and legacy code. The platform emphasizes fast, accurate, multilingual, multimodal, and context-aware search for both human and AI users, handling document processing through OCR and parsing so users can upload existing data without configuring search infrastructure. Mixedbread says its in-house research pipeline improves AI-assisted deep-search performance, reporting 16% fewer required LLM calls and 16% higher search accuracy than competing semantic search options on the BrowseComp-Plus benchmark using Gemini 2.5 Flash. The company is offering access through its platform and plans to refine the beta based on user feedback and real-world use cases.
Oct 01, 2025 593 words in the original blog post.