Episode

Multi-Vector Search with Amélie Chatelain and Antoine Chaffin - Weaviate Podcast #134!

Podcast
Weaviate Podcast
Published
Mar 23, 2026
Duration seconds
4873
Processing state
not_requested
Canonical source
https://podcasters.spotify.com/pod/show/weaviate/episodes/Multi-Vector-Search-with-Amlie-Chatelain-and-Antoine-Chaffin---Weaviate-Podcast-134-e3gq51u
Audio
https://anchor.fm/s/cffc3468/podcast/play/117297662/https%3A%2F%2Fd3ctxlq1ktw2nl.cloudfront.net%2Fstaging%2F2026-2-22%2Faad52788-4ad2-7dbb-f323-7ded21631204.mp3
JSON
/v1/public/podcasts/weaviate-podcast-6288219/episodes/multi-vector-search-with-am-lie-chatelain-and-antoine-chaffin-weaviate-podcast-134
Markdown
/podcast/weaviate-podcast-6288219/multi-vector-search-with-am-lie-chatelain-and-antoine-chaffin-weaviate-podcast-134.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/weaviate-podcast-6288219/episodes/multi-vector-search-with-am-lie-chatelain-and-antoine-chaffin-weaviate-podcast-134/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/weaviate-podcast-6288219/multi-vector-search-with-am-lie-chatelain-and-antoine-chaffin-weaviate-podcast-134.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Amélie Chatelain and Antoine Chaffin from LightOn are leading the way in the next generation of search powered by Multi-Vector representations and Late Interaction. The podcast begins with what motivates them to work on Multi-Vector Search, continuing to discuss particular details such as the combination between lexical and semantic search, as well as bi-encoder speed with cross encoder accuracy. This discussion continues to present insights about training multi-vector models and how they differ from their single-vector predecessors. The conversation continues into particular successes of Late Interaction such as code, reasoning-intensive, and multimodal retrieval. Agents are great at searching with grep, but they are even better with ColGrep! Reasoning-Intensive Retrieval is a step change in how we think about search systems, beautifully enabled by both Late Interaction models and Agentic Search. Further, Multimodal Search, such as matching text with videos, is seeing massive benefits from Multi-Vector representations. The podcast continues to dive into the cost of MaxSim and how efficient methods such as MUVERA and PLAID can help. The podcast concludes with a presentation of their recent work on ColBERT-Zero, pre-training with Late Interaction instead of Single-Vector Dense Embedding models. LightOn are also the developers of PyLate, the world's leading open-source library for training these kinds of models.Chapters0:00 An Introduction to Multi-Vector Search6:00 Multi- vs. Single-Vector8:55 Comparison with Cross Encoders15:55 ColGrep for Coding Agents30:34 Reasoning-Intensive Retrieval42:02 Multimodal Multi-Vector48:34 The Cost of Multi-Vector53:26 MUVERA and PLAID1:06:18 ColBERT-Zero and PyLate1:08:35 ColBERT-Zero and PyLate