mirror of
https://github.com/qdrant/landing_page.git
synced 2026-10-04 10:28:29 +02:00
980 B
980 B
title, description, weight
| title | description | weight |
|---|---|---|
| Pooling Techniques | Reduce the number of vectors per document using row/column pooling and hierarchical token pooling strategies. | 3 |
{{< date >}} Module 3 {{< /date >}}
Pooling Techniques
While quantization reduces the size of each vector, pooling reduces the number of vectors per document. By intelligently combining token embeddings, you can achieve significant memory savings while preserving retrieval quality.
Pooling is particularly effective when combined with quantization for maximum memory efficiency.
Now let's tackle the indexing challenge with MUVERA, enabling fast approximate search for multi-vector representations.