mirror of
https://github.com/qdrant/landing_page.git
synced 2026-09-28 15:38:33 +02:00
Update qdrant-landing/content/blog/binary-quantization-openai.md
This commit is contained in:
@@ -68,10 +68,6 @@ We use 100K random samples from the [OpenAI 1M](https://huggingface.co/datasets/
|
||||
|
||||
For each record, we run a parameter sweep over the number of oversampling, rescoring, and search limits. We can then understand the impact of these parameters on search accuracy and efficiency. Our experiment was designed to assess the impact of Binary Quantization under various conditions, based on the following parameters:
|
||||
|
||||
- Oversampling
|
||||
- Rescoring
|
||||
- Search limits
|
||||
|
||||
- **Oversampling**: By oversampling, we can limit the loss of information inherent in quantization. This also helps to preserve the semantic richness of your OpenAI embeddings. We experimented with different oversampling factors, and identified the impact on the accuracy and efficiency of search. Spoiler: higher oversampling factors tend to improve the accuracy of searches. However, they usually require more computational resources.
|
||||
|
||||
- **Rescoring**: Rescoring refines the first results of an initial binary search. This process leverages the original high-dimensional vectors to refine the search results, **always** improving accuracy. We toggled rescoring on and off to measure effectiveness, when combined with Binary Quantization. We also measured the impact on search performance.
|
||||
|
||||
Reference in New Issue
Block a user