mirror of
https://github.com/qdrant/landing_page.git
synced 2026-10-03 01:48:32 +02:00
Add notes about context window and vector size
This commit is contained in:
@@ -338,4 +338,8 @@ func main() {
|
|||||||
})
|
})
|
||||||
```
|
```
|
||||||
|
|
||||||
Usage examples, specific to each cluster and model can also be found in the Inference tab of the Cluster Detail page in the Qdrant Cloud Console.
|
Usage examples, specific to each cluster and model can also be found in the Inference tab of the Cluster Detail page in the Qdrant Cloud Console.
|
||||||
|
|
||||||
|
Note that, each model has a context window, which is the maximum number of tokens that can be processed by the model in a single request. If the input text exceeds the context window, it will be truncated to fit within the limit. The context window size is displayed in the Inference tab of the Cluster Detail page.
|
||||||
|
|
||||||
|
For dense vector models, you also have to ensure that the vector size configured in the collection matches the output size of the model. If the vector size does not match, the upsert will fail with an error.
|
||||||
|
|||||||
Reference in New Issue
Block a user