mirror of
https://github.com/qdrant/landing_page.git
synced 2026-09-28 07:28:30 +02:00
* docs(fastembed.md): add throughput comparison chart and mention quantized models in "Under the Hood" section
This commit is contained in:
@@ -97,6 +97,8 @@ FastEmbed is built for inference speed, without sacrificing (too much) performan
|
||||
|
||||
We use `BAAI/bge-small-en-v1.5` as our DefaultEmbedding, hence we've chosen that for comparison:
|
||||
|
||||

|
||||
|
||||
## Under the Hood
|
||||
|
||||
**Quantized Models**: We quantize the models for CPU (and Mac Metal) – giving you the best buck for your compute model. Our default model is so small, you can run this in AWS Lambda if you’d like!
|
||||
|
||||
Reference in New Issue
Block a user