Release 1 10 aggregated (#1005)

* Qdrant 1.10: describe IDF computation (#996)

* describe IDF computation

* add python snippet, fix typo

* docs: Java, Csharp

* add ts and rust examples

* docs: Formatting Java, Csharp indexing.md

* fix language

---------

Co-authored-by: Luis Cossío <luis.cossio@outlook.com>
Co-authored-by: Anush <anushshetty90@gmail.com>
Co-authored-by: davidmyriel <davidmyriel@gmail.com>

* Qdrant 1.10: add vectors page with description of vector types and datatypes (#995)

* add vectors page with description of vector types and datatypes

* docs: csharp vectors.md

* docs: Java vectors.md

* docs: Java, Csharp formatting

* wip: rust and typescript snippets

* Update qdrant-landing/content/documentation/concepts/vectors.md

Co-authored-by: Arnaud Gourlay <arnaud.gourlay@gmail.com>

* Update qdrant-landing/content/documentation/concepts/vectors.md

Co-authored-by: Arnaud Gourlay <arnaud.gourlay@gmail.com>

* add rest of snippets

* add proofread and edit text

---------

Co-authored-by: Anush <anushshetty90@gmail.com>
Co-authored-by: Arnaud Gourlay <arnaud.gourlay@gmail.com>
Co-authored-by: davidmyriel <davidmyriel@gmail.com>

* Qdrant 1.10.0: document S3 snapshots, reorganize storage text (#994)

* Move snapshots storage section down, add local fs and S3 sub sections

* Remove yaml suffix from S3 secrets

* Update S3 example configuration, it didn't match latest release

* Update GitHub Stars (#993)

Co-authored-by: generall <generall@users.noreply.github.com>

* Add documentation for the Query API [v1.10] (#990)

* Add documentation for the Query API

* review suggestions

* upd url

* Add rrf image, fix typos

* easier intro image to fusion

* add python snippets

* docs: java, csharp

* docs: csharp hybrid-queries.md

* docs: Format hybrid-queries.md

* Add Rust client search examples using Query API

* docs: Java hybrid-queries.md

* doc: Helper comment

* Add Rust examples in hybrid queries

* v1.10.0 typescript snippets (#999)

* v1.10.0 typescript snippets

* missed using

* missed comments

* missed docs

* Address @agourlay's review

* fix proofread & edit

---------

Co-authored-by: generall <andrey@vasnetsov.com>
Co-authored-by: Anush <anushshetty90@gmail.com>
Co-authored-by: timvisee <tim@visee.me>
Co-authored-by: Ivan Pleshkov <pleshkov.ivan@gmail.com>
Co-authored-by: davidmyriel <davidmyriel@gmail.com>

* [WIP] Version 1.10. Release Article (#992)

* add release article draft

* add wip links

* fix article linnk

* add more information

* fix link

* add parameters

* add benchmark info

* better issues window screenshot

* fix image and descriptions

* add last bits of info

* update release header

* fix title

* docs: Java, Csharp snippets

* docs: Fixed typos

* Removed first Java, C# snippets for a simpler intro

* Update S3 snapshot storage, link to configuration section

* Update S3 configuration example

* Add Rust examples

* Update new Rust client text

* Mention ColBERT Rust snippet as good client example

* Update qdrant-landing/content/blog/qdrant-1.10.x.md

Co-authored-by: Kacper Łukawski <kacperlukawski@users.noreply.github.com>

* Update qdrant-landing/content/blog/qdrant-1.10.x.md

Co-authored-by: Kacper Łukawski <kacperlukawski@users.noreply.github.com>

* Update qdrant-landing/content/blog/qdrant-1.10.x.md

Co-authored-by: Kacper Łukawski <kacperlukawski@users.noreply.github.com>

* Update qdrant-landing/content/blog/qdrant-1.10.x.md

Co-authored-by: Kacper Łukawski <kacperlukawski@users.noreply.github.com>

* Update qdrant-landing/content/blog/qdrant-1.10.x.md

Co-authored-by: Kacper Łukawski <kacperlukawski@users.noreply.github.com>

* Update qdrant-landing/content/blog/qdrant-1.10.x.md

Co-authored-by: Kacper Łukawski <kacperlukawski@users.noreply.github.com>

* Update qdrant-landing/content/blog/qdrant-1.10.x.md

Co-authored-by: Kacper Łukawski <kacperlukawski@users.noreply.github.com>

* Update qdrant-landing/content/blog/qdrant-1.10.x.md

Co-authored-by: Kacper Łukawski <kacperlukawski@users.noreply.github.com>

* Update qdrant-landing/content/blog/qdrant-1.10.x.md

Co-authored-by: Kacper Łukawski <kacperlukawski@users.noreply.github.com>

* add last few changes

* Fix internal link

* Update Rust documentation links to point to specific version

---------

Co-authored-by: Luis Cossío <luis.cossio@outlook.com>
Co-authored-by: Anush <anushshetty90@gmail.com>
Co-authored-by: timvisee <tim@visee.me>
Co-authored-by: Kacper Łukawski <kacperlukawski@users.noreply.github.com>

* upd links in release blog post

* dont use links with domain

* [draft] bm 42 article (#969)

* bm 42 article draft

* upd the article

* BPE -> wordpiece

* azeret mono as fallback font

* testing other fallback for monospace font

* self hosted fonts

* only proofread & edit

* Update qdrant-landing/content/articles/bm42.md

* last fixes

---------

Co-authored-by: trean <trean.mi@gmail.com>
Co-authored-by: davidmyriel <davidmyriel@gmail.com>

* fix link

---------

Co-authored-by: Luis Cossío <luis.cossio@outlook.com>
Co-authored-by: Anush <anushshetty90@gmail.com>
Co-authored-by: davidmyriel <davidmyriel@gmail.com>
Co-authored-by: Arnaud Gourlay <arnaud.gourlay@gmail.com>
Co-authored-by: Tim Visée <tim@visee.me>
Co-authored-by: generall <generall@users.noreply.github.com>
Co-authored-by: Luis Cossío <luis.cossio@qdrant.com>
Co-authored-by: Ivan Pleshkov <pleshkov.ivan@gmail.com>
Co-authored-by: Kacper Łukawski <kacperlukawski@users.noreply.github.com>
Co-authored-by: trean <trean.mi@gmail.com>
This commit is contained in:
Andrey Vasnetsov
2024-07-01 23:49:08 +02:00
committed by GitHub
co-authored by Luis Cossío Anush davidmyriel Arnaud Gourlay generall timvisee Ivan Pleshkov Kacper Łukawski trean Luis Cossío
parent ad6630b260
commit f5cd7a09fe
36 changed files with 4091 additions and 101 deletions
@@ -98,6 +98,15 @@ storage:
# Where to store snapshots
snapshots_path: ./snapshots
snapshots_config:
# "local" or "s3" - where to store snapshots
snapshots_storage: local
# s3_config:
# bucket: ""
# region: ""
# access_key: ""
# secret_key: ""
# Where to store temporary files
# If null, temporary snapshot are stored in: storage/snapshots_temp/
temp_path: null
@@ -130,10 +139,17 @@ storage:
performance:
# Number of parallel threads used for search operations. If 0 - auto selection.
max_search_threads: 0
# Max total number of threads, which can be used for running optimization processes across all collections.
# Note: Each optimization thread will also use `max_indexing_threads` for index building.
# So total number of threads used for optimization will be `max_optimization_threads * max_indexing_threads`
max_optimization_threads: 1
# Max number of threads (jobs) for running optimizations across all collections, each thread runs one job.
# If 0 - have no limit and choose dynamically to saturate CPU.
# Note: each optimization job will also use `max_indexing_threads` threads by itself for index building.
max_optimization_threads: 0
# CPU budget, how many CPUs (threads) to allocate for an optimization job.
# If 0 - auto selection, keep 1 or more CPUs unallocated depending on CPU size
# If negative - subtract this number of CPUs from the available CPUs.
# If positive - use this exact number of CPUs.
optimizer_cpu_budget: 0
# Prevent DDoS of too many concurrent updates in distributed mode.
# One external update usually triggers multiple internal updates, which breaks internal
@@ -141,6 +157,18 @@ storage:
# If null - auto selection.
update_rate_limit: null
# Limit for number of incoming automatic shard transfers per collection on this node, does not affect user-requested transfers.
# The same value should be used on all nodes in a cluster.
# Default is to allow 1 transfer.
# If null - allow unlimited transfers.
#incoming_shard_transfers_limit: 1
# Limit for number of outgoing automatic shard transfers per collection on this node, does not affect user-requested transfers.
# The same value should be used on all nodes in a cluster.
# Default is to allow 1 transfer.
# If null - allow unlimited transfers.
#outgoing_shard_transfers_limit: 1
optimizers:
# The minimal fraction of deleted vectors in a segment, required to perform segment optimization
deleted_threshold: 0.2
@@ -186,33 +214,76 @@ storage:
# Interval between forced flushes.
flush_interval_sec: 5
# Max number of threads, which can be used for optimization per collection.
# Note: Each optimization thread will also use `max_indexing_threads` for index building.
# So total number of threads used for optimization will be `max_optimization_threads * max_indexing_threads`
# If `max_optimization_threads = 0`, optimization will be disabled.
max_optimization_threads: 1
# Max number of threads (jobs) for running optimizations per shard.
# Note: each optimization job will also use `max_indexing_threads` threads by itself for index building.
# If null - have no limit and choose dynamically to saturate CPU.
# If 0 - no optimization threads, optimizations will be disabled.
max_optimization_threads: null
# This section has the same options as 'optimizers' above. All values specified here will overwrite the collections
# optimizers configs regardless of the config above and the options specified at collection creation.
#optimizers_overwrite:
# deleted_threshold: 0.2
# vacuum_min_vector_number: 1000
# default_segment_number: 0
# max_segment_size_kb: null
# memmap_threshold_kb: null
# indexing_threshold_kb: 20000
# flush_interval_sec: 5
# max_optimization_threads: null
# Default parameters of HNSW Index. Could be overridden for each collection or named vector individually
hnsw_index:
# Number of edges per node in the index graph. Larger the value - more accurate the search, more space required.
m: 16
# Number of neighbours to consider during the index building. Larger the value - more accurate the search, more time required to build index.
ef_construct: 100
# Minimal size (in KiloBytes) of vectors for additional payload-based indexing.
# If payload chunk is smaller than `full_scan_threshold_kb` additional indexing won't be used -
# in this case full-scan search should be preferred by query planner and additional indexing is not required.
# Note: 1Kb = 1 vector of size 256
full_scan_threshold_kb: 10000
# Number of parallel threads used for background index building. If 0 - auto selection.
# Number of parallel threads used for background index building.
# If 0 - automatically select.
# Best to keep between 8 and 16 to prevent likelihood of building broken/inefficient HNSW graphs.
# On small CPUs, less threads are used.
max_indexing_threads: 0
# Store HNSW index on disk. If set to false, index will be stored in RAM. Default: false
on_disk: false
# Custom M param for hnsw graph built for payload index. If not set, default M will be used.
payload_m: null
# Default shard transfer method to use if none is defined.
# If null - don't have a shard transfer preference, choose automatically.
# If stream_records, snapshot or wal_delta - prefer this specific method.
# More info: https://qdrant.tech/documentation/guides/distributed_deployment/#shard-transfer-method
shard_transfer_method: null
# Default parameters for collections
collection:
# Number of replicas of each shard that network tries to maintain
replication_factor: 1
# How many replicas should apply the operation for us to consider it successful
write_consistency_factor: 1
# Default parameters for vectors.
vectors:
# Whether vectors should be stored in memory or on disk.
on_disk: null
# shard_number_per_node: 1
# Default quantization configuration.
# More info: https://qdrant.tech/documentation/guides/quantization
quantization: null
service:
# Maximum size of POST data in a single request in megabytes
max_request_size_mb: 32
@@ -253,7 +324,7 @@ service:
#
# Uncomment to enable.
# api_key: your_secret_api_key_here
# Set an api-key for read-only operations.
# If set, all requests must include a header with the api-key.
# example header: `api-key: <API-KEY>`
@@ -265,6 +336,12 @@ service:
# Uncomment to enable.
# read_only_api_key: your_secret_read_only_api_key_here
# Uncomment to enable JWT Role Based Access Control (RBAC).
# If enabled, you can generate JWT tokens with fine-grained rules for access control.
# Use generated token instead of API key.
#
# jwt_rbac: true
cluster:
# Use `enabled: true` to run Qdrant in distributed deployment mode
enabled: false
@@ -331,4 +408,4 @@ WARN - storage.hnsw_index.m: value 1 invalid, must be from 4 to 10000
```
The server will continue to operate. Any validation errors should be fixed as
soon as possible though to prevent problematic behavior.
soon as possible though to prevent problematic behavior.