Commit Graph
2224 Commits
Author SHA1 Message Date
Dylan Couzon 7a4129be1e add third guide + update indices 2026-04-23 20:23:10 -04:00
Dylan Couzon 54ecb5b63c turn fundamentals into blog 2026-04-23 20:22:55 -04:00
Dylan Couzon c21062782b clean up golden set 2026-04-23 20:22:29 -04:00
Dylan Couzon e3e0516530 clean up Retrieval quality 2026-04-23 20:22:19 -04:00
Dylan Couzon 1fbc73c4b2 Merge branch 'master' into retrival-quality-guide-improvement 2026-04-23 12:24:29 -04:00
Anush 9dbd5a10f9 docs: Use gemini-embedding-2 in gemini.md (#2299) 2026-04-23 21:42:58 +05:30
Dylan Couzon 340c6528ca Clean up, tone 2026-04-23 11:06:18 -04:00
Dylan CouzonandClaude Opus 4.7 fb1672c102 retrieval-quality-fundamentals: tone pass
- Soften "how teams bridge this gap" to "common patterns for
  bridging this gap" so we don't imply we harvested real client
  pipelines for this writeup
- Reframe the layer-2/3 diagnostic and the "Isolate the component
  under test" bullet so they name multiple downstream consumer
  types (LLM generator, ranker, UI) rather than assuming RAG
- Add a one-line caveat that A/B design for RAG and agentic
  systems is still evolving to the Proxy KPIs paragraph
- Simplify the recall@k / precision@k equivalence note and link
  ann-benchmarks.com as the citation for community convention

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 22:39:03 -04:00
Dylan CouzonandClaude Opus 4.7 e2df842fb5 retrieval-quality-fundamentals: structure & scope pass
- Add "why measure retrieval quality" lead-in paragraph
- Expand ANN on first use; flag sparse vectors as out of scope
  for layer 1 (they use exact matching)
- Add LLM-as-judge to layer-2 ground-truth options
- Break the Tooling bullet into per-layer recommendations: Web UI
  for L1, ranx for L2, Ragas/Phoenix/DeepEval for L3
- Reorder Quality Metrics so layer 1 (ANN recall formula + exact
  kNN equivalence) comes before the generic layer-2 relevance
  metrics; trim a redundant sentence
- Add end-to-end answer quality as a distinct third layer in the
  prose intro so it matches the ladder table's four rows
- Standardize vocabulary on "layer" (was mixing "level" in the
  intro with "layer" everywhere else); update the section anchor
  to #connecting-the-layers-in-practice in this file and the two
  cross-linking tutorials

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 22:31:35 -04:00
Dylan Couzon 5748d6db8c update for moved URL 2026-04-22 22:17:07 -04:00
Dylan Couzon 92af150b00 retrieval-quality: rename to "Measuring ANN Precision"
- Update title, H1, and Time in retrieval-quality.md
    (30 min -> 15 min reflects the pivot to Web UI)
  - Rename references in tutorials-lp-overview.md and the headless
    tutorial index; swap the pill from Python to Web UI to reflect
    the new primary flow
  - Replace "ANN recall" with "ANN precision" in Fundamentals
    (4 places: intro, comparison note, ladder table, cross-link)
    and in the golden-set tutorial's layer-1 cross-reference
  - Filename kept as retrieval-quality.md so existing URLs and
    aliases still work
2026-04-22 17:13:27 -04:00
Dylan Couzon 175f92ddbe change recall back to precision@k
The Search Quality tab in the Web UI reports precision@k, not
  recall@k, so the whole tutorial now uses precision@k as the metric
  label: "ANN recall" -> "ANN precision" in the anchor, section
  headers, Python helper (avg_precision_at_k), and prose. Kept a
  one-line bridge note that ANN-benchmarks terminology calls this
  recall@k, since both searches return exactly k items.
2026-04-22 17:02:45 -04:00
Dylan Couzon 921b645e70 retrieval-quality: pivot to Web UI, add CI skeleton, trim setup
- Drop the dataset-setup walkthrough (HF loading, collection
    create, upload, wait-for-green). Readers at this phase already
    have a collection.
  - Replace the Python evaluation block with a "Measure ANN Recall
    with the Web UI" section built around the Search Quality tab.
    Default run is one-click (sample size 10); HNSW tuning uses the
    tab's advanced mode instead of update_collection. Three
    screenshot placeholders at
    /documentation/tutorials/retrieval-quality/*.png.
  - Reflect that the tab reports precision@k; note the recall@k
    equivalence already spelled out in the ANN Recall section.
  - Keep Python but move it to an "Automate in CI" section with a
    reusable skeleton function.
  - Collapse the standalone "Embeddings Quality" section into a
    one-sentence MTEB pointer inside ANN Recall.
  - Rewrite Wrapping Up to match the new scope.
  - Link the HNSW tuning section to Optimize Performance for the
    full parameter reference.
2026-04-22 16:34:16 -04:00
Dylan CouzonandClaude Opus 4.7 e5d9101227 retrieval-quality: tone pass on intro, section rename, drop RAG link
- Rewrite the intro so the ANN algorithm reads as one of several
  levers shaping retrieval quality (alongside the embedding model,
  retrieval strategy, filtering, reranking) rather than the only
  factor beyond embeddings. Addresses mrscoopers on the reductive
  "embeddings + ANN" framing.
- Rename the "Retrieval Quality" section to "ANN Recall" and
  rewrite its opening paragraph to match; ANN approximation quality
  isn't the same as retrieval quality broadly.
- Drop the RAG-evaluation-guide link from the three places it
  appeared in this file. This tutorial isn't RAG-specific.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 15:26:29 -04:00
Dylan CouzonandClaude Opus 4.7 fd095af85a retrieval-quality: layer-1 anchor and title-case headers
- Add a layer-1 anchor sentence under the Time/Level table
  pointing at the evaluation ladder in Fundamentals, mirroring
  the golden-set tutorial's anchor. Addresses abdonpijpelink's
  ask for a levels-table link and a "this tutorial focuses on
  level 1" framing.
- Title-case the six H2 headers for consistency across the
  tutorials-search-engineering set.

Deferred: streaming the 60K training items into upload_points
instead of materializing as a list (abdonpijpelink line 69) —
pending manager confirmation before proceeding.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 15:06:47 -04:00
Dylan Couzon 81883cbecc golden set: framing 2026-04-22 14:31:10 -04:00
Dylan CouzonandClaude Opus 4.7 a47c57ebdc golden set: use ranx for metrics, reframe Anthropic example
- Switch the metrics section from three hand-rolled functions
  (recall@k, MRR, NDCG@k) and a bespoke evaluate() loop to a
  single ranx-based example. Handles binary and graded labels
  in one call (mrscoopers, line 59).
- Reframe the Anthropic synthetic-generation snippet as a
  minimal prompt shape, point at Ragas for readers who want a
  maintained testset generator, and fix a bug (Anthropic() was
  called without importing the class; switched to
  anthropic.Anthropic()). Addresses abdonpijpelink line 53 and
  mrscoopers line 31.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 13:52:23 -04:00
Dylan CouzonandClaude Opus 4.7 acc85e65ca golden set: anchor intro at layer 2, link layer 1 to Web UI
- Open the tutorial by anchoring it at layer 2 of the evaluation
  ladder and linking to the levels table in retrieval-quality-
  fundamentals (abdonpijpelink, line 14).
- Point layer-1 readers (ANN recall vs exact kNN) at the Search
  Quality tab in the Qdrant Web UI instead of the retrieval-quality
  tutorial (mrscoopers, line 13).
- Trim the duplicate layer-1 pointer at the end of "Using the
  Golden Set" (mrscoopers, line 113).
- Open cross-tutorial links in a new tab to match the repo
  convention.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 10:56:34 -04:00
Dylan Couzon 6eba7e2ceb docs(golden-set): anchor intro at layer 2 2026-04-22 10:30:25 -04:00
Dylan CouzonandAbdon Pijpelink 4d99153c90 make Anthropic API key explicit
Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-04-22 09:17:40 -04:00
Abdon Pijpelink 434a451c2b Update Hybrid Search with Reranking tutorial (#2274)
* Update for Cloud Inference and data ingestion

* Fix link

* Review feedback

* Make code snippets testable

* Add C# code snippets

* Add Go code snippets

* Add Java code snippets

* Add Rust code snippets

* Add TS code snippets

* Move CSV streaming/parsing to separate function
2026-04-22 11:03:05 +02:00
Abdon PijpelinkandClaude Sonnet 4.6 b93c8379e2 Use section _index.md titles in docs breadcrumbs instead of humanized URL slugs (#2290)
Co-authored-by: Claude Sonnet 4.6 <noreply@anthropic.com>
2026-04-22 09:16:27 +02:00
Dylan Couzon 1a9821cc95 golden set: reduce verbosity 2026-04-21 12:06:48 -04:00
Dylan Couzon bb8cac5a6e update completion times 2026-04-21 12:00:10 -04:00
Abdon Pijpelink b23b94b3c5 "Things to Check Before Taking Qdrant into Production" guide (#2219)
* Update Qdrant logo in tradeoff image

* Add production checklist

* Add payload indexing

* Review feedback

* Review feedback

* Small edits

* Fix links
2026-04-21 14:36:33 +02:00
Abdon PijpelinkandTim Visée ca7aaa7499 Tweak prevent unoptimized docs (#2284)
* Add stronger wording about wait=true

* Position indexed_only and prevent_unoptimized as alternatives

* Review feedback

* Apply suggestions from code review

Co-authored-by: Tim Visée <tim+github@visee.me>

* Small edit

---------

Co-authored-by: Tim Visée <tim+github@visee.me>
2026-04-21 14:10:27 +02:00
Dylan Couzon 8cdc9acd14 golden dataset - improve tutorial quality 2026-04-20 12:42:53 -04:00
Dylan Couzon 02710a1f9b Add using golden set instructions 2026-04-20 12:41:21 -04:00
Dylan Couzon 5503b005b6 improve wording 2026-04-20 12:32:33 -04:00
Dylan Couzon d6f06d6b60 clean up tables 2026-04-20 12:22:45 -04:00
Dylan Couzon e9e0e4bddd Syntax 2026-04-20 12:21:05 -04:00
kanungleandAbdon Pijpelink 544708f293 Restructure Docs - Stage 4a (#2280)
* create Develop and Deploy tabs; move Operations; re-weight pages

* move capacity planning page; create section dropdown content

* added aliases to frontmatter

* update link references to new canonical links; maintain anchoring

* address remaining link issues and errors

* fix outlier tutorial reference issue

* Treat 'develop' and 'deploy' as a unified search space

* fix some frontmatter aliases

* add section header redirects

* fix 'Operations' redirect to go to 'Deploy' tab

* update redirects file for * pattern

* add :splat to redirect references

* Add wildcard to each entry in _redirects file

---------

Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-04-20 18:03:45 +02:00
Dylan Couzon 814bdc1a48 clean up messaging 2026-04-20 11:06:41 -04:00
Dylan CouzonandClaude Sonnet 4.6 52a38d255c Link new retrieval quality tutorials from navigation indexes
Updates both search engineering index files (the headless partial and
the tutorials-lp overview) to list Retrieval Quality Fundamentals and
Building a Golden Query Set alongside the existing Retrieval Quality
Evaluation row. The Evaluation row is also retitled from "Measure
quality and tune HNSW parameters" to "Measure ANN recall and tune
HNSW parameters" to match the refactored page.

Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>
2026-04-20 10:47:02 -04:00
Dylan CouzonandClaude Sonnet 4.6 7d3d44b727 Refactor Retrieval Quality Evaluation to recall terminology
Renames precision to recall throughout the ANN-evaluation tutorial so
the page aligns with the ANN-benchmarks convention and with the new
Retrieval Quality Fundamentals page. The numerical formula is
unchanged: when ANN and exact search both return exactly k items,
recall@k and precision@k are numerically identical.

Other changes:
- Remove the Quality metrics subsection, now covered by the
  Fundamentals page, and replace it with a short link across.
- Bump weight from 4 to 6 so the three retrieval-quality pages
  order as Fundamentals, Golden Query Set, Evaluation.
- Fix a pre-existing prose/code mismatch: the prose said "first
  50000 items" while the code uses range(60000).

Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>
2026-04-20 10:46:19 -04:00
Dylan CouzonandClaude Sonnet 4.6 5db9d106aa Add Retrieval Quality Fundamentals and Golden Query Set tutorials
Introduces two new conceptual tutorials under tutorials-search-engineering:

- Retrieval Quality Fundamentals covers the three-level evaluation
  framework (ANN recall, retrieval relevance, business impact), the
  evaluation ladder that connects them in practice, and a which-metric-
  when decision table keyed by scenario and available ground truth.
- Building a Golden Query Set covers query generation at scale (logs,
  LLM synthesis, human annotation) and the failure modes commonly
  lumped together as data leakage: synthetic-query unrealism,
  embedding-model contamination, near-duplicate documents, temporal
  drift, and reviewer reproducibility.

Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>
2026-04-20 10:42:05 -04:00
Derrick MwitiandAbdon Pijpelink dc564bbab0 Time based sharding tutorial (#1805)
* add snippets

* add snippets

* add snippets

* add snippets descriptions

* Initial commit

* Updated tutorial

* Expand on shard_number

* Add sentence about optimal batch size

* Show how to query multiple shards

* Small edits

* Fixes

* Add to landing page; adjust weight

* Clean up code snippets

* Don't set shard_key in code snippet

* Review feedback

* Update links to moved pages

* Add C# code snippets

* Add Go code snippets

* Add Java code snippets

* Add Rust code snippets

* Add TS code snippets

* Add note about changing collection settings

* Move CSV parsing to separate function

* Trigger Build

* Remove unused dependency

---------

Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-04-16 22:19:08 +02:00
Bastian Hofmann 6d041d78ba Merge pull request #2257 from qdrant/feat/bashofmann/multi-az
Add documentation for Multi AZ clusters
2026-04-15 15:57:14 +02:00
Filip Makraduli 44cf24a4aa Add Superlinked embeddings provider page (#2276)
Adds a documentation page for Superlinked (SIE) as a Qdrant embedding
provider. The sie-qdrant package provides SIEVectorizer for dense
embeddings and SIENamedVectorizer for multi-type (dense, sparse, and
multivector/ColBERT) embeddings, enabling hybrid search via Qdrant's
Reciprocal Rank Fusion and native MaxSim retrieval via MultiVectorConfig.
Python-only.
2026-04-14 15:57:41 +05:30
Abdon Pijpelink 2b586cadb1 Clarify upload queue in Edge synchronization guide (#2259) 2026-04-10 15:49:56 +02:00
Bastian Hofmann 2b99c06227 Merge pull request #2255 from qdrant/feat/bashofmann/audit-logging
Add audit logging documentation for cluster configuration
2026-04-07 14:18:49 +02:00
Bastian Hofmann 55657e5f95 Merge pull request #2256 from qdrant/feat/bashofmann/gpu
Add documentation for GPU clusters
2026-04-07 14:18:36 +02:00
Abdon Pijpelink 5705d7e2a2 Fix HTTP code snippets (#2231) 2026-04-07 11:35:51 +02:00
Abdon Pijpelink cbbd26e9b0 Document how to rotate an API key without any downtime (#2260) 2026-04-07 11:03:26 +02:00
Bastian Hofmann c29622a642 Clarified that you can only add 1 gpu 2026-04-07 10:44:09 +02:00
Bastian Hofmann 07e7b8270b add link 2026-04-07 10:34:31 +02:00
Bastian HofmannandWilliam Godfrey 6e5342aaf9 Update qdrant-landing/content/documentation/cloud/create-cluster.md
Co-authored-by: William Godfrey <will@catalyst-tech.io>
2026-04-02 17:43:26 +02:00
Bastian HofmannandWilliam Godfrey 1726f8d7c5 Update qdrant-landing/content/documentation/cloud/configure-cluster.md
Co-authored-by: William Godfrey <will@catalyst-tech.io>
2026-04-02 16:03:00 +02:00
Bastian HofmannandWilliam Godfrey ee8c7542e1 Update qdrant-landing/content/documentation/cloud/configure-cluster.md
Co-authored-by: William Godfrey <will@catalyst-tech.io>
2026-04-02 16:02:35 +02:00
Bastian Hofmann a74fefb77c The logs endpoint returns logs from all nodes 2026-04-02 11:55:27 +02:00