Commit Graph
110 Commits
Author SHA1 Message Date
Dylan CouzonandClaude Opus 4.7 d615b8fc30 measuring-retrieval-relevance: accessibility pass on ranking metrics
Expands MRR with Mean Reciprocal Rank and a Wikipedia link where the
metric becomes operational, matching the inline expansion treatment
already given to NDCG. Drops the IR acronym in favor of the plainer
"ranking metrics" phrasing. Strips the parenthetical subtitle from the
Pitfalls heading and normalizes the intro to use "golden sets".

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-23 23:40:40 -04:00
Dylan CouzonandClaude Opus 4.7 4d5a3b8f60 ann-precision: add Prerequisites block and rename closing to Next Steps
Adds an explicit Prerequisites line matching the pattern in Measuring
Retrieval Relevance and Evaluating Pipeline Output Quality. Renames
"Wrapping Up" to "Next Steps" for naming consistency across the series.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-23 23:40:09 -04:00
Dylan CouzonandClaude Opus 4.7 6c10996b66 rename Building a Golden Query Set to Measuring Retrieval Relevance
Aligns the layer-2 tutorial title with the parallel "Measuring X" /
"Evaluating X" pattern used by the other two and maps directly to the
four-layer framework. Slug stays the same to preserve URLs and the
golden-set artifact identity in the path. Also updates the nav descriptions
to reflect the tutorial's full scope (build + score, not just build).

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-23 23:29:44 -04:00
Dylan CouzonandClaude Opus 4.7 e02bfb0bf1 delete retrieval-quality-fundamentals and remove from nav indexes
The four-layer framework, metric-selection table, and business-impact
guidance now live inside the three execution tutorials. The fundamentals
page is no longer needed as a shared reference and readers don't have to
leave the tutorial flow to get context.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-23 23:22:40 -04:00
Dylan CouzonandClaude Opus 4.7 cbda4a6c9c pipeline-output: replace Next Steps with Connecting to Business Impact
Folds the business-impact framework (KPI selection, pre-registered decision
rules, offline-online calibration, proxy signals) into the tutorial so it
stands alone without the external fundamentals page. The ladder pointer in
the intro now references the four-layer section inside ANN Precision.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-23 23:19:17 -04:00
Dylan CouzonandClaude Opus 4.7 1ab6ff384e golden-set: inline metric-selection table and choosing-k guidance
Replaces the external pointer to retrieval-quality-fundamentals with an
inline scenario-to-metric table plus guidance on picking k. The ladder
pointer now references the four-layer section inside ANN Precision rather
than the fundamentals page.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-23 23:16:37 -04:00
Dylan CouzonandClaude Opus 4.7 29640488d9 ann-precision: inline the four-layer framework and MTEB ceiling note
Replaces the external pointer to retrieval-quality-fundamentals with a
self-contained section introducing the four evaluation layers. The tutorial
no longer depends on fundamentals for orientation.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-23 23:15:34 -04:00
Dylan Couzon 07fb044ed4 clean up tutorial 3 2026-04-23 23:09:41 -04:00
Dylan Couzon 5ca06869c2 Improve dataset generation for next guide. 2026-04-23 23:07:32 -04:00
Dylan Couzon 0e6e3e3c5c golden-set final update 2026-04-23 21:41:16 -04:00
Dylan Couzon 0ed547e945 update code snippets 2026-04-23 20:37:11 -04:00
Dylan Couzon 7a4129be1e add third guide + update indices 2026-04-23 20:23:10 -04:00
Dylan Couzon 54ecb5b63c turn fundamentals into blog 2026-04-23 20:22:55 -04:00
Dylan Couzon c21062782b clean up golden set 2026-04-23 20:22:29 -04:00
Dylan Couzon e3e0516530 clean up Retrieval quality 2026-04-23 20:22:19 -04:00
Dylan Couzon 1fbc73c4b2 Merge branch 'master' into retrival-quality-guide-improvement 2026-04-23 12:24:29 -04:00
Dylan Couzon 340c6528ca Clean up, tone 2026-04-23 11:06:18 -04:00
Dylan CouzonandClaude Opus 4.7 fb1672c102 retrieval-quality-fundamentals: tone pass
- Soften "how teams bridge this gap" to "common patterns for
  bridging this gap" so we don't imply we harvested real client
  pipelines for this writeup
- Reframe the layer-2/3 diagnostic and the "Isolate the component
  under test" bullet so they name multiple downstream consumer
  types (LLM generator, ranker, UI) rather than assuming RAG
- Add a one-line caveat that A/B design for RAG and agentic
  systems is still evolving to the Proxy KPIs paragraph
- Simplify the recall@k / precision@k equivalence note and link
  ann-benchmarks.com as the citation for community convention

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 22:39:03 -04:00
Dylan CouzonandClaude Opus 4.7 e2df842fb5 retrieval-quality-fundamentals: structure & scope pass
- Add "why measure retrieval quality" lead-in paragraph
- Expand ANN on first use; flag sparse vectors as out of scope
  for layer 1 (they use exact matching)
- Add LLM-as-judge to layer-2 ground-truth options
- Break the Tooling bullet into per-layer recommendations: Web UI
  for L1, ranx for L2, Ragas/Phoenix/DeepEval for L3
- Reorder Quality Metrics so layer 1 (ANN recall formula + exact
  kNN equivalence) comes before the generic layer-2 relevance
  metrics; trim a redundant sentence
- Add end-to-end answer quality as a distinct third layer in the
  prose intro so it matches the ladder table's four rows
- Standardize vocabulary on "layer" (was mixing "level" in the
  intro with "layer" everywhere else); update the section anchor
  to #connecting-the-layers-in-practice in this file and the two
  cross-linking tutorials

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 22:31:35 -04:00
Dylan Couzon 5748d6db8c update for moved URL 2026-04-22 22:17:07 -04:00
Dylan Couzon 92af150b00 retrieval-quality: rename to "Measuring ANN Precision"
- Update title, H1, and Time in retrieval-quality.md
    (30 min -> 15 min reflects the pivot to Web UI)
  - Rename references in tutorials-lp-overview.md and the headless
    tutorial index; swap the pill from Python to Web UI to reflect
    the new primary flow
  - Replace "ANN recall" with "ANN precision" in Fundamentals
    (4 places: intro, comparison note, ladder table, cross-link)
    and in the golden-set tutorial's layer-1 cross-reference
  - Filename kept as retrieval-quality.md so existing URLs and
    aliases still work
2026-04-22 17:13:27 -04:00
Dylan Couzon 175f92ddbe change recall back to precision@k
The Search Quality tab in the Web UI reports precision@k, not
  recall@k, so the whole tutorial now uses precision@k as the metric
  label: "ANN recall" -> "ANN precision" in the anchor, section
  headers, Python helper (avg_precision_at_k), and prose. Kept a
  one-line bridge note that ANN-benchmarks terminology calls this
  recall@k, since both searches return exactly k items.
2026-04-22 17:02:45 -04:00
Dylan Couzon 921b645e70 retrieval-quality: pivot to Web UI, add CI skeleton, trim setup
- Drop the dataset-setup walkthrough (HF loading, collection
    create, upload, wait-for-green). Readers at this phase already
    have a collection.
  - Replace the Python evaluation block with a "Measure ANN Recall
    with the Web UI" section built around the Search Quality tab.
    Default run is one-click (sample size 10); HNSW tuning uses the
    tab's advanced mode instead of update_collection. Three
    screenshot placeholders at
    /documentation/tutorials/retrieval-quality/*.png.
  - Reflect that the tab reports precision@k; note the recall@k
    equivalence already spelled out in the ANN Recall section.
  - Keep Python but move it to an "Automate in CI" section with a
    reusable skeleton function.
  - Collapse the standalone "Embeddings Quality" section into a
    one-sentence MTEB pointer inside ANN Recall.
  - Rewrite Wrapping Up to match the new scope.
  - Link the HNSW tuning section to Optimize Performance for the
    full parameter reference.
2026-04-22 16:34:16 -04:00
Dylan CouzonandClaude Opus 4.7 e5d9101227 retrieval-quality: tone pass on intro, section rename, drop RAG link
- Rewrite the intro so the ANN algorithm reads as one of several
  levers shaping retrieval quality (alongside the embedding model,
  retrieval strategy, filtering, reranking) rather than the only
  factor beyond embeddings. Addresses mrscoopers on the reductive
  "embeddings + ANN" framing.
- Rename the "Retrieval Quality" section to "ANN Recall" and
  rewrite its opening paragraph to match; ANN approximation quality
  isn't the same as retrieval quality broadly.
- Drop the RAG-evaluation-guide link from the three places it
  appeared in this file. This tutorial isn't RAG-specific.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 15:26:29 -04:00
Dylan CouzonandClaude Opus 4.7 fd095af85a retrieval-quality: layer-1 anchor and title-case headers
- Add a layer-1 anchor sentence under the Time/Level table
  pointing at the evaluation ladder in Fundamentals, mirroring
  the golden-set tutorial's anchor. Addresses abdonpijpelink's
  ask for a levels-table link and a "this tutorial focuses on
  level 1" framing.
- Title-case the six H2 headers for consistency across the
  tutorials-search-engineering set.

Deferred: streaming the 60K training items into upload_points
instead of materializing as a list (abdonpijpelink line 69) —
pending manager confirmation before proceeding.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 15:06:47 -04:00
Dylan Couzon 81883cbecc golden set: framing 2026-04-22 14:31:10 -04:00
Dylan CouzonandClaude Opus 4.7 a47c57ebdc golden set: use ranx for metrics, reframe Anthropic example
- Switch the metrics section from three hand-rolled functions
  (recall@k, MRR, NDCG@k) and a bespoke evaluate() loop to a
  single ranx-based example. Handles binary and graded labels
  in one call (mrscoopers, line 59).
- Reframe the Anthropic synthetic-generation snippet as a
  minimal prompt shape, point at Ragas for readers who want a
  maintained testset generator, and fix a bug (Anthropic() was
  called without importing the class; switched to
  anthropic.Anthropic()). Addresses abdonpijpelink line 53 and
  mrscoopers line 31.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 13:52:23 -04:00
Dylan CouzonandClaude Opus 4.7 acc85e65ca golden set: anchor intro at layer 2, link layer 1 to Web UI
- Open the tutorial by anchoring it at layer 2 of the evaluation
  ladder and linking to the levels table in retrieval-quality-
  fundamentals (abdonpijpelink, line 14).
- Point layer-1 readers (ANN recall vs exact kNN) at the Search
  Quality tab in the Qdrant Web UI instead of the retrieval-quality
  tutorial (mrscoopers, line 13).
- Trim the duplicate layer-1 pointer at the end of "Using the
  Golden Set" (mrscoopers, line 113).
- Open cross-tutorial links in a new tab to match the repo
  convention.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-04-22 10:56:34 -04:00
Dylan Couzon 6eba7e2ceb docs(golden-set): anchor intro at layer 2 2026-04-22 10:30:25 -04:00
Dylan CouzonandAbdon Pijpelink 4d99153c90 make Anthropic API key explicit
Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-04-22 09:17:40 -04:00
Abdon Pijpelink 434a451c2b Update Hybrid Search with Reranking tutorial (#2274)
* Update for Cloud Inference and data ingestion

* Fix link

* Review feedback

* Make code snippets testable

* Add C# code snippets

* Add Go code snippets

* Add Java code snippets

* Add Rust code snippets

* Add TS code snippets

* Move CSV streaming/parsing to separate function
2026-04-22 11:03:05 +02:00
Dylan Couzon 1a9821cc95 golden set: reduce verbosity 2026-04-21 12:06:48 -04:00
Dylan Couzon bb8cac5a6e update completion times 2026-04-21 12:00:10 -04:00
Dylan Couzon 8cdc9acd14 golden dataset - improve tutorial quality 2026-04-20 12:42:53 -04:00
Dylan Couzon 02710a1f9b Add using golden set instructions 2026-04-20 12:41:21 -04:00
Dylan Couzon 5503b005b6 improve wording 2026-04-20 12:32:33 -04:00
Dylan Couzon d6f06d6b60 clean up tables 2026-04-20 12:22:45 -04:00
Dylan Couzon e9e0e4bddd Syntax 2026-04-20 12:21:05 -04:00
kanungleandAbdon Pijpelink 544708f293 Restructure Docs - Stage 4a (#2280)
* create Develop and Deploy tabs; move Operations; re-weight pages

* move capacity planning page; create section dropdown content

* added aliases to frontmatter

* update link references to new canonical links; maintain anchoring

* address remaining link issues and errors

* fix outlier tutorial reference issue

* Treat 'develop' and 'deploy' as a unified search space

* fix some frontmatter aliases

* add section header redirects

* fix 'Operations' redirect to go to 'Deploy' tab

* update redirects file for * pattern

* add :splat to redirect references

* Add wildcard to each entry in _redirects file

---------

Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-04-20 18:03:45 +02:00
Dylan Couzon 814bdc1a48 clean up messaging 2026-04-20 11:06:41 -04:00
Dylan CouzonandClaude Sonnet 4.6 7d3d44b727 Refactor Retrieval Quality Evaluation to recall terminology
Renames precision to recall throughout the ANN-evaluation tutorial so
the page aligns with the ANN-benchmarks convention and with the new
Retrieval Quality Fundamentals page. The numerical formula is
unchanged: when ANN and exact search both return exactly k items,
recall@k and precision@k are numerically identical.

Other changes:
- Remove the Quality metrics subsection, now covered by the
  Fundamentals page, and replace it with a short link across.
- Bump weight from 4 to 6 so the three retrieval-quality pages
  order as Fundamentals, Golden Query Set, Evaluation.
- Fix a pre-existing prose/code mismatch: the prose said "first
  50000 items" while the code uses range(60000).

Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>
2026-04-20 10:46:19 -04:00
Dylan CouzonandClaude Sonnet 4.6 5db9d106aa Add Retrieval Quality Fundamentals and Golden Query Set tutorials
Introduces two new conceptual tutorials under tutorials-search-engineering:

- Retrieval Quality Fundamentals covers the three-level evaluation
  framework (ANN recall, retrieval relevance, business impact), the
  evaluation ladder that connects them in practice, and a which-metric-
  when decision table keyed by scenario and available ground truth.
- Building a Golden Query Set covers query generation at scale (logs,
  LLM synthesis, human annotation) and the failure modes commonly
  lumped together as data leakage: synthetic-query unrealism,
  embedding-model contamination, near-duplicate documents, temporal
  drift, and reviewer reproducibility.

Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>
2026-04-20 10:42:05 -04:00
Mohamed Arbi 0b6ccaf51a Merge branch 'master' into fix/monitoring-logging-doc-typos 2026-03-31 02:11:16 +02:00
goodnight 48d8d761f2 fix(docs): correct typos and improve clarity in monitoring and some relevance docs 2026-03-31 00:53:46 +01:00
1a40961d62 fix: linkchecker include filter port mismatch (#2242)
* initial commit; fixed anchor links on internal docs pages

* add back in absolute paths for links in code comments

* fix: update linkchecker include filter to match server port 1314

PR #1629 changed the Hugo server to port 1314 but forgot to update
the --include filter, which still matched port 1313. This caused all
links to be excluded, making the checker a no-op (0 checked, 82277 excluded).

Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com>

* Fix url rewrite regex so images are not impacted

* Fix links from non-documentation pages

* Fix broken links

* more broken links

* more broken links

* broken link

* Add srcset width descriptor to .lycheeignore

* Ignore URLs that contain a % character

* Anchor regex so it matches the entire URL

---------

Co-authored-by: kanungle <neil.kanungo@gmail.com>
Co-authored-by: Claude Opus 4.6 (1M context) <noreply@anthropic.com>
Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-03-30 17:21:32 +02:00
Mohamed ArbiandAbdon Pijpelink c0db45ed8f Fix Docs : minor grammar fixes (#2218)
* fix(docs): fix typos and some links in documentation

* fix(docs): correct typo in filtering.md

* Update qdrant-landing/content/documentation/headless/snippets/inference/jinaai-upsert/generated/typescript.md

Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>

* Update qdrant-landing/content/documentation/headless/snippets/inference/multiple/generated/typescript.md

Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>

* Update qdrant-landing/content/documentation/hybrid-cloud/configure-scale-upgrade.md

Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>

* Update qdrant-landing/content/documentation/cloud-api.md

Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>

---------

Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-03-26 17:23:36 +01:00
kanungle e2108a796f Reorg of User Manual section; refactor weights 2026-03-14 23:37:29 -07:00
Evgeniya Sukhodolskaya 8fc7402118 more clear structure 2026-03-01 16:15:13 +01:00
JennyandAbdon Pijpelink ed178cb49c added relevance feedback tutorial v1 (#2171)
* added relevance feedback tutorial v1

* Small edits

* Add link to tutorial from docs

---------

Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-02-26 11:33:27 +01:00
Abdon Pijpelink 87180dc77f Resolve merge conflicts 2026-02-05 17:22:56 +01:00