Commit Graph
6335 Commits
Author SHA1 Message Date
Dylan Couzon f2e66d326f Expand hybrid-queries with DBSF, weighted RRF tuning, FormulaQuery 2026-05-15 18:00:05 -04:00
Dylan Couzon 8cdce983ed upload snippets 2026-05-15 16:49:56 -04:00
daniel-azoulai ced6f2c513 Update 2026-05-15 13:29:44 -07:00
daniel-azoulai 2019d92fdd Update 2026-05-15 11:30:01 -07:00
daniel-azoulai f02cdb4c8a Update case-study-go-perfect.md 2026-05-15 11:21:53 -07:00
daniel-azoulai c2515f8500 update 2026-05-15 11:17:36 -07:00
Dylan Couzon 1d36116d77 Merge pull request #2334 from qdrant/hybrid-search-gap-2
Multi-representation search tutorial + supporting doc cross-links
2026-05-15 14:17:28 -04:00
Dylan CouzonandClaude Opus 4.7 f680674669 sharpen dense_title rationale and alert linear-fusion warning
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-05-15 14:14:21 -04:00
Tim Visée db67ab203f Merge pull request #2263 from qdrant/correct-param-index-note
Correct parameterized index note
2026-05-15 17:46:49 +02:00
Abdon Pijpelink 1a0f06231c Add links to TurboQuant article (#2354)
* Add links to TQ article

* Add link to articles for other quantization methods

* Update TQ article preview images
2026-05-15 16:26:18 +02:00
daniel-azoulai f3d108c413 Create case-study-go-perfect.md 2026-05-14 09:31:59 -07:00
kanungle 33410a4795 Merge pull request #2349 from qdrant/maddie-qdrant-patch-6
Update top-banner.md
2026-05-13 14:43:27 -07:00
m-qdrant fc74dab48d Update top-banner.md 2026-05-13 16:31:12 -04:00
m-qdrant 96bf9222cf Update top-banner.md 2026-05-13 16:30:19 -04:00
Dylan CouzonandClaude Opus 4.7 d355db954e link keyword index, price storage tradeoff, add Discord CTA
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-05-13 16:05:19 -04:00
Dylan Couzon e9aa78796f fix miss matched key value pair 2026-05-13 15:54:13 -04:00
kanungle 10299579aa Merge pull request #2328 from qdrant/turboquant-article
TurboQuant article
2026-05-13 10:49:50 -07:00
jojii 3baa204409 Update date of the article 2026-05-13 17:21:19 +02:00
kanungle 209991c3c2 Merge pull request #2323 from qdrant/meta-descriptions
initial commit; Claude generated SEO descriptions
2026-05-13 08:18:46 -07:00
jojii cb477c16c5 Add link to visual reprentation of TQ 2026-05-13 11:03:19 +02:00
jojii f9d760c05e Less em dashes 2026-05-13 10:37:39 +02:00
kanungle b013a0b2e4 Merge branch 'master' into meta-descriptions 2026-05-12 15:44:44 -07:00
trean d0bd23c266 fix typo 2026-05-12 18:22:21 +02:00
Ivan Pleshkov 1ad8138ec5 fix bullet padding 2026-05-12 17:52:56 +02:00
Ivan Pleshkov 4e5361d5b3 use x as letter 2026-05-12 17:16:38 +02:00
daniel-azoulai dbb82c7006 Merge pull request #2342 from qdrant/case-study-sapu
case-study-sapu
2026-05-12 08:13:52 -07:00
trean 9bbfb79295 vsd page: update speaker lineup copy (#2347) 2026-05-12 16:48:14 +02:00
daniel-azoulai d2af0904eb Update case-study-sapu.md 2026-05-12 07:28:02 -07:00
jojii 962ac60a8f Crop and scale images 2026-05-12 14:53:20 +02:00
Ivan PleshkovandAbdon Pijpelink 98c3bd3a75 Update qdrant-landing/content/articles/turboquant-quantization.md
Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-05-12 13:56:55 +02:00
Ivan PleshkovandAbdon Pijpelink 3025dc6415 Update qdrant-landing/content/articles/turboquant-quantization.md
Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-05-12 13:56:28 +02:00
Ivan PleshkovandAbdon Pijpelink 18327e309e Update qdrant-landing/content/articles/turboquant-quantization.md
Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-05-12 13:56:17 +02:00
brone1323andAbdon Pijpelink 5aa98e5959 fix(docs): correct C# MatchExcept snippet (was using Match instead of MatchExcept) (#2345)
* fix(docs): correct C# MatchExcept snippet to use MatchExcept instead of Match

The C# code example in the match-except filter condition snippet was
incorrectly using Match() instead of MatchExcept(), making it identical
to the match-any example. This is misleading for developers trying to
implement the MatchExcept (NOT IN) filter condition in C#.

Fixes #1656

* Add Markdownified code snippet

---------

Co-authored-by: Abdon Pijpelink <abdon.pijpelink@qdrant.com>
2026-05-12 10:26:57 +02:00
Ivan Pleshkov 1258c41d0d fix links 2026-05-12 09:51:51 +02:00
Dylan Couzon 31a3c88379 add document_id payload index and tune upload_points
Adds the missing keyword index on document_id so grouping works under
strict mode (Cloud default), and tunes the upload_points call to
batch_size=256, parallel=2 for faster ingestion. Mirrors the notebook
in qdrant/examples#103.
2026-05-11 22:05:24 -04:00
Dylan Couzon 5502b91064 point notebook links at Colab for direct run
Switches the two body references to the accompanying notebook from
GitHub URLs to githubtocolab so readers can run it without cloning.
The header table's GitHub link stays for readers who want the source
view.
2026-05-11 21:14:59 -04:00
Ivan Pleshkov 29b04098ef rename deep dive article 2026-05-12 03:05:35 +02:00
Dylan Couzon def2b3a39b improve clarity FormulaQuery syntax 2026-05-11 20:51:07 -04:00
Dylan Couzon 7f9ad2149f update visual 2026-05-11 20:50:42 -04:00
Dylan CouzonandClaude Opus 4.7 3de7b1ddfb tighten intro, dataset, and wrap-up; add notebook pointer before setup
Tightens awkward and stale phrasing across the intro and Dataset
section (per a full review pass), repositions FormulaQuery in Wrapping
Up as an alternative rather than part of the default pipeline, drops
the redundant 'document-side equivalent' closing line, and adds a
one-line pointer to the notebook right before the Setup section so
readers can pivot to runnable code.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-05-11 20:23:42 -04:00
Dylan CouzonandClaude Opus 4.7 b30140820f tighten retrieval section and remove vector-description duplication
Adds dense_abstract as a fourth prefetch in retrieve() since we
ingest it already. Removes the unactionable dense_abstract hedge
paragraph, the When to Group section (the prefetch-limit gotcha
folds into a code comment), the duplicate schema-section vector
bullets, and the awkward transition sentence between the code and
the design subsections. Replaces 'summary' with 'abstract as a
whole' in descriptions for consistency with the schema rename done
earlier.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-05-11 19:53:39 -04:00
Dylan CouzonandClaude Opus 4.7 479d305aca explain query_points_groups inline and remove pattern-doesnt-fit section
Adds a one-line comment above the query_points_groups call so readers
see what the function does without waiting for the When to Group
subsection (which now sits at the end of the design-decisions list).

Removes the entire 'Where This Pattern Doesn't Fit' section: the
short/homogeneous paragraph read as obvious (a reader who's deep into
this tutorial wouldn't try to multi-rep a tweet), and the
inconsistent-metadata paragraph was too vague to be actionable.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-05-11 19:08:37 -04:00
Dylan CouzonandClaude Opus 4.7 8ea4653ee8 move categories from BM25 input to filterable payload
Renames sparse_keywords to sparse_title (the vector now indexes only
the title text), adds a keyword index on the tags payload field, and
gives retrieve() an optional tags parameter that builds a query_filter
when set. Pre-filtering on the tags payload is faster and more precise
than mixing categories into BM25 lexical matching.

Updates the schema description bullets, the prefetch justification,
the intro failure-mode list, and the dataset framing to match the
new design. Drops avg_len from 15 to 10 to reflect title-only word
counts.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-05-11 18:44:37 -04:00
Dylan CouzonandClaude Opus 4.7 8abae12a8c consolidate chunking note and remove open ends section
Moves the standalone chunking-strategy sentence into the Dataset
paragraph that already explains why we chunk, and drops the dedicated
Open Ends section. The chunking-method link list (POMA-AI VST, Jina,
Chonkie) goes away with the section, and the BM25F note is also
removed since the workaround it describes is what the tutorial
already demonstrates. Cleans up two stale step-number references in
Wrapping Up.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-05-11 18:18:51 -04:00
Dylan CouzonandClaude Opus 4.7 a9610d2a30 restructure fusion section to cover four named strategies
Renames "How to Fuse" to "Which Fusion to Use" and expands it to
cover RRF (default), Weighted RRF, DBSF, and custom formulas as four
named options, with the FormulaQuery code block moved in from what
was the boosting section. Softens RRF framing from "stick with RRF
unless..." to "reasonable starting point; variants often do better
once you have an eval set." Adds the distribution-alignment
explanation in the custom-formula paragraph and links the in-repo
Decay Functions and Score Boosting references along with the RRF vs
DBSF FAQ entry.

The "When to Boost, When to Rerank" section now only covers true
boosting (recency, authority, decay) and reranking, and is reordered
to sit immediately after fusion so the ranking decisions stay
together. The "When to Group, When Not To" section moves to the end
as a presentation concern.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-05-11 17:33:32 -04:00
daniel-azoulai 1f2197b7c0 Update case-study-sapu.md 2026-05-11 13:57:08 -07:00
daniel-azoulai 484061ff80 Update case-study-sapu.md 2026-05-11 13:49:12 -07:00
Ivan Pleshkov d91e94a365 review remark 2026-05-11 22:29:43 +02:00
Dylan CouzonandClaude Opus 4.7 68cec300f2 set avg_len for BM25 ingestion and trim parameter notes
Calibrates BM25 length normalization for the short title+categories
sparse field with a comment on why. Removes redundant k1/b/avg_len
prose from the tutorial Open Ends section and the cross-link paragraph
in text-search.md, since the BM25 Parameters subsection above already
documents calibration with a working example.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-05-11 16:06:03 -04:00
Dylan CouzonandClaude Opus 4.7 376202a113 remove lookup_from references, point to with_lookup for payload splits
The two lookup_from mentions were misleading: the feature is for
querying by ID across collections, not for splitting representation
storage. The line-107 paragraph now points readers to the documented
with_lookup pattern for the payload-split case and stays silent on
vector splits, which are a separate design problem (multiple queries
plus client-side fusion) that doesn't fit this tutorial's scope.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
2026-05-11 15:24:08 -04:00