Skip to content

Add guidance on aggregating over large result sets - #4855

Open
thomasht86 wants to merge 1 commit into
masterfrom
thomasht86/corpus-level-aggregations
Open

Add guidance on aggregating over large result sets#4855
thomasht86 wants to merge 1 commit into
masterfrom
thomasht86/corpus-level-aggregations

Conversation

@thomasht86

Copy link
Copy Markdown
Contributor

Grouping over a large matched set, such as category counts over the full corpus, is a common thing users want to do, causing issues if done frequently.

Adds a section to the grouping guide describing how to reduce the work per query, how to compute approximate counts from a random 1% sample using a synthetic random field, and when to cache aggregates in a Searcher. Link to it from the FAQ and the performance docs.

I confirm that this contribution is made under the terms of the license found in the root directory of this repository's source tree and that I have the authority necessary to make this contribution on behalf of its copyright owner.

Grouping over a large matched set, such as category counts over the
full corpus, is evaluated on every matched document on every content
node. Add a section to the grouping guide describing how to reduce the
work per query, how to compute approximate counts from a random 1%
sample using a synthetic random field, and when to cache aggregates in
a Searcher. Link to it from the FAQ and the performance docs.

Co-Authored-By: Claude Fable 5.1 <[email protected]>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant