discourse-ai

Commit Graph

Author	SHA1	Message	Date
Roman Rizzi	9a79afcdbf	DEV: Better strategies for summarization (#88 ) * DEV: Better strategies for summarization The strategy responsibility needs to be "Given a collection of texts, I know how to summarize them most efficiently, using the minimum amount of requests and maximizing token usage". There are different token limits for each model, so it all boils down to two different strategies: Fold all these texts into a single one, doing the summarization in chunks, and then build a summary from those. Build it by combining texts in a single prompt, and truncate it according to your token limits. While the latter is less than ideal, we need it for "bart-large-cnn-samsum" and "flan-t5-base-samsum", both with low limits. The rest will rely on folding. * Expose summarized chunks to users	2023-06-27 12:26:33 -03:00
Rafael dos Santos Silva	8742535024	FEATURE: Allow using large context OpenAI models for summarization (#86 )	2023-06-13 15:23:48 -03:00
Roman Rizzi	3364fec425	DEV: Remove the summarization feature (#83 ) * DEV: Remove the summarization feature Instead, we'll register summarization implementations for OpenAI, Anthropic, and Discourse AI using the API defined in discourse/discourse#21813. Core and chat will implement features on top of these implementations instead of this plugin extending them. * Register instances that contain the model, requiring less site settings	2023-06-13 14:32:26 -03:00

Author

SHA1

Message

Date

Roman Rizzi

9a79afcdbf

DEV: Better strategies for summarization (#88 )

* DEV: Better strategies for summarization

The strategy responsibility needs to be "Given a collection of texts, I know how to summarize them most efficiently, using the minimum amount of requests and maximizing token usage".

There are different token limits for each model, so it all boils down to two different strategies:

Fold all these texts into a single one, doing the summarization in chunks, and then build a summary from those.
Build it by combining texts in a single prompt, and truncate it according to your token limits.

While the latter is less than ideal, we need it for "bart-large-cnn-samsum" and "flan-t5-base-samsum", both with low limits. The rest will rely on folding.

* Expose summarized chunks to users

2023-06-27 12:26:33 -03:00

Rafael dos Santos Silva

8742535024

FEATURE: Allow using large context OpenAI models for summarization (#86 )

2023-06-13 15:23:48 -03:00

Roman Rizzi

3364fec425

DEV: Remove the summarization feature (#83 )

* DEV: Remove the summarization feature

Instead, we'll register summarization implementations for OpenAI, Anthropic, and Discourse AI using the API defined in discourse/discourse#21813.

Core and chat will implement features on top of these implementations instead of this plugin extending them.

* Register instances that contain the model, requiring less site settings

2023-06-13 14:32:26 -03:00

3 Commits