discourse-ai

Commit Graph

Author	SHA1	Message	Date
Rafael dos Santos Silva	b25daed60b	FEATURE: Llama2 for summarization (#116 )	2023-07-27 13:55:32 -03:00
Sam	4b0c077ce5	FEATURE: port to use claude-2 for chat bot (#114 ) Claude 1 costs the same and is less good than Claude 2. Make use of Claude 2 in all spots ... This also fixes streaming so it uses the far more efficient streaming protocol.	2023-07-27 11:24:44 +10:00
Rafael dos Santos Silva	e3b4a73267	FEATURE: Cache Related Topics for longer (#110 )	2023-07-18 11:27:06 -03:00
Roman Rizzi	473732c18a	FIX: Return base prompt instead of nil (#106 )	2023-07-13 21:48:25 -03:00
Rafael dos Santos Silva	703762a7a9	PERF: .find_each instead of .find to save us from memory allocation peaks also Fix embeddings rake task for new db structure	2023-07-13 18:59:25 -03:00
Roman Rizzi	5f0c617880	REFACTOR: Cohesive narrative for single-chunk summaries. (#103 ) Single and multi-chunk summaries end using different prompts for the last summary. This change detects when the summarized content fits in a single chunk and uses a slightly different prompt, which leads to more consistent summary formats. This PR also moves the chunk-splitting step to the `FoldContent` strategy as preparation for implementing streamed summaries.	2023-07-13 17:05:41 -03:00
Rafael dos Santos Silva	5e3f4e1b78	FEATURE: Embeddings to main db (#99 ) * FEATURE: Embeddings to main db This commit moves our embeddings store from an external configurable PostgreSQL instance back into the main database. This is done to simplify the setup. There is a migration that will try to import the external embeddings into the main DB if it is configured and there are rows. It removes support from embeddings models that aren't all_mpnet_base_v2 or OpenAI text_embedding_ada_002. However it will now be easier to add new models. It also now takes into account: - topic title - topic category - topic tags - replies (as much as the model allows) We introduce an interface so we can eventually support multiple strategies for handling long topics. This PR severely damages the semantic search performance, but this is a temporary until we can get adapt HyDE to make semantic search use the same embeddings we have for semantic related with good performance. Here we also have some ground work to add post level embeddings, but this will be added in a future PR. Please note that this PR will also block Discourse from booting / updating if this plugin is installed and the pgvector extension isn't available on the PostgreSQL instance Discourse uses.	2023-07-13 12:41:36 -03:00
Rafael dos Santos Silva	9d10a152b9	FEATURE: Claude 2 for summarization and AIHelper (#101 )	2023-07-13 12:32:08 -03:00
Roman Rizzi	fbe1bab980	FIX: typo while updating a section (#98 )	2023-06-27 17:57:58 -03:00
Roman Rizzi	1b568f2391	FIX: Claude's max_tookens_to_sample is a required field (#97 )	2023-06-27 14:42:33 -03:00
Roman Rizzi	9a79afcdbf	DEV: Better strategies for summarization (#88 ) * DEV: Better strategies for summarization The strategy responsibility needs to be "Given a collection of texts, I know how to summarize them most efficiently, using the minimum amount of requests and maximizing token usage". There are different token limits for each model, so it all boils down to two different strategies: Fold all these texts into a single one, doing the summarization in chunks, and then build a summary from those. Build it by combining texts in a single prompt, and truncate it according to your token limits. While the latter is less than ideal, we need it for "bart-large-cnn-samsum" and "flan-t5-base-samsum", both with low limits. The rest will rely on folding. * Expose summarized chunks to users	2023-06-27 12:26:33 -03:00
Sam	9390fba768	FIX: adjust token limits to account for functions (#96 ) Reduce maximum replies to 2500 tokens and make them even for both GPT-3.5 and 4 Account for 400+ tokens in function definitions (this was unaccounted for)	2023-06-23 10:02:04 +10:00
Sam	f8cabfad6b	FEATURE: Try to hone search so it reduces search terms in subsequent rounds (#95 ) Teach via system message that you can reduce search terms to get more results	2023-06-21 20:07:55 +10:00
Sam	a028309cbd	FEATURE: add ai_bot_enabled_chat commands and tune search (#94 ) * FEATURE: add ai_bot_enabled_chat commands and tune search This allows admins to disable/enable GPT command integrations. Also hones search results which were looping cause the result did not denote the failure properly (it lost context) * include more context for google command include more context for time command * type	2023-06-21 17:10:30 +10:00
Sam	d1ab79e82f	FEATURE: Add Azure cognitive service support (#93 ) The new site settings: ai_openai_gpt35_url : distribution for GPT 16k ai_openai_gpt4_url: distribution for GPT 4 ai_openai_embeddings_url: distribution for ada2 If untouched we will simply use OpenAI endpoints. Azure requires 1 URL per model, OpenAI allows a single URL to serve multiple models. Hence the new settings.	2023-06-21 10:39:51 +10:00
Sam	30778d8af8	FIX: avoid storing corrupt prompts (#92 ) ``` prompt << build_message(bot_user.username, reply) ``` Would store a "cooked" prompt which is invalid, instead just store the raw values which are later passed to build_message Additionally: 1. Disable summary command which needs honing 2. Stop storing decorations (searched for X) in prompt which leads to straying 3. Ship username directly to model, avoiding "user: content" in prompts. This was causing GPT to stray	2023-06-20 15:44:03 +10:00
Sam	70c158cae1	FEATURE: add full bot support for GPT 3.5 (#87 ) Given latest GPT 3.5 16k which is both better steered and supports functions we can now support rich bot integration. Clunky system message based steering is removed and instead we use the function framework provided by Open AI	2023-06-20 08:45:31 +10:00
Rafael dos Santos Silva	e457c687ca	FIX: OpenAI Tokenizer was failing to truncate mid emojis (#91 ) * FIX: OpenAI Tokenizer was failing to truncate mid emojis * Update spec/shared/tokenizer.rb Co-authored-by: Joffrey JAFFEUX <j.jaffeux@gmail.com> --------- Co-authored-by: Joffrey JAFFEUX <j.jaffeux@gmail.com>	2023-06-16 15:15:36 -03:00
Rafael dos Santos Silva	8742535024	FEATURE: Allow using large context OpenAI models for summarization (#86 )	2023-06-13 15:23:48 -03:00
Roman Rizzi	3364fec425	DEV: Remove the summarization feature (#83 ) * DEV: Remove the summarization feature Instead, we'll register summarization implementations for OpenAI, Anthropic, and Discourse AI using the API defined in discourse/discourse#21813. Core and chat will implement features on top of these implementations instead of this plugin extending them. * Register instances that contain the model, requiring less site settings	2023-06-13 14:32:26 -03:00
Sam	081231a6eb	FIX: support multiple command executions (#85 ) Previous to this change we were chaining stuff too late and would execute commands serially leading to very unexpected results This corrects this and allows us to run stuff like: > Search google 3/4 times on various permutations of QUERY and answer this question. We limit at 5 commands to ensure there are not pathological user cases where you lean on the LLM to flood us with results.	2023-06-06 07:09:33 +10:00
Sam	840968630e	FEATURE: disable smart commands on Claude and GPT 3.5 (#84 ) For the time being smart commands only work consistently on GPT 4. Avoid using any smart commands on the earlier models. Additionally adds better error handling to Claude which sometimes streams partial json and slightly tunes the search command.	2023-06-01 09:10:33 +10:00
Sam	96d521198b	FIX: missing localization (#81 ) blog.start_gpt_chat -> was on my blog This also slightly tunes the search prompt to support filtering by oldest and try a tiny bit harder to guide GPT 3.5 which is a bit of a losing battle Co-authored-by: Krzysztof Kotlarek <kotlarek.krzysztof@gmail.com>	2023-05-25 11:05:02 +10:00
Rafael dos Santos Silva	cfc6e388df	FIX: Ensure embeddings database outages are handled gracefully (#80 ) The rails_failover middleware will intercept all `PG::ConnectionBad` errors and put the cluster into readonly mode. It does not have any handling for multiple databases. Therefore, an issue with the embeddings database was taking the whole cluster into readonly. This commit fixes the issue by rescuing `PG::Error` from all AI database accesses, and re-raises errors with a different class. It also adds a spec to ensure that an embeddings database outage does not affect the functionality of the topics/show route. Co-authored-by: David Taylor <david@taylorhq.com>	2023-05-23 22:57:52 +01:00
Rafael dos Santos Silva	b213fe7f94	FIX: Give up trying to reuse the DB connection and rely on pgbouncer (#79 )	2023-05-23 15:12:59 -03:00
Sam	d85b503ed4	FIX: guide GPT 3.5 better (#77 ) * FIX: guide GPT 3.5 better This limits search results to 10 cause we were blowing the whole token budget on search results, additionally it includes a quick exchange at the start of a session to try and guide GPT 3.5 to follow instructions Sadly GPT 3.5 drifts off very quickly but this does improve stuff a bit. It also attempts to correct some issues with anthropic, though it still is surprisingly hard to ground * add status:public, this is a bit of a hack but ensures that we can search for any filter provided * fix specs	2023-05-23 23:08:17 +10:00
Sam	b82fc1e692	FIX: ensure we only attempt embedding once every 15 minutes (#76 ) This also heavily reduced log noise and ensures our exception handling is more surgical.	2023-05-23 10:43:24 +10:00
Sam	074d00ca32	FEATURE: improve search prompt (#75 ) - We only support searching public topics - make it clear - Stop using bug/feature, cause is poisons system - these may not exist - Add after: and before: which are very handy for bounding search results	2023-05-23 07:52:14 +10:00
Sam	e0cf7b7d70	FIX: results will be nil for invalid queries (#74 ) Previous to this change invalid searches would break the command.	2023-05-22 15:14:26 +10:00
Sam	92fb84e24d	iterate commands (#73 ) * FEATURE: introduce a more efficient formatter Previous formatting style was space inefficient given JSON consumes lots of tokens, the new format is now used consistently across commands Also fixes - search limited to 10 - search breaking on limit: non existent directive * Slight improvement to summarizer Stop blowing up context with custom prompts * ensure we include the guiding message * correct spec * langchain style summarizer ... much more accurate (albeit more expensive) * lint	2023-05-22 12:09:14 +10:00
Sam	d59ed1091b	FEATURE: add support for GPT <-> Forum integration This change-set connects GPT based chat with the forum it runs on. Allowing it to perform search, lookup tags and categories and summarize topics. The integration is currently restricted to public portions of the forum. Changes made: - Do not run ai reply job for small actions - Improved composable system prompt - Trivial summarizer for topics - Image generator - Google command for searching via Google - Corrected trimming of posts raw (was replacing with numbers) - Bypass of problem specs The feature works best with GPT-4 --------- Co-authored-by: Roman Rizzi <rizziromanalejandro@gmail.com>	2023-05-20 17:45:54 +10:00
Rafael dos Santos Silva	262ed4753e	FEATURE: Basic StableDiffusion text2img support (#72 )	2023-05-20 09:38:08 +10:00
Rafael dos Santos Silva	739b314312	Fixes for embeddings and truncate (#67 )	2023-05-18 09:21:28 +10:00
Rafael dos Santos Silva	e9ae28f773	FIX: Non instructor OSS embeddings was broken (#65 )	2023-05-17 12:10:10 -03:00
Roman Rizzi	362f6167d1	FEATURE: Less friction for starting a conversation with an AI bot. (#63 ) * FEATURE: Less friction for starting a conversation with an AI bot. This PR adds a new header icon as a shortcut to start a conversation with one of our AI Bots. After clicking and selecting one from the dropdown menu, we'll open the composer with some fields already filled (recipients and title). If you leave the title as is, we'll queue a job after five minutes to update it using a bot suggestion. * Update assets/javascripts/initializers/ai-bot-replies.js Co-authored-by: Rafael dos Santos Silva <xfalcox@gmail.com> * Update assets/javascripts/initializers/ai-bot-replies.js Co-authored-by: Rafael dos Santos Silva <xfalcox@gmail.com> --------- Co-authored-by: Rafael dos Santos Silva <xfalcox@gmail.com>	2023-05-16 14:38:21 -03:00
Rafael dos Santos Silva	2ed1f874c2	Use correct API signature for instructor embeddings (#62 )	2023-05-15 17:18:11 -03:00
Rafael dos Santos Silva	3c9513e754	Refinements to embeddings and tokenizers (#61 ) * Refinements to embeddings and tokenizers * lint * Truncate with tokenizers for summary * fix	2023-05-15 15:10:42 -03:00
Rafael dos Santos Silva	97124b30de	FEATURE: Update summarization token count and add Claude 100k (#58 )	2023-05-11 15:35:58 -03:00
Rafael dos Santos Silva	66bf4c74c6	FEATURE: Handle invalid media in NSFW module (#57 ) * FEATURE: Handle invalid media in NSFW module * fix lint	2023-05-11 15:35:39 -03:00
Roman Rizzi	7e3cb0ea16	FEATURE: Multi-model support for the AI Bot module. (#56 ) We'll create one bot user for each available model. When listed in the `ai_bot_enabled_chat_bots` setting, they will reply. This PR lets us use Claude-v1 in stream mode.	2023-05-11 10:03:03 -03:00
Rafael dos Santos Silva	e5537d4c77	FEATURE: Allow excluding closed topics from semantic related (#55 )	2023-05-09 15:30:50 -03:00
Rafael dos Santos Silva	f1133f66a6	Updates to embedding rake tasks (#54 ) - Creates embeddings in topic ID order, so it's easier to stop and restart from where we stopped - Update index parameters with current best practices	2023-05-09 13:45:16 -03:00
Sam	e76fc77189	fixes (#53 ) * Minor... use username suggester in case username already exists * FIX: ensure we truncate long prompts Previously we 1. Used raw length instead of token counts for counting length 2. We totally dropped a prompt if it was too long New implementation will truncate "raw" if it gets too long maintaining meaning.	2023-05-06 07:31:53 -03:00
Roman Rizzi	71b105a1bb	FEATURE: Introduce the ai-bot module (#52 ) This module lets you chat with our GPT bot inside a PM. The bot only replies to members of the groups listed on the ai_bot_allowed_groups setting and only if you invite it to participate in the PM.	2023-05-05 15:28:31 -03:00
Rafael dos Santos Silva	c96edc8a72	FIX: Pass correct API Key to summarization service (#50 )	2023-05-02 21:41:11 -03:00
Rafael dos Santos Silva	89ac5d720a	FIX: Only send supported image types for classification (#49 ) * FIX: Only send supported image types for classification	2023-04-27 17:52:20 -03:00
Sam	2cd60a4b3b	FEATURE: add a table to audit OpenAI usage (#45 ) Still need to build a job to purge logs	2023-04-26 11:44:29 +10:00
David Taylor	a0542d1859	DEV: Resolve add_to_serializer deprecations (#46 ) `26b7f8a63b`	2023-04-24 16:07:17 +01:00
Sam	057fbe1ce6	FEATURE: add internal support for streaming mode (#42 ) Also adds some tests around completions and supports additional params such as top_p, temperature and max_tokens This also migrates off Faraday to using Net::HTTP directly	2023-04-21 16:54:25 +10:00
Meghna	14b21b4f4d	UX: add a custom sparkles icon for AI action buttons (#44 )	2023-04-20 20:41:24 +05:30
Roman Rizzi	38e007a3a5	FEATURE: Topic summarization (#41 ) * FEATURE: Topic summarization Summarize topics using the TopicView's "summary" filter. The UI is similar to what we do for chat, but we don't allow the user to select a timeframe. Co-authored-by: Rafael dos Santos Silva <xfalcox@gmail.com>	2023-04-19 17:57:31 -03:00
Rafael dos Santos Silva	9783e3b025	FEATURE: Add a basic tokenizer API (#37 ) * FEATURE: Add a basic tokenizer API * Add tests * lint	2023-04-19 11:55:59 -03:00
Rafael dos Santos Silva	4368ef29d8	FIX: Sometimes Claude sends all titles suggestions in a single ai tag (#40 )	2023-04-10 16:02:44 -03:00
Rafael dos Santos Silva	bb0b829634	FEATURE: Anthropic Claude for AIHelper and Summarization modules (#39 )	2023-04-10 11:04:42 -03:00
Rafael dos Santos Silva	5549e4d5b3	FEATURE: Chat channel summarization. (#32 ) * start summary module * chat channel summarization * FEATURE: modal for channel summarization --------- Co-authored-by: Roman Rizzi <rizziromanalejandro@gmail.com>	2023-04-04 11:24:09 -03:00
Roman Rizzi	7a54455cf6	FIX: Use correct variable and method for embeddings (#35 )	2023-03-31 16:15:10 -03:00
Roman Rizzi	4e05763a99	FEATURE: Semantic assymetric full-page search (#34 ) Depends on discourse/discourse#20915 Hooks to the full-page-search component using an experimental API and performs an assymetric similarity search using our embeddings database.	2023-03-31 15:29:56 -03:00
Sam	6543c50758	FIX: stop returning self as a candidate for related topics (#31 )	2023-03-31 11:04:17 +10:00
Sam	0d80d9ec49	FEATURE: allow limiting results in related topics section (#30 ) Also: - Normalizes behavior between logged in and anon, we only show related topics in the related topic section - Renames "suggested" to "related" given this only exists in related section - Adds a spec section to ensure anon does not regress - Adds `ai_embeddings_semantic_related_topics` to limit related topics Renamed settings: ai_embeddings_semantic_suggested_model -> ai_embeddings_semantic_related_model ai_embeddings_semantic_suggested_topics_enabled -> ai_embeddings_semantic_related_topics_enabled Plugins is still in an experimental phase and not much is overidden hence avoiding adding site setting migrations. Co-authored-by: Krzysztof Kotlarek <kotlarek.krzysztof@gmail.com>	2023-03-31 11:04:34 +11:00
Sam	1d097b9d82	FEATURE: attempt to include related topics above suggested (#28 ) Allows related topics to show up for logged on users - Introduces a new "Related Topics" block above suggested when related topics exist - Renames `ai_embeddings_semantic_suggested_topics_anons_enabled` -> `ai_embeddings_semantic_suggested_topics_enabled` (given it is only deployed on 1 site not bothering with a migration) - Adds an integration test to ensure data arrives correctly on the client	2023-03-31 09:07:22 +11:00
Rafael dos Santos Silva	b942a18298	FEATURE: Support for GPT-4 in AI Helper module (#29 )	2023-03-28 23:22:34 -03:00
Rafael dos Santos Silva	45950f1bb4	FIX: Only show public visible topics as suggested for anons (#27 ) * FIX: Only show public visible topics as suggested for anons * DEV: Add tests for embeddings * Update spec/lib/modules/embeddings/semantic_suggested_spec.rb Co-authored-by: Bianca Nenciu <nbianca@users.noreply.github.com> * Update spec/lib/modules/embeddings/semantic_suggested_spec.rb Co-authored-by: Bianca Nenciu <nbianca@users.noreply.github.com> * move to top --------- Co-authored-by: Bianca Nenciu <nbianca@users.noreply.github.com>	2023-03-23 17:28:01 -03:00
Roman Rizzi	4c960970fa	DEV: Log information about errors from the completions OpenAI API (#26 )	2023-03-22 16:00:28 -03:00
Sam	1d14f7ffaf	FEATURE: Add a markdown table AI helper (#25 )	2023-03-22 13:16:29 -03:00
Rafael dos Santos Silva	bd342f538d	FEATURE: Try to generate embeddings for a topic when those aren't found (#23 )	2023-03-21 18:20:46 -03:00
Roman Rizzi	39f7f1f29e	FEATURE: Prompts can consist of multiple messages. (#21 ) A prompt with multiple messages leads to better results, as the AI can learn for given examples. Alongside this change, we provide a better default proofreading prompt.	2023-03-21 12:04:59 -03:00
Rafael dos Santos Silva	6bdbc0e32d	FIX: Proper flow when a topic doesn't have embeddings (#20 )	2023-03-20 16:44:55 -03:00
Roman Rizzi	fea9041ee1	DEV: Use 10s timeout when using the completions API (#19 )	2023-03-20 16:43:51 -03:00
Roman Rizzi	320ac6e84b	REFACTOR: Store prompts in a dedicated table. (#14 ) This change makes it easier to add new prompts to our AI helper. We don't have a UI for it yet. You'll have to do it through a console.	2023-03-17 15:14:19 -03:00
Joffrey JAFFEUX	edfdc6dfae	DEV: applies chat namespacing (#12 )	2023-03-17 15:15:38 +01:00
Roman Rizzi	75aa595105	FIX: Use cooked suggestion when generating a diff (#13 )	2023-03-16 11:09:28 -03:00
Rafael dos Santos Silva	80d662e9e8	FEATURE: Semantic Suggested Topics (#10 )	2023-03-15 17:21:45 -03:00
Roman Rizzi	f99fe7e1ed	FEATURE: Composer AI helper (#8 ) * FEATURE: Composer AI helper This change introduces a new composer button for the group members listed in the `ai_helper_allowed_groups` site setting. Users can use chatGPT to review, improve, or translate their posts to English. * Add a safeguard for PMs and don't rely on parentView	2023-03-15 17:02:20 -03:00
Roman Rizzi	aa2fca6086	DEV: DiscourseAI -> DiscourseAi rename to have consistent folders and files (#9 )	2023-03-14 16:03:50 -03:00
Rafael dos Santos Silva	510c6487e3	DEV: Preparation work for multiple inference providers (#5 )	2023-03-07 16:14:39 -03:00
Roman Rizzi	a838116cd5	FEATURE: Use dedicated reviewables for AI flags. (#4 ) This change adds two new reviewable types: ReviewableAIPost and ReviewableAIChatMessage. They have the same actions as their existing counterparts: ReviewableFlaggedPost and ReviewableChatMessage. We'll display the model used and their accuracy when showing these flags in the review queue and adjust the latter after staff performs an action, tracking a global accuracy per existing model in a separate table. * FEATURE: Dedicated reviewables for AI flags * Store and adjust model accuracy * Display accuracy in reviewable templates	2023-03-07 15:39:28 -03:00
Roman Rizzi	676d3ce6b2	DEV: Rename XClassification --> XClassificator to make it more obvious (#3 )	2023-02-28 11:17:03 -03:00
Roman Rizzi	b9a650fde4	DEV: Dedicated table for saving classification results (#1 )	2023-02-27 16:21:40 -03:00
Roman Rizzi	5f9597474c	REFACTOR: Streamline flag and classification process	2023-02-24 13:25:02 -03:00
Roman Rizzi	85768cfb1c	FEATURE: Classify posts looking for NSFW images	2023-02-24 09:11:58 -03:00
Roman Rizzi	94933f3c58	DEV: Add missing specs for the toxicity module	2023-02-24 07:53:43 -03:00
Roman Rizzi	e8bffcdd64	DEV: Add tests for the sentiment module	2023-02-23 15:50:10 -03:00
Roman Rizzi	ef6c785aca	DEV: Move jobs undear each module lib directory	2023-02-23 14:09:52 -03:00
Roman Rizzi	1afa274b99	DEV: Reorganize files and add an entry point for each module	2023-02-23 12:25:00 -03:00
Rafael dos Santos Silva	a73931c151	Refactoring of nsfw and flagger	2023-02-23 12:13:26 -03:00
Roman Rizzi	6f0c141062	FEATURE: Introduce NSFW content detection basic flow.	2023-02-23 11:08:34 -03:00
Rafael dos Santos Silva	f572a7cc2c	fix lint	2023-02-22 20:48:51 -03:00
Rafael dos Santos Silva	6cf411ec90	add toxicity and sentiment modules	2023-02-22 20:46:53 -03:00

1 2 3 4

188 Commits