discourse-ai

mirror of https://github.com/discourse/discourse-ai.git synced 2025-06-29 11:02:17 +00:00

Author	SHA1	Message	Date
Sam	471f96f972	FEATURE: allow seeing configured LLM on feature page (#1460 ) This is an interim fix so we can at least tell what feature is being used for what LLM. It also adds some test coverage to the feature page.	2025-06-24 17:42:47 +10:00
Keegan George	baaa3d199a	FIX: streaming related specs (#1448 ) ## 🔍 Overview This update fixes an issue where message bus streaming related specs were not working correctly. To do so we pass the `last_id` when subscribing to `MessageBus` which allows us to unskip those broken tests. --------- Co-authored-by: Joffrey JAFFEUX <j.jaffeux@gmail.com>	2025-06-19 07:41:18 -07:00
Joffrey JAFFEUX	6a33e5154d	DEV: makes ai menu helper a standalone menu (#1434 ) The current menu was rendering inside the post text toolbar (on desktop). This is not ideal as the post text toolbar rendering is conditioned on the presence of text selection, when you click a button on the toolbar, by design of the web browsers you will lose your text selection, making all of this super tricky. This commit makes desktop and mobile behave in the same way by rendering their own menu and capturing the quote state when we render the post text selection toolbar, this allows us to reason a much simpler way about the AI helper. This commit also removes what appears to be an unused file and corrects which was seemingly copy/paste mistakes. ⚠️ Technical note, this commit is correcting the message bus subscription which amongst other things allows to write specs which are not flaky. However due to the current implementation we have a channel per post, which means we need to serialize on last message bus id per post. We have two possible solutions here: - subscribe at the topic level - refactor the code to be able to use `MessageBus.last_ids` to be able to grab multiple posts at once instead of having to call `MessageBus.last_id` and done one Redis call per post --------- Co-authored-by: Keegan George <kgeorge13@gmail.com>	2025-06-19 11:56:00 +02:00
Rafael dos Santos Silva	bc8e57d7e8	DEV: Move title suggestion to an array (#1435 )	2025-06-16 18:06:54 -03:00
Roman Rizzi	9b7f1e6ee9	FIX: Helper wasn't working when the persona doesn't use structured output (#1433 )	2025-06-12 12:33:12 -03:00
Sam	6817866de9	FEATURE: allow access to assigns from forum researcher (#1412 ) * FEATURE: allow access to assigns from forum researcher * FIX: should properly be checking for empty * finish PR	2025-06-06 16:59:00 +10:00
Roman Rizzi	c885e5697f	review feedback	2025-06-04 14:23:00 -03:00
Roman Rizzi	0338dbea23	FEATURE: Use different personas to power AI helper features. You can now edit each AI helper prompt individually through personas, limit access to specific groups, set different LLMs, etc.	2025-06-04 14:23:00 -03:00
Sam	cf220c530c	FIX: Improve MessageBus efficiency and correctly stop streaming (#1362 ) * FIX: Improve MessageBus efficiency and correctly stop streaming This commit enhances the message bus implementation for AI helper streaming by: - Adding client_id targeting for message bus publications to ensure only the requesting client receives streaming updates - Limiting MessageBus backlog size (2) and age (60 seconds) to prevent Redis bloat - Replacing clearTimeout with Ember's cancel method for proper runloop management, we were leaking a stop - Adding tests for client-specific message delivery These changes improve memory usage and make streaming more reliable by ensuring messages are properly directed to the requesting client. * composer suggestion needed a fix as well. * backlog size of 2 is risky here cause same channel name is reused between clients	2025-05-23 16:23:06 +10:00
Keegan George	9ee82fd8be	DEV: Temporarily suppress diff animation as we fix issues (#1341 ) The diff animation introduced in https://github.com/discourse/discourse-ai/pull/1332 and with attempts to improve it in https://github.com/discourse/discourse-ai/pull/1338 still has various issues. As we work on a fix, we want to revert the animation to simply stream the diff without animation so users are not left with a janky unusable experience.	2025-05-15 14:55:30 -07:00
Keegan George	dfea784fc4	DEV: Improve diff streaming accuracy with safety checker (#1338 ) This update adds a safety checker which scans the streamed updates. It ensures that incomplete segments of text are not sent yet over message bus as this will cause breakage with the diff streamer. It also updates the diff streamer to handle a thinking state for when we are waiting for message bus updates.	2025-05-15 11:38:46 -07:00
Sam	17f04c76d8	FEATURE: add OpenAI image generation and editing capabilities (#1293 ) This commit enhances the AI image generation functionality by adding support for: 1. OpenAI's GPT-based image generation model (gpt-image-1) 2. Image editing capabilities through the OpenAI API 3. A new "Designer" persona specialized in image generation and editing 4. Two new AI tools: CreateImage and EditImage Technical changes include: - Renaming `ai_openai_dall_e_3_url` to `ai_openai_image_generation_url` with a migration - Adding `ai_openai_image_edit_url` setting for the image edit API endpoint - Refactoring image generation code to handle both DALL-E and the newer GPT models - Supporting multipart/form-data for image editing requests * wild guess but maybe quantization is breaking the test sometimes this increases distance * Update lib/personas/designer.rb Co-authored-by: Alan Guo Xiang Tan <gxtan1990@gmail.com> * simplify and de-flake code * fix, in chat we need enough context so we know exactly what uploads a user uploaded. * Update lib/personas/tools/edit_image.rb Co-authored-by: Alan Guo Xiang Tan <gxtan1990@gmail.com> * cleanup downloaded files right away * fix implementation --------- Co-authored-by: Alan Guo Xiang Tan <gxtan1990@gmail.com>	2025-04-29 17:38:54 +10:00
Keegan George	1300cc8a36	FEATURE: Add streaming to composer helper (#1256 ) This update adding streaming to the AI helper inside the composer.	2025-04-14 08:18:50 -07:00
Sam	5b6d39a206	FEATURE: flexible image handling within messages (#1214 ) * DEV: refactor bot internals This introduces a proper object for bot context, this makes it simpler to improve context management as we go cause we have a nice object to work with Starts refactoring allowing for a single message to have multiple uploads throughout * transplant method to message builder * chipping away at inline uploads * image support is improved but not fully fixed yet partially working in anthropic, still got quite a few dialects to go * open ai and claude are now working * Gemini is now working as well * fix nova * more dialects... * fix ollama * fix specs * update artifact fixed * more tests * spam scanner * pass more specs * bunch of specs improved * more bug fixes. * all the rest of the tests are working * improve tests coverage and ensure custom tools are aware of new context object * tests are working, but we need more tests * resolve merge conflict * new preamble and expanded specs on ai tool * remove concept of "standalone tools" This is no longer needed, we can set custom raw, tool details are injected into tool calls	2025-03-31 12:39:07 -03:00
Keegan George	6aaf8a0619	DEV: Use existing topic embeddings when suggesting tags/categories on edit (#1189 ) When editing a topic (instead of creating one) and using the tag/category suggestion buttons. We want to use existing topic embeddings instead of creating new ones.	2025-03-12 18:52:07 -07:00
Roman Rizzi	6765a13a40	FEATURE: Experimental search results from an AI Persona. (#1139 ) * FEATURE: Experimental search results from an AI Persona. When a user searches discourse, we'll send the query to an AI Persona to provide additional context and enrich the results. The feature depends on the user being a member of a group to which the persona has access. * Update assets/stylesheets/common/ai-blinking-animation.scss Co-authored-by: Keegan George <kgeorge13@gmail.com> --------- Co-authored-by: Keegan George <kgeorge13@gmail.com>	2025-02-20 14:37:58 -03:00
Rafael dos Santos Silva	37bf160d26	FIX: Add workaround to pgvector HNSW search limitations (#1133 ) From [pgvector/pgvector](https://github.com/pgvector/pgvector) README > With approximate indexes, filtering is applied after the index is scanned. If a condition matches 10% of rows, with HNSW and the default hnsw.ef_search of 40, only 4 rows will match on average. For more rows, increase hnsw.ef_search. > > Starting with 0.8.0, you can enable [iterative index scans](https://github.com/pgvector/pgvector#iterative-index-scans), which will automatically scan more of the index when needed. Since we are stuck on 0.7.0 we are going the first option for now.	2025-02-19 16:30:01 -03:00
Sam	5e80f93e4c	FEATURE: PDF support for rag pipeline (#1118 ) This PR introduces several enhancements and refactorings to the AI Persona and RAG (Retrieval-Augmented Generation) functionalities within the discourse-ai plugin. Here's a breakdown of the changes: 1. LLM Model Association for RAG and Personas: - New Database Columns: Adds `rag_llm_model_id` to both `ai_personas` and `ai_tools` tables. This allows specifying a dedicated LLM for RAG indexing, separate from the persona's primary LLM. Adds `default_llm_id` and `question_consolidator_llm_id` to `ai_personas`. - Migration: Includes a migration (`20250210032345_migrate_persona_to_llm_model_id.rb`) to populate the new `default_llm_id` and `question_consolidator_llm_id` columns in `ai_personas` based on the existing `default_llm` and `question_consolidator_llm` string columns, and a post migration to remove the latter. - Model Changes: The `AiPersona` and `AiTool` models now `belong_to` an `LlmModel` via `rag_llm_model_id`. The `LlmModel.proxy` method now accepts an `LlmModel` instance instead of just an identifier. `AiPersona` now has `default_llm_id` and `question_consolidator_llm_id` attributes. - UI Updates: The AI Persona and AI Tool editors in the admin panel now allow selecting an LLM for RAG indexing (if PDF/image support is enabled). The RAG options component displays an LLM selector. - Serialization: The serializers (`AiCustomToolSerializer`, `AiCustomToolListSerializer`, `LocalizedAiPersonaSerializer`) have been updated to include the new `rag_llm_model_id`, `default_llm_id` and `question_consolidator_llm_id` attributes. 2. PDF and Image Support for RAG: - Site Setting: Introduces a new hidden site setting, `ai_rag_pdf_images_enabled`, to control whether PDF and image files can be indexed for RAG. This defaults to `false`. - File Upload Validation: The `RagDocumentFragmentsController` now checks the `ai_rag_pdf_images_enabled` setting and allows PDF, PNG, JPG, and JPEG files if enabled. Error handling is included for cases where PDF/image indexing is attempted with the setting disabled. - PDF Processing: Adds a new utility class, `DiscourseAi::Utils::PdfToImages`, which uses ImageMagick (`magick`) to convert PDF pages into individual PNG images. A maximum PDF size and conversion timeout are enforced. - Image Processing: A new utility class, `DiscourseAi::Utils::ImageToText`, is included to handle OCR for the images and PDFs. - RAG Digestion Job: The `DigestRagUpload` job now handles PDF and image uploads. It uses `PdfToImages` and `ImageToText` to extract text and create document fragments. - UI Updates: The RAG uploader component now accepts PDF and image file types if `ai_rag_pdf_images_enabled` is true. The UI text is adjusted to indicate supported file types. 3. Refactoring and Improvements: - LLM Enumeration: The `DiscourseAi::Configuration::LlmEnumerator` now provides a `values_for_serialization` method, which returns a simplified array of LLM data (id, name, vision_enabled) suitable for use in serializers. This avoids exposing unnecessary details to the frontend. - AI Helper: The `AiHelper::Assistant` now takes optional `helper_llm` and `image_caption_llm` parameters in its constructor, allowing for greater flexibility. - Bot and Persona Updates: Several updates were made across the codebase, changing the string based association to a LLM to the new model based. - Audit Logs: The `DiscourseAi::Completions::Endpoints::Base` now formats raw request payloads as pretty JSON for easier auditing. - Eval Script: An evaluation script is included. 4. Testing: - The PR introduces a new eval system for LLMs, this allows us to test how functionality works across various LLM providers. This lives in `/evals`	2025-02-14 12:15:07 +11:00
Roman Rizzi	e52045ebdc	DEV: Robust check for embeddings enabled (#1116 )	2025-02-06 12:18:55 -03:00
Roman Rizzi	f5cf1019fb	FEATURE: configurable embeddings (#1049 ) * Use AR model for embeddings features * endpoints * Embeddings CRUD UI * Add presets. Hide a couple more settings * system specs * Seed embedding definition from old settings * Generate search bit index on the fly. cleanup orphaned data * support for seeded models * Fix run test for new embedding * fix selected model not set correctly	2025-01-21 12:23:19 -03:00
Sam	11d0f60f1e	FEATURE: smart date support for AI helper (#1044 ) * FEATURE: smart date support for AI helper This feature allows conversion of human typed in dates and times to smart "Discourse" timezone friendly dates. * fix specs and lint * lint * address feedback * add specs	2024-12-31 08:04:25 +11:00
Rafael dos Santos Silva	792df58fbc	FIX: AI Helper category / tag suggestion when user does not categories muted (#1042 )	2024-12-23 15:58:26 -03:00
Roman Rizzi	534b0df391	REFACTOR: Separation of concerns for embedding generation. (#1027 ) In a previous refactor, we moved the responsibility of querying and storing embeddings into the `Schema` class. Now, it's time for embedding generation. The motivation behind these changes is to isolate vector characteristics in simple objects to later replace them with a DB-backed version, similar to what we did with LLM configs.	2024-12-16 09:55:39 -03:00
Roman Rizzi	eae527f99d	REFACTOR: A Simpler way of interacting with embeddings tables. (#1023 ) * REFACTOR: A Simpler way of interacting with embeddings' tables. This change adds a new abstraction called `Schema`, which acts as a repository that supports the same DB features `VectorRepresentation::Base` has, with the exception that removes the need to have duplicated methods per embeddings table. It is also a bit more flexible when performing a similarity search because you can pass it a block that gives you access to the builder, allowing you to add multiple joins/where conditions.	2024-12-13 10:15:21 -03:00
Keegan George	fc88bb08ab	FIX: Tag suggester is suggesting already assigned tags (#990 ) This PR fixes an issue where the tag suggester for edit title topic area was suggesting tags that are already assigned on a post. It also updates the amount of suggested tags to 7 so that there is still a decent amount of tags suggested when tags are already assigned.	2024-12-03 07:25:04 +11:00
Sam	0cb2c413ba	FEATURE: exclude muted categories from category suggester (#979 ) The logic here is that users do not particularly care about topics in the category so we can exclude them from tag and category suggestions	2024-11-29 12:17:28 +11:00
Rafael dos Santos Silva	80adefa1c1	FIX: Move count logic to the end for tag suggestions (#978 )	2024-11-28 16:27:38 -03:00
Keegan George	f1c7ee8624	DEV: Better control what prompts can appear in post/composer (#969 ) This PR updates the logic for the location map so it permits only the desired prompts through to the composer/post menu. Anything else won't be shown by default. This PR also adds relevant tests to prevent regression.	2024-11-27 16:14:21 -08:00
Keegan George	dabef02919	DEV: Prevent `detect_text_locale` from appearing in menus (#967 ) ### 🔍 Overview With the recent changes to allow DiscourseAi in the translator plugin, `detect_text_locale` was needed as a CompletionPrompt. However, it is leaking into composer/post helper menus. This PR ensures we don't not show it in those menus.	2024-11-28 09:27:08 +11:00
Keegan George	6b7d7c1179	REFACTOR: Helper suggestions (#914 ) This PR adds some updates to the Helper suggestions to improve it's functionality and modernize some of the codebase.	2024-11-27 12:21:03 -08:00
Rafael dos Santos Silva	aef9a03d4c	FEATURE: Truncate AI Captions to a reasonable max size (#907 )	2024-11-12 15:52:46 -03:00
Rafael dos Santos Silva	96f5f8cbd0	FIX: Basic cleanup of AI Caption to remove line breaks and pipes (#857 )	2024-10-23 18:38:29 -03:00
Kris	25cc03809a	DEV: ensure far-copy icon is included in subset (#841 )	2024-10-16 12:32:13 -04:00
Roman Rizzi	db5cbfb148	FIX: Bail earlier when a chat thread has no messages (#789 )	2024-08-30 17:17:14 -03:00
Roman Rizzi	c6aeabbfc0	FIX: Malformed message in systemless + inline img scenario (#771 )	2024-08-23 16:41:57 -03:00
Keegan George	f72ab12761	DEV: Clearly separate post/composer helper settings (#747 )	2024-08-12 15:40:23 -07:00
Keegan George	0be292f247	FIX: `auto_image_caption` not always present for current user. (#746 )	2024-08-08 13:37:26 -07:00
Roman Rizzi	5c196bca89	FEATURE: Track if a model can do vision in the llm_models table (#725 ) * FEATURE: Track if a model can do vision in the llm_models table * Data migration	2024-07-24 16:29:47 -03:00
PangBo	4ebbdc043e	fix: locale handling in assistant.rb (#705 )	2024-07-05 11:16:09 +02:00
Keegan George	eab2f74b58	DEV: Use site locale for composer helper translations (#698 )	2024-07-04 08:23:37 -07:00
Rafael dos Santos Silva	a708d4dfa2	FIX: Use base64 encoded images in AI Image Caption via LLaVa (#693 ) * FIX: Use base64 encoded images in AI Image Caption via LLaVa This fixed a regression introduced in #646 where we started sending schemaless URLs for our LLaVa service, which doesn't handle it well. Moving to base64 encoded images solves: - The service needing to download images Now the service running LLaVa doesn't need internet access - Secure uploads compat Every image is treated the same, less branching for secure uploads - Image Size problems Discourse is now responsible for ensure a max size for images - Troublesome dev env Previously to this commit you would need a dev env that was internet acessible to use llava image captions	2024-06-27 16:24:44 -03:00
Sam	b487de933d	FEATURE: add support for all vision models (#646 ) Previoulsy on GPT-4-vision was supported, change introduces support for Google/Anthropic and new OpenAI models Additionally this makes vision work properly in dev environments cause we sent the encoded payload via prompt vs sending urls	2024-05-28 10:31:15 -03:00
Loïc Guitaut	dd4e305ff7	DEV: Update rubocop-discourse to version 3.8.0 (#641 )	2024-05-28 11:15:42 +02:00
Keegan George	dae9d6f14e	FIX: Move image caption group check logic to server side (#645 ) * FIX: User groups error before initialization * DEV: Move group check logic to server side	2024-05-28 10:29:11 +10:00
Keegan George	a1c649965f	FEATURE: Auto image captions (#637 )	2024-05-27 10:49:24 -07:00
Sam	8eee6893d6	FEATURE: GPT4o support and better auditing (#618 ) - Introduce new support for GPT4o (automation / bot / summary / helper) - Properly account for token counts on OpenAI models - Track feature that was used when generating AI completions - Remove custom llm support for summarization as we need better interfaces to control registration and de-registration	2024-05-14 13:28:46 +10:00
Sam	85734fef52	FIX: properly cache user locale (#593 ) This blob is localized according to user locale, so we can end up bleeding incorrect data in the cache	2024-04-26 09:28:35 -03:00
Rafael dos Santos Silva	595cde0fd6	FIX: Users with empty locales would error out during prompt localization (#584 )	2024-04-22 13:55:10 -03:00
Sam	bd6f5caeac	FEATURE: Stable diffusion 3 support (#582 ) - Adds support for sd3 and sd3 turbo models - this requires new endpoints - Adds a hack to normalize arrays in the tool calls - Removes some leftover code - Adds support for aspect ratio as well so you can generate wide or tall images	2024-04-19 18:08:16 +10:00
Rafael dos Santos Silva	176a4458f2	FIX: Prevent AI chat thread titles from being created before replies are posted (#517 ) Chat thread replies draft trigger the thread_created event, which we relied on to trigger the AI generated title. Because of that we now will use the noisier chat_message_created event, and manually check for thread and replies existence. See https://github.com/discourse/discourse/pull/26033	2024-03-07 16:14:17 -03:00

1 2

69 Commits