discourse-ai

mirror of https://github.com/discourse/discourse-ai.git synced 2025-03-01 14:59:22 +00:00

Author	SHA1	Message	Date
Sam	fe19133dd4	FEATURE: full support for Sonnet 3.7 (#1151 ) * FEATURE: full support for Sonnet 3.7 - Adds support for Sonnet 3.7 with reasoning on bedrock and anthropic - Fixes regression where provider params were not populated Note. reasoning tokens are hardcoded to minimum of 100 maximum of 65536 * FIX: open ai non reasoning models need to use deprecate max_tokens	2025-02-25 17:32:12 +11:00
Sam	84e791a941	FIX: legacy reasoning models not working, missing provider params (#1149 ) * FIX: legacy reasoning models not working, missing provider params 1. Legacy reasoning models (o1-preview / o1-mini) do not support developer or system messages, do not use them. 2. LLM editor form not showing all provider params due to missing remap * add system test	2025-02-24 16:38:23 +11:00
Natalie Tay	2486e0e2dd	DEV: Extract configs to a yml file and allow local config (#1142 )	2025-02-24 16:22:19 +11:00
Keegan George	08377bab35	DEV: Sentiment analysis report follow-up updates (#1145 ) * DEV: make include subcategories checkbox operational * DEV: add pagination for post requests * WIP: selected chart UX improvements * DEV: Functional sentiment filters * DEV: Reset filters after going back * DEV: Add category colors, improve UX * DEV: Update spec	2025-02-24 16:21:10 +11:00
Discourse Translator Bot	43cbb7f45f	Update translations (#1148 )	2025-02-24 16:20:25 +11:00
Jarek Radosz	4f94675a72	DEV: Update license (#1147 )	2025-02-24 11:20:06 +08:00
Kris	55dde0a9e6	UX: minor adjustments to search bot (#1146 )	2025-02-21 19:40:53 -05:00
Roman Rizzi	09f4895302	UI: Custom icon for Discobot discoveries (#1144 )	2025-02-21 12:22:36 -03:00
Rafael dos Santos Silva	04678cab04	FIX: Discovery search would break normal search for anons (#1143 )	2025-02-21 12:14:47 -03:00
Roman Rizzi	f922012499	UX: Display a tooltip signalling this is an AI powered feature (#1141 )	2025-02-20 16:21:26 -03:00
Roman Rizzi	6765a13a40	FEATURE: Experimental search results from an AI Persona. (#1139 ) * FEATURE: Experimental search results from an AI Persona. When a user searches discourse, we'll send the query to an AI Persona to provide additional context and enrich the results. The feature depends on the user being a member of a group to which the persona has access. * Update assets/stylesheets/common/ai-blinking-animation.scss Co-authored-by: Keegan George <kgeorge13@gmail.com> --------- Co-authored-by: Keegan George <kgeorge13@gmail.com>	2025-02-20 14:37:58 -03:00
Keegan George	24f0e1262d	FEATURE: New sentiment analysis visualization report (#1109 ) ## 🔍 Overview This update adds a new report page at `admin/reports/sentiment_analysis` where admins can see a sentiment analysis report for the forum grouped by either category or tags. ## ➕ More details The report can breakdown either category or tags into positive/negative/neutral sentiments based on the grouping (category/tag). Clicking on the doughnut visualization will bring up a post list of all the posts that were involved in that classification with further sentiment classifications by post. The report can additionally be sorted in alphabetical order or by size, as well as be filtered by either category/tag based on the grouping. ## 👨🏽‍💻 Technical Details The new admin report is registered via the pluginAPi with `api.registerReportModeComponent` to register the custom sentiment doughnut report. However, when each doughnut visualization is clicked, a new endpoint found at: `/discourse-ai/sentiment/posts` is fetched to showcase posts classified by sentiments based on the respective params. ## 📸 Screenshots ![Screenshot 2025-02-14 at 11 11 35](https://github.com/user-attachments/assets/a63b5ab8-4fb2-477d-bd29-92545f44ff09)	2025-02-20 09:14:10 -08:00
Keegan George	1f9f330ce2	DEV: Add summarization type to eval (#1138 ) Adds `type: summarization` for topic summarization eval: https://github.com/discourse/discourse-ai-evals/pull/4	2025-02-20 09:07:23 -08:00
Roman Rizzi	70248ccfca	DEV: Update annotations for models using Core tables (#1140 )	2025-02-20 11:49:50 -03:00
Keegan George	af47873f28	FIX: hardcoded require for evals (#1137 ) The require for `discourse/config/environment` should point to the local user's core Discourse environment instead of the hardcoded path for Sam's env.	2025-02-19 11:56:52 -08:00
Rafael dos Santos Silva	37bf160d26	FIX: Add workaround to pgvector HNSW search limitations (#1133 ) From [pgvector/pgvector](https://github.com/pgvector/pgvector) README > With approximate indexes, filtering is applied after the index is scanned. If a condition matches 10% of rows, with HNSW and the default hnsw.ef_search of 40, only 4 rows will match on average. For more rows, increase hnsw.ef_search. > > Starting with 0.8.0, you can enable [iterative index scans](https://github.com/pgvector/pgvector#iterative-index-scans), which will automatically scan more of the index when needed. Since we are stuck on 0.7.0 we are going the first option for now.	2025-02-19 16:30:01 -03:00
Kris	3a755ca883	DEV: add summary button wrapper removed from core (#1136 )	2025-02-19 12:58:10 -05:00
Sam	12f00a62d2	FIX: use max_completion_tokens for open ai models (#1134 ) max_tokens is now deprecated per API	2025-02-19 15:49:15 +11:00
Sam	0c9466059c	DEV: improve artifact editing and eval system (#1130 ) - Add non-contiguous search/replace support using ... syntax - Add judge support for evaluating LLM outputs with ratings - Improve error handling and reporting in eval runner - Add full section replacement support without search blocks - Add fabricators and specs for artifact diffing - Track failed searches to improve debugging - Add JS syntax validation for artifact versions in eval system - Update prompt documentation with clear guidelines * improve eval output * move error handling * llm as a judge * fix spec * small note on evals	2025-02-19 15:44:33 +11:00
Keegan George	02f0908963	DEV: Misconfigured llm should go to edit page (#1132 )	2025-02-18 10:34:18 -08:00
Discourse Translator Bot	2c0a8e7f5c	Update translations (#1131 )	2025-02-18 14:51:56 +01:00
Martin Brennan	0582e75205	DEV: Skip PDF tests (#1129 ) CI=1 is not set on our internal build system, we need to fix that but for now let's unblock the build	2025-02-18 10:17:11 +10:00
Sam	ce79a18790	FEATURE: Native PDF support (#1127 ) * FEATURE: Native PDF support This amends it so we use PDF Reader gem to extract text from PDFs * This means that our simple pdf eval passes at last * fix spec * skip test in CI * test file support * Update lib/utils/image_to_text.rb Co-authored-by: Alan Guo Xiang Tan <gxtan1990@gmail.com> * address pr comments --------- Co-authored-by: Alan Guo Xiang Tan <gxtan1990@gmail.com>	2025-02-18 09:22:57 +11:00
Sam	9a6aec2cf6	DEV: eval support for tool calls (#1128 ) Also fixes anthropic with no params, streaming calls	2025-02-18 07:58:54 +11:00
Sam	5e80f93e4c	FEATURE: PDF support for rag pipeline (#1118 ) This PR introduces several enhancements and refactorings to the AI Persona and RAG (Retrieval-Augmented Generation) functionalities within the discourse-ai plugin. Here's a breakdown of the changes: 1. LLM Model Association for RAG and Personas: - New Database Columns: Adds `rag_llm_model_id` to both `ai_personas` and `ai_tools` tables. This allows specifying a dedicated LLM for RAG indexing, separate from the persona's primary LLM. Adds `default_llm_id` and `question_consolidator_llm_id` to `ai_personas`. - Migration: Includes a migration (`20250210032345_migrate_persona_to_llm_model_id.rb`) to populate the new `default_llm_id` and `question_consolidator_llm_id` columns in `ai_personas` based on the existing `default_llm` and `question_consolidator_llm` string columns, and a post migration to remove the latter. - Model Changes: The `AiPersona` and `AiTool` models now `belong_to` an `LlmModel` via `rag_llm_model_id`. The `LlmModel.proxy` method now accepts an `LlmModel` instance instead of just an identifier. `AiPersona` now has `default_llm_id` and `question_consolidator_llm_id` attributes. - UI Updates: The AI Persona and AI Tool editors in the admin panel now allow selecting an LLM for RAG indexing (if PDF/image support is enabled). The RAG options component displays an LLM selector. - Serialization: The serializers (`AiCustomToolSerializer`, `AiCustomToolListSerializer`, `LocalizedAiPersonaSerializer`) have been updated to include the new `rag_llm_model_id`, `default_llm_id` and `question_consolidator_llm_id` attributes. 2. PDF and Image Support for RAG: - Site Setting: Introduces a new hidden site setting, `ai_rag_pdf_images_enabled`, to control whether PDF and image files can be indexed for RAG. This defaults to `false`. - File Upload Validation: The `RagDocumentFragmentsController` now checks the `ai_rag_pdf_images_enabled` setting and allows PDF, PNG, JPG, and JPEG files if enabled. Error handling is included for cases where PDF/image indexing is attempted with the setting disabled. - PDF Processing: Adds a new utility class, `DiscourseAi::Utils::PdfToImages`, which uses ImageMagick (`magick`) to convert PDF pages into individual PNG images. A maximum PDF size and conversion timeout are enforced. - Image Processing: A new utility class, `DiscourseAi::Utils::ImageToText`, is included to handle OCR for the images and PDFs. - RAG Digestion Job: The `DigestRagUpload` job now handles PDF and image uploads. It uses `PdfToImages` and `ImageToText` to extract text and create document fragments. - UI Updates: The RAG uploader component now accepts PDF and image file types if `ai_rag_pdf_images_enabled` is true. The UI text is adjusted to indicate supported file types. 3. Refactoring and Improvements: - LLM Enumeration: The `DiscourseAi::Configuration::LlmEnumerator` now provides a `values_for_serialization` method, which returns a simplified array of LLM data (id, name, vision_enabled) suitable for use in serializers. This avoids exposing unnecessary details to the frontend. - AI Helper: The `AiHelper::Assistant` now takes optional `helper_llm` and `image_caption_llm` parameters in its constructor, allowing for greater flexibility. - Bot and Persona Updates: Several updates were made across the codebase, changing the string based association to a LLM to the new model based. - Audit Logs: The `DiscourseAi::Completions::Endpoints::Base` now formats raw request payloads as pretty JSON for easier auditing. - Eval Script: An evaluation script is included. 4. Testing: - The PR introduces a new eval system for LLMs, this allows us to test how functionality works across various LLM providers. This lives in `/evals`	2025-02-14 12:15:07 +11:00
Joffrey JAFFEUX	e2afbc26d3	FIX: correctly handle provider edit (#1125 ) Prior to this commit, editing the provider wouldn't recompute the provider params. It would also not correctly recompute the "canEditURL" property. To make possible this commit has: - made a fix in core: https://github.com/discourse/discourse/pull/31329 - ensures the provider params are recomputed when provider is changed - made the check on `canEditURL` based on form state and not initial model value Tests have been added to confirm the expected behavior.	2025-02-13 12:03:13 +01:00
dependabot[bot]	35f15629fb	Build(deps-dev): Bump rack from 3.1.6 to 3.1.10 (#1124 ) Bumps [rack](https://github.com/rack/rack) from 3.1.6 to 3.1.10. - [Release notes](https://github.com/rack/rack/releases) - [Changelog](https://github.com/rack/rack/blob/main/CHANGELOG.md) - [Commits](https://github.com/rack/rack/compare/v3.1.6...v3.1.10) --- updated-dependencies: - dependency-name: rack dependency-type: indirect ... Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2025-02-13 12:09:04 +11:00
David Battersby	1bfad17b9a	FIX: update draft key for new PM with AI bot (#1123 ) * FIX: update draft key for new PM with AI bot * allow multiple drafts with ai bot * fix linting	2025-02-13 12:08:31 +11:00
Rafael dos Santos Silva	77c6543c5d	FIX: Embeddings backfill job compat when transitioning models (#1122 ) When you already have embeddings for a model stored and change models, our backfill script was failing to backfill the newly configured model. Regression introduced most likely in 1686a8a	2025-02-12 10:37:45 -03:00
Rafael dos Santos Silva	708a3bd2b8	UX: Better tooltips for embeddings task instructions prefixes (#1121 )	2025-02-11 15:01:29 -03:00
Discourse Translator Bot	8c653fdb25	Update translations (#1120 )	2025-02-11 16:43:19 +01:00
Martin Brennan	7b1bdbde6d	FIX: Check post action creator result when flagging spam (#1119 ) Currently in core re-flagging something that is already flagged as spam is not supported, long term we may want to support this but in the meantime we should not be silencing/hiding if the PostActionCreator fails when flagging things as spam. --------- Co-authored-by: Ted Johansson <drenmi@gmail.com>	2025-02-11 13:29:27 +10:00
Hoa Nguyen	b60926c6e6	FEATURE: Tool name validation (#842 ) * FEATURE: Tool name validation - Add unique index to the name column of the ai_tools table - correct our tests for AiToolController - tool_name field which will be used to represent to LLM - Add tool_name to Tools's presets - Add duplicate tools validation for AiPersona - Add unique constraint to the name column of the ai_tools table * DEV: Validate duplicate tool_name between builin tools and custom tools * lint * chore: fix linting * fix conlict mistakes * chore: correct icon class * chore: fix failed specs * Add max_length to tool_name * chore: correct the option name * lintings * fix lintings	2025-02-07 14:34:47 +11:00
David Taylor	551f674c43	DEV: Bump dependencies and fix linting (#1115 )	2025-02-06 17:42:32 +01:00
Roman Rizzi	90bcb8b503	DEV: Build sentiment clients outside of promises (#1117 )	2025-02-06 13:11:10 -03:00
Roman Rizzi	e52045ebdc	DEV: Robust check for embeddings enabled (#1116 )	2025-02-06 12:18:55 -03:00
Kris	83c7919856	UX: clarify embeddings description (#1113 )	2025-02-06 08:50:01 -05:00
David Taylor	a996aa45bc	DEV: Pin version for Discourse <3.5.0.beta1-dev (#1114 )	2025-02-05 19:57:52 +01:00
Discourse Translator Bot	bdef136080	Update translations (#1112 )	2025-02-04 15:18:08 +01:00
Roman Rizzi	1b1b44353b	FEATURE: Changes to summaries' outdated logic. (#1108 ) Before this change, a summary was only outdated when new content appeared, for topics with "best replies", when the query returned different results. The intent behind this change is to detect when a summary is outdated as a result of an edit. Additionally, we are changing the backfill candidates query to compare "ai_summary_backfill_topic_max_age_days" against "last_posted_at" instead of "created_at", to catch long-lived, active topics. This was discussed here: https://meta.discourse.org/t/ai-summarization-backfill-is-stuck-keeps-regenerating-the-same-topic/347088/14?u=roman_rizzi	2025-02-04 09:31:11 -03:00
Joffrey JAFFEUX	d3b93f984d	UX: include none false for provider params (#1111 ) This change has been forgotten in `40e996b174`	2025-02-04 12:38:22 +01:00
Joffrey JAFFEUX	40e996b174	DEV: converts llm admin page to use form kit (#1099 ) This also converts the quota editor, and the quota modal.	2025-02-04 11:51:01 +01:00
Sam	43c56d7c92	FIX: need to be able to search replace within lines (#1110 ) (this is needed for very simple diffs and HTML)	2025-02-04 18:16:52 +11:00
Sam	a7d032fa28	DEV: artifact system update (#1096 ) ### Why This pull request fundamentally restructures how AI bots create and update web artifacts to address critical limitations in the previous approach: 1. Improved Artifact Context for LLMs: Previously, artifact creation and update tools included the entire artifact source code directly in the tool arguments. This overloaded the Language Model (LLM) with raw code, making it difficult for the LLM to maintain a clear understanding of the artifact's current state when applying changes. The LLM would struggle to differentiate between the base artifact and the requested modifications, leading to confusion and less effective updates. 2. Reduced Token Usage and History Bloat: Including the full artifact source code in every tool interaction was extremely token-inefficient. As conversations progressed, this redundant code in the history consumed a significant number of tokens unnecessarily. This not only increased costs but also diluted the context for the LLM with less relevant historical information. 3. Enabling Updates for Large Artifacts: The lack of a practical diff or targeted update mechanism made it nearly impossible to efficiently update larger web artifacts. Sending the entire source code for every minor change was both computationally expensive and prone to errors, effectively blocking the use of AI bots for meaningful modifications of complex artifacts. This pull request addresses these core issues by: * Introducing methods for the AI bot to explicitly read and understand the current state of an artifact. * Implementing efficient update strategies that send targeted changes rather than the entire artifact source code. * Providing options to control the level of artifact context included in LLM prompts, optimizing token usage. ### What The main changes implemented in this PR to resolve the above issues are: 1. `Read Artifact` Tool for Contextual Awareness: - A new `read_artifact` tool is introduced, enabling AI bots to fetch and process the current content of a web artifact from a given URL (local or external). - This provides the LLM with a clear and up-to-date representation of the artifact's HTML, CSS, and JavaScript, improving its understanding of the base to be modified. - By cloning local artifacts, it allows the bot to work with a fresh copy, further enhancing context and control. 2. Refactored `Update Artifact` Tool with Efficient Strategies: - The `update_artifact` tool is redesigned to employ more efficient update strategies, minimizing token usage and improving update precision: - `diff` strategy: Utilizes a search-and-replace diff algorithm to apply only the necessary, targeted changes to the artifact's code. This significantly reduces the amount of code sent to the LLM and focuses its attention on the specific modifications. - `full` strategy: Provides the option to replace the entire content sections (HTML, CSS, JavaScript) when a complete rewrite is required. - Tool options enhance the control over the update process: - `editor_llm`: Allows selection of a specific LLM for artifact updates, potentially optimizing for code editing tasks. - `update_algorithm`: Enables choosing between `diff` and `full` update strategies based on the nature of the required changes. - `do_not_echo_artifact`: Defaults to true, and by not echoing the artifact in prompts, it further reduces token consumption in scenarios where the LLM might not need the full artifact context for every update step (though effectiveness might be slightly reduced in certain update scenarios). 3. System and General Persona Tool Option Visibility and Customization: - Tool options, including those for system personas, are made visible and editable in the admin UI. This allows administrators to fine-tune the behavior of all personas and their tools, including setting specific LLMs or update algorithms. This was previously limited or hidden for system personas. 4. Centralized and Improved Content Security Policy (CSP) Management: - The CSP for AI artifacts is consolidated and made more maintainable through the `ALLOWED_CDN_SOURCES` constant. This improves code organization and future updates to the allowed CDN list, while maintaining the existing security posture. 5. Codebase Improvements: - Refactoring of diff utilities, introduction of strategy classes, enhanced error handling, new locales, and comprehensive testing all contribute to a more robust, efficient, and maintainable artifact management system. By addressing the issues of LLM context confusion, token inefficiency, and the limitations of updating large artifacts, this pull request significantly improves the practicality and effectiveness of AI bots in managing web artifacts within Discourse.	2025-02-04 16:27:27 +11:00
Roman Rizzi	7ad922331b	FIX: Make sure DiscoursePrometheus is installed when collecting metrics (#1107 )	2025-02-03 11:04:25 -03:00
Sam	cf86d274a0	FEATURE: improve o3-mini support (#1106 ) * DEV: raise timeout for reasoning LLMs * FIX: use id to identify llms, not model_name model_name is not unique, in the case of reasoning models you may configure the same llm multiple times using different reasoning levels.	2025-02-03 08:45:56 +11:00
Sam	381a2715c8	FEATURE: o3-mini supports (#1105 ) 1. Adds o3-mini presets 2. Adds support for reasoning effort 3. properly use "developer" messages for reasoning models	2025-02-01 14:08:34 +11:00
Rafael dos Santos Silva	8c22540e61	FEATURE: Use persona default LLM for Discord integration (#1104 )	2025-01-31 18:11:20 -03:00
Keegan George	13f9a1908a	DEV: Only permit config of allowed seeded models (#1103 ) This update resolves a regression that was introduced in https://github.com/discourse/discourse-ai/pull/1036/files. Previously, only seeded models that were allowed could be configured for model settings. However, in our attempts to prevent unreachable LLM errors from not allowing settings to persist, it also unknowingly allowed seeded models that were not allowed to be configured. This update resolves this issue, while maintaining the ability to still set unreachable LLMs.	2025-01-31 13:01:15 -08:00
Discourse Translator Bot	980f17a338	Update translations (#1097 )	2025-01-31 10:48:55 +01:00

1 2 3 4 5 ...

1112 Commits