discourse-ai

mirror of https://github.com/discourse/discourse-ai.git synced 2025-02-18 17:34:52 +00:00

Author	SHA1	Message	Date
Roman Rizzi	ed97827f49	FIX: Correctly display errors when parent module needs to be disabled first (#788 ) * FIX: Correctly display errors when parent module needs to be disabled first * Update spec/configuration/llm_validator_spec.rb Co-authored-by: Penar Musaraj <pmusaraj@gmail.com> --------- Co-authored-by: Penar Musaraj <pmusaraj@gmail.com>	2024-08-30 17:16:11 -03:00
Sam	584753cf60	FIX: we were never reindexing old content (#786 ) * FIX: we were never reindexing old content Embedding backfill contains logic for searching for old content change and then backfilling. Unfortunately it was excluding all topics that had embedding unconditionally, leading to no backfill ever happening. This change adds a test and ensures we backfill. * over select results, this ensures we will be more likely to find ai results when filtered	2024-08-30 14:37:55 +10:00
Sam	eee8e72756	FEATURE: API scope for semantic search (#785 ) The new API scope allows restricting access to semantic search only.	2024-08-30 09:35:20 +10:00
Sam	41054c4fb8	FEATURE: improve site setting search (#780 ) This improves the site setting search so it performs a somewhat fuzzy match. Previously it did not handle seperators such as "space" and a term such as "min_post_length" would not find "min_first_post_length" A more liberal search algorithm makes it easier to the AI to navigate settings. * Minor fix, {{and parameter.enum parameter.enum.length}} is non obviously broken. If parameter.enum is a tracked array it will return the object cause embers and helper implementation. This corrects an issue where enum keeps on selecting itself by mistake.	2024-08-29 16:05:38 +10:00
Rafael dos Santos Silva	a08d168740	FEATURE: Initial support for seeded LLMs (#756 )	2024-08-28 15:57:58 -03:00
Sam	0687ec75c3	FEATURE: allow embedding based search without hyde (#777 ) This allows callers of embedding based search to bypass hyde. Hyde will expand the search term using an LLM, but if an LLM is performing the search we can skip this expansion. It also introduced some tests for the controller which we did not have	2024-08-28 14:17:34 +10:00
Roman Rizzi	c6aeabbfc0	FIX: Malformed message in systemless + inline img scenario (#771 )	2024-08-23 16:41:57 -03:00
Roman Rizzi	eac83eb619	FIX: Triage's search_for_text should be case-insensitive (#767 )	2024-08-22 18:32:42 -03:00
Roman Rizzi	64641b6175	FEATURE: LLM Triage support for systemless models. (#757 ) * FEATURE: LLM Triage support for systemless models. This change adds support for OSS models without support for system messages. LlmTriage's system message field is no longer mandatory. We now send the post contents in a separate user message. * Models using Ollama can also disable system prompts	2024-08-21 11:41:55 -03:00
Sam	97fc822cb6	FEATURE: Allow specific groups access to summary feature on PMs (#760 ) New `ai_pm_summarization_allowed_groups` can be used to allow visibility of the summarization feature on PMs. This can be useful on forums where a lot of communication happens inside PMs.	2024-08-21 07:58:24 +10:00
Roman Rizzi	f789d3ee96	FIX: Triage-flagged posts didn't have a score. (#752 ) The score will contain the LLM result, and make sure the flag isn't displayed when a minimum score threshold is present.	2024-08-14 15:54:09 -03:00
Keegan George	f72ab12761	DEV: Clearly separate post/composer helper settings (#747 )	2024-08-12 15:40:23 -07:00
Sam	cf1e15ef12	FIX: gemini 0801 tool calls (#748 ) Gemini experimental model requires tool_config. Previously defaults would apply. This corrects prompts containing multiple tools on gemini.	2024-08-12 16:10:16 +10:00
Keegan George	0be292f247	FIX: `auto_image_caption` not always present for current user. (#746 )	2024-08-08 13:37:26 -07:00
Rafael dos Santos Silva	1686a8a683	DEV: Move to single table per embeddings type (#561 ) Also move us to halfvecs for speed and disk usage gains	2024-08-08 11:55:20 -03:00
Roman Rizzi	20efc9285e	FIX: Correctly save provider-specific params for new models. (#744 ) Creating a new model, either manually or from presets, doesn't initialize the `provider_params` object, meaning their custom params won't persist. Additionally, this change adds some validations for Bedrock params, which are mandatory, and a clear message when a completion fails because we cannot build the URL.	2024-08-07 16:08:56 -03:00
Rafael dos Santos Silva	3b5245dc54	FIX: Handle nil reply_to_post in AI Bot event handler (#743 ) Somehow it's possible to have a nil `post.reply_to_post` with a non-nil `post.reply_to_post_number`.	2024-08-07 14:22:32 -03:00
Keegan George	2ff3fe3a9f	UX: Use stacked line chart for post sentiment (#737 )	2024-08-02 14:23:29 -07:00
Sam	948cf893a9	FIX: Add tool support to open ai compatible dialect and vllm (#734 ) * FIX: Add tool support to open ai compatible dialect and vllm Automatic tools are in progress in vllm see: https://github.com/vllm-project/vllm/pull/5649 Even when they are supported, initial support will be uneven, only some models have native tool support notably mistral which has some special tokens for tool support. After the above PR lands in vllm we will still need to swap to XML based tools on models without native tool support. * fix specs	2024-08-02 09:52:33 -03:00
Rafael dos Santos Silva	7dc0a2a20a	DEV: Log error responses on Gemini / Cloudflare embeddings requests (#736 )	2024-08-01 17:25:24 -03:00
Roman Rizzi	bed044448c	DEV: Remove old code now that features rely on LlmModels. (#729 ) * DEV: Remove old code now that features rely on LlmModels. * Hide old settings and migrate persona llm overrides * Remove shadowing special URL + seeding code. Use srv:// prefix instead.	2024-07-30 13:44:57 -03:00
Roman Rizzi	5c196bca89	FEATURE: Track if a model can do vision in the llm_models table (#725 ) * FEATURE: Track if a model can do vision in the llm_models table * Data migration	2024-07-24 16:29:47 -03:00
Rafael dos Santos Silva	3502f0f1cd	FEATURE: GPT4o Tokenizer (#721 )	2024-07-22 15:26:14 -03:00
Roman Rizzi	f328b81c78	FIX: Make sure custom tool enums follow json-schema. (#718 ) Enums didn't work as expected because we the dialect couldn't translate them correctly. It doesn't understand what "enum_values" is.	2024-07-16 14:23:17 -03:00
Roman Rizzi	0a8195242b	FIX: Limit system message size to 60% of available tokens. (#714 ) Using RAG fragments can lead to considerably big system messages, which becomes problematic when models have a smaller context window. Before this change, we only look at the rest of the conversation to make sure we don't surpass the limit, which could lead to two unwanted scenarios when having large system messages: All other messages are excluded due to size. The system message already exceeds the limit. As a result, I'm putting a hard-limit of 60% of available tokens. We don't want to aggresively truncate because if rag fragments are included, the system message contains a lot of context to improve the model response, but we also want to make room for the recent messages in the conversation.	2024-07-12 15:09:01 -03:00
PangBo	4ebbdc043e	fix: locale handling in assistant.rb (#705 )	2024-07-05 11:16:09 +02:00
Roman Rizzi	442681a3d3	FIX: Mixtral models have system role support. (#703 ) Using assistant role for system produces an error because they expect alternating roles like user/assistant/user and so on. Prompts cannot start with the assistant role.	2024-07-04 13:23:03 -03:00
Keegan George	eab2f74b58	DEV: Use site locale for composer helper translations (#698 )	2024-07-04 08:23:37 -07:00
Sam	1320eed9b2	FEATURE: move summary to use llm_model (#699 ) This allows summary to use the new LLM models and migrates of API key based model selection Claude 3.5 etc... all work now. --------- Co-authored-by: Roman Rizzi <rizziromanalejandro@gmail.com>	2024-07-04 10:48:18 +10:00
Roman Rizzi	fc081d9da6	FIX: Restore ability to fold summaries, which was accidentally removed (#700 )	2024-07-03 18:10:31 -03:00
Keegan George	1b0ba9197c	DEV: Add summarization logic from core (#658 )	2024-07-02 08:51:59 -07:00
Sam	b671ffe7fa	FIX: info not working, not suppressing hidden tags from report (#696 ) 2 small fixes 1. The info button was not properly working post refactor 2. Suppress any secured tags from report input	2024-07-02 16:38:33 +10:00
Rafael dos Santos Silva	a708d4dfa2	FIX: Use base64 encoded images in AI Image Caption via LLaVa (#693 ) * FIX: Use base64 encoded images in AI Image Caption via LLaVa This fixed a regression introduced in #646 where we started sending schemaless URLs for our LLaVa service, which doesn't handle it well. Moving to base64 encoded images solves: - The service needing to download images Now the service running LLaVa doesn't need internet access - Secure uploads compat Every image is treated the same, less branching for secure uploads - Image Size problems Discourse is now responsible for ensure a max size for images - Troublesome dev env Previously to this commit you would need a dev env that was internet acessible to use llava image captions	2024-06-27 16:24:44 -03:00
Sam	b863ddc94b	FEATURE: custom user defined tools (#677 ) Introduces custom AI tools functionality. 1. Why it was added: The PR adds the ability to create, manage, and use custom AI tools within the Discourse AI system. This feature allows for more flexibility and extensibility in the AI capabilities of the platform. 2. What it does: - Introduces a new `AiTool` model for storing custom AI tools - Adds CRUD (Create, Read, Update, Delete) operations for AI tools - Implements a tool runner system for executing custom tool scripts - Integrates custom tools with existing AI personas - Provides a user interface for managing custom tools in the admin panel 3. Possible use cases: - Creating custom tools for specific tasks or integrations (stock quotes, currency conversion etc...) - Allowing administrators to add new functionalities to AI assistants without modifying core code - Implementing domain-specific tools for particular communities or industries 4. Code structure: The PR introduces several new files and modifies existing ones: a. Models: - `app/models/ai_tool.rb`: Defines the AiTool model - `app/serializers/ai_custom_tool_serializer.rb`: Serializer for AI tools b. Controllers: - `app/controllers/discourse_ai/admin/ai_tools_controller.rb`: Handles CRUD operations for AI tools c. Views and Components: - New Ember.js components for tool management in the admin interface - Updates to existing AI persona management components to support custom tools d. Core functionality: - `lib/ai_bot/tool_runner.rb`: Implements the custom tool execution system - `lib/ai_bot/tools/custom.rb`: Defines the custom tool class e. Routes and configurations: - Updates to route configurations to include new AI tool management pages f. Migrations: - `db/migrate/20240618080148_create_ai_tools.rb`: Creates the ai_tools table g. Tests: - New test files for AI tool functionality and integration The PR integrates the custom tools system with the existing AI persona framework, allowing personas to use both built-in and custom tools. It also includes safety measures such as timeouts and HTTP request limits to prevent misuse of custom tools. Overall, this PR significantly enhances the flexibility and extensibility of the Discourse AI system by allowing administrators to create and manage custom AI tools tailored to their specific needs. Co-authored-by: Martin Brennan <martin@discourse.org>	2024-06-27 17:27:40 +10:00
Sam	af4f871096	FIX: never provide tools with invalid UTF-8 strings (#692 ) Previous to this change, on truncation we could return invalid UTF-8 strings to caller This also allows tools to read up to 30 megs vs the old 4 megs.	2024-06-27 14:06:52 +10:00
Roman Rizzi	f622e2644f	FEATURE: Store provider-specific parameters. (#686 ) Previously, we stored request parameters like the OpenAI organization and Bedrock's access key and region as site settings. This change stores them in the `llm_models` table instead, letting us drop more settings while also becoming more flexible.	2024-06-25 08:26:30 +10:00
Sam	1d5fa0ce6c	FIX: when creating an llm we were not creating user (#685 ) This meant that if you toggle ai user early it surprisingly did not work. Also remove safety settings from gemini, it is overly cautious	2024-06-24 09:59:42 +10:00
Roman Rizzi	091dc626e8	FIX: Make sure LlmEnumerator always return value hashes using symbols (#684 )	2024-06-21 16:35:31 -03:00
Sam	e04a7be122	FEATURE: LLM presets for model creation (#681 ) * FEATURE: LLM presets for model creation Previous to this users needed to look up complicated settings when setting up models. This introduces and extensible preset system with Google/OpenAI/Anthropic presets. This will cover all the most common LLMs, we can always add more as we go. Additionally: - Proper support for Anthropic Claude Sonnet 3.5 - Stop blurring api keys when navigating away - this made it very complex to reuse keys	2024-06-21 17:32:15 +10:00
Roman Rizzi	558574fa87	DEV: Use LlmModels as options in automation rules (#676 )	2024-06-21 08:07:17 +10:00
Rafael dos Santos Silva	714caf34fe	FEATURE: Support for Claude 3.5 Sonnet via AWS Bedrock (#680 )	2024-06-20 17:51:46 -03:00
Roman Rizzi	8849caf136	DEV: Transition "Select model" settings to only use LlmModels (#675 ) We no longer support the "provider:model" format in the "ai_helper_model" and "ai_embeddings_semantic_search_hyde_model" settings. We'll migrate existing values and work with our new data-driven LLM configs from now on.	2024-06-19 18:01:35 -03:00
Sam	0d6d9a6ef5	FEATURE: allow access to private topics if tool permits (#673 ) Previously read tool only had access to public topics, this allows access to all topics user has access to, if admin opts for the option Also - Fixes VLLM migration - Display which llms have bot enabled	2024-06-19 15:49:36 +10:00
Roman Rizzi	8d5f901a67	DEV: Rewire AI bot internals to use LlmModel (#638 ) * DRAFT: Create AI Bot users dynamically and support custom LlmModels * Get user associated to llm_model * Track enabled bots with attribute * Don't store bot username. Minor touches to migrate default values in settings * Handle scenario where vLLM uses a SRV record * Made 3.5-turbo-16k the default version so we can remove hack	2024-06-18 14:32:14 -03:00
Sam	460f5c4553	FIX: display search correctly, bug when stripping XML (#668 ) - Display filtered search correctly, so it is not confusing - When XML stripping, if a chunk was `<` it would crash - SQL Helper improved to be better aware of Data Explorer	2024-06-14 15:28:40 +10:00
Sam	f642a27f11	FIX: Dall E / Artist broken when tool_details is disabled (#667 ) We were missing logic to handle custom_html from tools This also fixes image generation in chat	2024-06-12 17:58:28 +10:00
Sam	52a7dd2a4b	FEATURE: optional tool detail blocks (#662 ) This is a rather huge refactor with 1 new feature (tool details can be suppressed) Previously we use the name "Command" to describe "Tools", this unifies all the internal language and simplifies the code. We also amended the persona UI to use less DToggles which aligns with our design guidelines. Co-authored-by: Martin Brennan <martin@discourse.org>	2024-06-11 18:14:14 +10:00
Sam	875bb04467	FIX: summarize is not working remove for now (#661 ) Summarize is not working right, remove it for now. Also it was specified twice breaking tool calls.	2024-06-08 12:25:30 +10:00
Roman Rizzi	212ee23698	FIX: Use new report color keys defined in discourse/discourse#27240 (#660 )	2024-06-07 15:08:06 -03:00
Sam	8b81ff45b8	FIX: switch off native tools on Anthropic Claude Opus (#659 ) Native tools do not work well on Opus. Chain of Thought prompting means it consumes enormous amounts of tokens and has poor latency. This commit introduce and XML stripper to remove various chain of thought XML islands from anthropic prompts when tools are involved. This mean Opus native tools is now functions (albeit slowly) From local testing XML just works better now. Also fixes enum support in Anthropic native tools	2024-06-07 10:52:01 -03:00

1 2 3 4 5 ...

425 Commits