discourse-ai

Commit Graph

Author	SHA1	Message	Date
Keegan George	6aaf8a0619	DEV: Use existing topic embeddings when suggesting tags/categories on edit (#1189 ) When editing a topic (instead of creating one) and using the tag/category suggestion buttons. We want to use existing topic embeddings instead of creating new ones.	2025-03-12 18:52:07 -07:00
Rafael dos Santos Silva	37bf160d26	FIX: Add workaround to pgvector HNSW search limitations (#1133 ) From [pgvector/pgvector](https://github.com/pgvector/pgvector) README > With approximate indexes, filtering is applied after the index is scanned. If a condition matches 10% of rows, with HNSW and the default hnsw.ef_search of 40, only 4 rows will match on average. For more rows, increase hnsw.ef_search. > > Starting with 0.8.0, you can enable [iterative index scans](https://github.com/pgvector/pgvector#iterative-index-scans), which will automatically scan more of the index when needed. Since we are stuck on 0.7.0 we are going the first option for now.	2025-02-19 16:30:01 -03:00
Roman Rizzi	e52045ebdc	DEV: Robust check for embeddings enabled (#1116 )	2025-02-06 12:18:55 -03:00
Roman Rizzi	f5cf1019fb	FEATURE: configurable embeddings (#1049 ) * Use AR model for embeddings features * endpoints * Embeddings CRUD UI * Add presets. Hide a couple more settings * system specs * Seed embedding definition from old settings * Generate search bit index on the fly. cleanup orphaned data * support for seeded models * Fix run test for new embedding * fix selected model not set correctly	2025-01-21 12:23:19 -03:00
Rafael dos Santos Silva	792df58fbc	FIX: AI Helper category / tag suggestion when user does not categories muted (#1042 )	2024-12-23 15:58:26 -03:00
Roman Rizzi	534b0df391	REFACTOR: Separation of concerns for embedding generation. (#1027 ) In a previous refactor, we moved the responsibility of querying and storing embeddings into the `Schema` class. Now, it's time for embedding generation. The motivation behind these changes is to isolate vector characteristics in simple objects to later replace them with a DB-backed version, similar to what we did with LLM configs.	2024-12-16 09:55:39 -03:00
Roman Rizzi	eae527f99d	REFACTOR: A Simpler way of interacting with embeddings tables. (#1023 ) * REFACTOR: A Simpler way of interacting with embeddings' tables. This change adds a new abstraction called `Schema`, which acts as a repository that supports the same DB features `VectorRepresentation::Base` has, with the exception that removes the need to have duplicated methods per embeddings table. It is also a bit more flexible when performing a similarity search because you can pass it a block that gives you access to the builder, allowing you to add multiple joins/where conditions.	2024-12-13 10:15:21 -03:00
Keegan George	fc88bb08ab	FIX: Tag suggester is suggesting already assigned tags (#990 ) This PR fixes an issue where the tag suggester for edit title topic area was suggesting tags that are already assigned on a post. It also updates the amount of suggested tags to 7 so that there is still a decent amount of tags suggested when tags are already assigned.	2024-12-03 07:25:04 +11:00
Sam	0cb2c413ba	FEATURE: exclude muted categories from category suggester (#979 ) The logic here is that users do not particularly care about topics in the category so we can exclude them from tag and category suggestions	2024-11-29 12:17:28 +11:00
Rafael dos Santos Silva	80adefa1c1	FIX: Move count logic to the end for tag suggestions (#978 )	2024-11-28 16:27:38 -03:00
Keegan George	6b7d7c1179	REFACTOR: Helper suggestions (#914 ) This PR adds some updates to the Helper suggestions to improve it's functionality and modernize some of the codebase.	2024-11-27 12:21:03 -08:00
Sam	6ddc17fd61	DEV: port directory structure to Zeitwerk (#319 ) Previous to this change we relied on explicit loading for a files in Discourse AI. This had a few downsides: - Busywork whenever you add a file (an extra require relative) - We were not keeping to conventions internally ... some places were OpenAI others are OpenAi - Autoloader did not work which lead to lots of full application broken reloads when developing. This moves all of DiscourseAI into a Zeitwerk compatible structure. It also leaves some minimal amount of manual loading (automation - which is loading into an existing namespace that may or may not be there) To avoid needing /lib/discourse_ai/... we mount a namespace thus we are able to keep /lib pointed at ::DiscourseAi Various files were renamed to get around zeitwerk rules and minimize usage of custom inflections Though we can get custom inflections to work it is not worth it, will require a Discourse core patch which means we create a hard dependency.	2023-11-29 15:17:46 +11:00

12 Commits