mirror of
https://github.com/Mintplex-Labs/anything-llm.git
synced 2026-07-19 22:23:50 -04:00
Closed
opened 2026-02-22 18:24:50 -05:00 by yindo
·
4 comments
No Branch/Tag Specified
master
5846-bug-when-scrolling-up-scrolling-jumps
refactor-remove-workspace-pfp
5969-bug-stopgenerationbutton-disappears-on-the-first-prompt-that-initiates-an-agent-session
2235-bug-how-to-upload-a-folder-with-subfolders-with-files-to-anythingllm
5990-workspace-update-fails-with-unknown-argument-router_id-v1130v1150-intel-mac
opencomputer-examples
pg
feat/image-generation-translations
feat/image-generation
5924-bug-meeting-summary-fails-with-sincludes-is-not-a-function-when-default-llm-is-anthropic-claude
render
feat/uniform-modal-component
5883-bug-unescaped-content-in-json-strings-being-passed-to-document-generator-tools
5901-bug-api-update-embeddings-fails-prisma-argument-filename-is-missing-on-workspace_documentscreate-desktop-windows-v1141
hybrid-search
1981-translations
5752-bug-prompts-to-local-jan-endpoint-unresponsive
feat-disable-native-tool-calling-env-var
5676-bug-non-ollama-agent-providers-do-not-parse-and-present-reasoning-content
feat/markdown-web-scraping
5717-bug-apiv1documentupload-silently-drops-metadata-field-in-desktop-1130-arg-count-mismatch-nested-payload-key
fix/aibitat-context-overflow
5711-bug-erratic-deepseek-v4-flash-the-agent-model-failed-to-respond-400-the-reasoning_content
5631-feat-custom-api-request-timeouts-for-ai-providers
feat-reasoning-control
feat-agent-clarifying-questions-translations
5583-bug-lm-studio-provider-does-not-present-reasoning-output
5313-normalize-translations
feat/memory-translations
5305-lemonade-embedding-engine-swallows-errors-falsely-reports-documents-as-embedded
5060-bug-agent-interactions-agent-are-not-persisted-to-thread-history-via-api
pptx-subagent
feat-render-images-from-mcp-tool-results
stt-provider-expansion-openai-api-compatible
feat-file-search-agent-tool
feat-native-embedder-job-queue
5189-normalize-translations
feat-file-search-agent-tool-translations
i18n-eslint
5140-auto-migration
5112-bug-openrouter-failed-message-bug
3506-feat-parameters-for-openrouter-models
4992-feat-preserve-scroll-position
4973-bug-markdown-numbered-list-display-in-reasoning-pane
desktop
4938-bug-pending-chat-rerendering-ui-bug
quickstart-env
node-llama-cpp-in-container-cuda
node-llama-cpp-in-container
ollama-in-container
4817-feat-set-cooldown-per-mcp-server
4845-keyboard-shortcuts-to-navigate-in-chat
4844-feat-reorder-threads-by-latest-interaction
standardize-username-constraints-normalize-translations
4792-feat-refactor-workspacepfp-image
1382-embed-ip-improvements
1382-bug-embed-api-improvements
refactor-eslint-frontend
4687-feat-refactor-vector-db-providers
4615-feat-disable-apidocs-with-environment-variable
4559-feat-agent-web-search-enable-ordering-of-results
4599-bug-ollama-race-condition-bug
4572-bug-lmstudio-provided-llm-stopped-working-with-anything-llm-after-upgrading-to-190
4508-agent-youtube-transcript-analysis
4497-feat-workspace-names
frontend-eslint
ollama-lmstudio-auto-context-window
4431-validate-vector-database-connectioN
2019-slash-command-keyboard-selection
microsoft-foundry-provider
4431-validate-vector-database-connection
4325-sys-prompt-var-improvements
3209-feat-apiv1workspacestream-chat-sources-citations
4210-bug-voice-to-text-overwrite
4136-feat-jan-as-a-backend-server-option
4172-feat-openai-o3-support
1.8.3-rerelease
web-push-notifications-service
tasks
3955-feat-jinaai-embedder-provider-support
3921-feat-agent-skills-uiux-improvements
3901-bug-validfunccall-checks-optional-arguments
keyboard-dev
1787-custom-roles-and-permissions
add-jira-slack-data-connector
office-extension-wip
lightmode-dropdown-color-update
3586-bug-agent-flow-function-description-provided-by-user-is-not-seen-in-the-llm-query
3463-bug-agent-continues-to-run-if-request-failed-even-after-exit
3439-feat-call-variables-within-the-flow-api-block-url-field
3282-manager-view-models-workspace
3280-token-counting-server-side-truncation-improvements
3147-bug-embedded-chat-widget---not-considering-query-mode-option-always-working-in-chat-mode
2995-feat-disable-temperature-setting-for-deepseek-r1-deepseek-reasoner-model
2827-feat-perplexity-citations
2866-feat-finally-a-gemini-models-endpoint
2647-feat-hpp-header-for-a-c++-code-file-mime-addition
lancedb-revert
1656-feat-implement-tooltip-ui-designs
2011-feat-bump-perplexity-models
1873-feat-auto-add-and-watch-folder-for-document-uploads
1297-feat-gemini-agent-support
1759-bug-ui-bug-fixes
1686-feat-implement-winston-for-logging
1536-bug-toggling-on-users-can-delete-workspaces-does-not-take-effect
agent-ui-mobile-styles
1522-feat-chromadb-support
1595-bug-unable-to-get-live-web-search-and-browsing-agent-working-using-google-custom-search-engine-error-getaddrinfo-enotfound-http-errno-3008
1582-bug-lm-studio-does-not-allow-for-different-model-selection
1312-bug-usernames-should-not-be-case-sensitive-when-logging-in
1029-feat-hf-serverless-inference-api
1086-feat-implement-normalized-input-fields
knowledge-graph-support
644-bug-uploaded-file-name-does-not-match-the-displayed-file-name-after-the-upload
v1.15.0
v1.14.2
v1.14.1
v1.14.0
v1.13.0
v1.12.1
v1.12.0
v1.11.2
v1.11.1
v1.11.0
v1.10.0
v1.9.1
v1.9.0
v1.8.5
v1.8.4
v1.8.3
v1.8.2
v1.8.1
v1.8.0
v1.7.8
v1.7.6
v1.7.5
v1.7.4
v1.4.0
v1.3.0
v1.2.4
v1.2.3
v1.2.2
v1.2.1
v1.2.0
v1.1.1
v1.1.0
v1.0.0
Labels
Clear labels
Desktop
Docker
Integration Request
Integration Request
OS: Linux
OS: Mobile
OS: Windows
UI/UX
blocked
bug
bug
core-team-only
documentation
duplicate
embed-widget
enhancement
feature request
github_actions
good first issue
investigating
needs info / can't replicate
possible bug
pull-request
question
stage: specifications
wontfix
Mirrored from GitHub Pull Request
Milestone
No items
No Milestone
Projects
Clear projects
No project
Notifications
Due Date
No due date set.
Dependencies
No dependencies set.
Reference: Mintplex-Labs/anything-llm#1446
Reference in New Issue
Block a user
Blocking a user prevents them from interacting with repositories, such as opening or commenting on pull requests or issues. Learn more about blocking a user.
Delete Branch "%!s()"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
Originally created by @goodglitch on GitHub (Sep 3, 2024).
Original GitHub issue: https://github.com/Mintplex-Labs/anything-llm/issues/2212
What would you like to see?
The main problem with LLMs are their hallucinations. The best and the easiest way to have it under control is to double check results ourselves. That would be much easier if LLM will cite the exact source the information was retrieved from. For this I would suggest instead of using [CONTEXT #] to use something like [CONTEXT SRC="Attention Is All You Need"] blah, blah. [CONTEXT END]. Then in the LLM Prompt tell it directly to show the source that it has used and provide few examples like that:
Given the following conversation, relevant context, and a follow up question, reply with an answer to the current question the user is asking. Return only your response to the question given the above information following the users instructions as needed.
[CONTEXT SRC="Training strategies for improving listeners' comprehension of foreign-accented speech" page=5]
blah, blah.
[CONTEXT END]
[CONTEXT SRC="Teaching native speakers to listen to foreign-accented speech" page 23]
blah, blah.
[CONTEXT END]
ASSISTANT:
A few researchers in the linguistics field have developed training programs designed to improve native speakers' ability to understand accented speech (source 1). Their training techniques are based on the research described above indicating that comprehension improves with exposure to non-native speech. Other researchers conducted their training with students preparing to be social workers, but note that other professionals who work with non-native speakers could benefit from a similar program. (source 2)
References
Source 1: "Training strategies for improving listeners' comprehension of foreign-accented speech", page 5.
Source 2: "Teaching native speakers to listen to foreign-accented speech", page 23.
For htmls I think a relevant link will be sufficient.
@timothycarambat commented on GitHub (Sep 3, 2024):
You cannot trust an LLM to do the above task, if it hallucinates a response even with valid context, it will hallucinate a validation for it as well. The reason the citation exists is that the only trustful way to verify information from a response is by directly viewing the context used by the LLM. This is why we include it in citations since you cannot, nor should you, trust an LLM to always give you perfect recall and answers.
When a citation is present we not only show the text snippets but also name the source document it was derived from. You get both of these details in the UI currently. I am not sure what more there is to accomplish on this detail that does not end with a black box evaluating a black box
@goodglitch commented on GitHub (Sep 4, 2024):
Thank you for the answer, missed Citations part. Happy that you have it, although there is no exact snippet and what is also important page number in the current reply:
Btw, just recently stumbled upon your two competitors Kotaemon and Ragflow. May be you can draw some inspiration from them. I especially like that you can immediately see relevant page of PDF in the UI in Kotaemon. Still haven't time to check Ragflow though.
@goodglitch commented on GitHub (Sep 4, 2024):
Btw, I am really curious how do you know what source has LLM used? I mean, you know what snippets you gave to it, but it is unclear to me how one can guess what exactly was used in the answer.
@david-morris commented on GitHub (Sep 4, 2024):
I would like to move this to a discussion, but...
I understand that "hey where was your attention inside the loaded context" is the responsibility of the language model.
But the embedder decided to load a specific context from the available documents.
I get that the context window isn't going to line up with a page or paragraph, but it would be nice to spit out the part of the document that was loaded, even if it needs to be output as text. If we have some sort of "start of two pages worth of topic ... and that's why foo was chosen as a metasyntactic variable" citation mode, then we can look at how we could implement turning that into some sort of link that highlights relevant info.
I get that it's not what we really want (exact location citation), but it would make reviewing citations of large documents like books possible.
[FEAT]: Cite source when using RAGto [GH-ISSUE #2212] [FEAT]: Cite source when using RAG