mirror of
https://github.com/Mintplex-Labs/anything-llm.git
synced 2026-07-22 01:55:24 -04:00
Closed
opened 2026-02-22 18:18:26 -05:00 by yindo
·
8 comments
No Branch/Tag Specified
master
5968-chore-update-default-system-prompt-for-better-tool-use
5846-bug-when-scrolling-up-scrolling-jumps
2235-translations
2235-bug-how-to-upload-a-folder-with-subfolders-with-files-to-anythingllm
refactor-remove-workspace-pfp
5969-bug-stopgenerationbutton-disappears-on-the-first-prompt-that-initiates-an-agent-session
5990-workspace-update-fails-with-unknown-argument-router_id-v1130v1150-intel-mac
opencomputer-examples
pg
feat/image-generation-translations
feat/image-generation
5924-bug-meeting-summary-fails-with-sincludes-is-not-a-function-when-default-llm-is-anthropic-claude
render
feat/uniform-modal-component
5883-bug-unescaped-content-in-json-strings-being-passed-to-document-generator-tools
5901-bug-api-update-embeddings-fails-prisma-argument-filename-is-missing-on-workspace_documentscreate-desktop-windows-v1141
hybrid-search
1981-translations
5752-bug-prompts-to-local-jan-endpoint-unresponsive
feat-disable-native-tool-calling-env-var
5676-bug-non-ollama-agent-providers-do-not-parse-and-present-reasoning-content
feat/markdown-web-scraping
5717-bug-apiv1documentupload-silently-drops-metadata-field-in-desktop-1130-arg-count-mismatch-nested-payload-key
fix/aibitat-context-overflow
5711-bug-erratic-deepseek-v4-flash-the-agent-model-failed-to-respond-400-the-reasoning_content
5631-feat-custom-api-request-timeouts-for-ai-providers
feat-reasoning-control
feat-agent-clarifying-questions-translations
5583-bug-lm-studio-provider-does-not-present-reasoning-output
5313-normalize-translations
feat/memory-translations
5305-lemonade-embedding-engine-swallows-errors-falsely-reports-documents-as-embedded
5060-bug-agent-interactions-agent-are-not-persisted-to-thread-history-via-api
pptx-subagent
feat-render-images-from-mcp-tool-results
stt-provider-expansion-openai-api-compatible
feat-file-search-agent-tool
feat-native-embedder-job-queue
5189-normalize-translations
feat-file-search-agent-tool-translations
i18n-eslint
5140-auto-migration
5112-bug-openrouter-failed-message-bug
3506-feat-parameters-for-openrouter-models
4992-feat-preserve-scroll-position
4973-bug-markdown-numbered-list-display-in-reasoning-pane
desktop
4938-bug-pending-chat-rerendering-ui-bug
quickstart-env
node-llama-cpp-in-container-cuda
node-llama-cpp-in-container
ollama-in-container
4817-feat-set-cooldown-per-mcp-server
4845-keyboard-shortcuts-to-navigate-in-chat
4844-feat-reorder-threads-by-latest-interaction
standardize-username-constraints-normalize-translations
4792-feat-refactor-workspacepfp-image
1382-embed-ip-improvements
1382-bug-embed-api-improvements
refactor-eslint-frontend
4687-feat-refactor-vector-db-providers
4615-feat-disable-apidocs-with-environment-variable
4559-feat-agent-web-search-enable-ordering-of-results
4599-bug-ollama-race-condition-bug
4572-bug-lmstudio-provided-llm-stopped-working-with-anything-llm-after-upgrading-to-190
4508-agent-youtube-transcript-analysis
4497-feat-workspace-names
frontend-eslint
ollama-lmstudio-auto-context-window
4431-validate-vector-database-connectioN
2019-slash-command-keyboard-selection
microsoft-foundry-provider
4431-validate-vector-database-connection
4325-sys-prompt-var-improvements
3209-feat-apiv1workspacestream-chat-sources-citations
4210-bug-voice-to-text-overwrite
4136-feat-jan-as-a-backend-server-option
4172-feat-openai-o3-support
1.8.3-rerelease
web-push-notifications-service
tasks
3955-feat-jinaai-embedder-provider-support
3921-feat-agent-skills-uiux-improvements
3901-bug-validfunccall-checks-optional-arguments
keyboard-dev
1787-custom-roles-and-permissions
add-jira-slack-data-connector
office-extension-wip
lightmode-dropdown-color-update
3586-bug-agent-flow-function-description-provided-by-user-is-not-seen-in-the-llm-query
3463-bug-agent-continues-to-run-if-request-failed-even-after-exit
3439-feat-call-variables-within-the-flow-api-block-url-field
3282-manager-view-models-workspace
3280-token-counting-server-side-truncation-improvements
3147-bug-embedded-chat-widget---not-considering-query-mode-option-always-working-in-chat-mode
2995-feat-disable-temperature-setting-for-deepseek-r1-deepseek-reasoner-model
2827-feat-perplexity-citations
2866-feat-finally-a-gemini-models-endpoint
2647-feat-hpp-header-for-a-c++-code-file-mime-addition
lancedb-revert
1656-feat-implement-tooltip-ui-designs
2011-feat-bump-perplexity-models
1873-feat-auto-add-and-watch-folder-for-document-uploads
1297-feat-gemini-agent-support
1759-bug-ui-bug-fixes
1686-feat-implement-winston-for-logging
1536-bug-toggling-on-users-can-delete-workspaces-does-not-take-effect
agent-ui-mobile-styles
1522-feat-chromadb-support
1595-bug-unable-to-get-live-web-search-and-browsing-agent-working-using-google-custom-search-engine-error-getaddrinfo-enotfound-http-errno-3008
1582-bug-lm-studio-does-not-allow-for-different-model-selection
1312-bug-usernames-should-not-be-case-sensitive-when-logging-in
1029-feat-hf-serverless-inference-api
1086-feat-implement-normalized-input-fields
knowledge-graph-support
644-bug-uploaded-file-name-does-not-match-the-displayed-file-name-after-the-upload
v1.15.0
v1.14.2
v1.14.1
v1.14.0
v1.13.0
v1.12.1
v1.12.0
v1.11.2
v1.11.1
v1.11.0
v1.10.0
v1.9.1
v1.9.0
v1.8.5
v1.8.4
v1.8.3
v1.8.2
v1.8.1
v1.8.0
v1.7.8
v1.7.6
v1.7.5
v1.7.4
v1.4.0
v1.3.0
v1.2.4
v1.2.3
v1.2.2
v1.2.1
v1.2.0
v1.1.1
v1.1.0
v1.0.0
Labels
Clear labels
Desktop
Docker
Integration Request
Integration Request
OS: Linux
OS: Mobile
OS: Windows
UI/UX
blocked
bug
bug
core-team-only
documentation
duplicate
embed-widget
enhancement
feature request
github_actions
good first issue
investigating
needs info / can't replicate
possible bug
pull-request
question
stage: specifications
wontfix
Mirrored from GitHub Pull Request
Milestone
No items
No Milestone
Projects
Clear projects
No project
Notifications
Due Date
No due date set.
Dependencies
No dependencies set.
Reference: Mintplex-Labs/anything-llm#233
Reference in New Issue
Block a user
Blocking a user prevents them from interacting with repositories, such as opening or commenting on pull requests or issues. Learn more about blocking a user.
Delete Branch "%!s()"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
Originally created by @Justin-KM on GitHub (Dec 3, 2023).
Original GitHub issue: https://github.com/Mintplex-Labs/anything-llm/issues/406
Originally assigned to: @Justin-KM on GitHub.
I deployed the docker image from the link:
docker pull mintplexlabs/anythingllm:master
and run it with the following cmd in my MacBook Pro intelCPU
docker run -d -p 3001:3001 mintplexlabs/anythingllm:master
and I got this error when trying to vectorize a document, I try to use embedding models from both openai and azure, it doesn't make any difference:
Error: 1 documents could not be embedded.
@timothycarambat commented on GitHub (Dec 4, 2023):
Can you tail the docker logs to report why the embedding failed? It should be in the logs and from there can trace what was the root cause.
@35develr commented on GitHub (Dec 7, 2023):
Same Problem: here the Log from Docker:
@Justin-KM commented on GitHub (Dec 7, 2023):
here you go:
Chunks created from document: 7
Error: Could not embed document chunks! This document will not be recorded.
at Object.addDocumentToNamespace (/app/server/utils/vectorDbProviders/lance/index.js:211:15)
at process.processTicksAndRejections (node:internal/process/task_queues:95:5)
at async Object.addDocuments (/app/server/models/documents.js:56:26)
at async /app/server/endpoints/workspaces.js:161:33
addDocumentToNamespace Could not embed document chunks! This document will not be recorded.
Failed to vectorize custom-documents/8807-20231115-77f486ad-c2dc-4697-8421-101055077607.json
@timothycarambat commented on GitHub (Dec 7, 2023):
@Justin-KM @roedlmax Both of your errors are coming from the embedding engine that is being used for vectorization. The logs do not include that information - could you provide it?
If you are using a non-hosted solution for embedding (localAI or LMStudio) this error likely lies in the connection information to those services or the model that is being used for embedding is not correct or responding. If so then we should further debug why to determine if the bug lies on our side or if it is the embedding endpoint.
The only other variable is that the document being embedded has issues like not parseable text, however, that is likely not the case here.
@Justin-KM commented on GitHub (Dec 10, 2023):
thanks for your reply.
I spent some time today to trace down into source code and found the followingh http retturn of openai embedding api.
My guest is it is using async http calling openai embedding api, and the api doesn't like multiple calls.
Error: Request failed with status code 429
at createError (/Users/justin/anythingllm/anything-llm/server/node_modules/axios/lib/core/createError.js:16:15)
at settle (/Users/justin/anythingllm/anything-llm/server/node_modules/axios/lib/core/settle.js:17:12)
at IncomingMessage.handleStreamEnd (/Users/justin/anythingllm/anything-llm/server/node_modules/axios/lib/adapters/http.js:322:11)
at IncomingMessage.emit (node:events:531:35)
at endReadableNT (node:internal/streams/readable:1696:12)
at process.processTicksAndRejections (node:internal/process/task_queues:82:21) {
config: {
transitional: {
silentJSONParsing: true,
forcedJSONParsing: true,
clarifyTimeoutError: false
},
adapter: [Function: httpAdapter],
transformRequest: [ [Function: transformRequest] ],
transformResponse: [ [Function: transformResponse] ],
timeout: 0,
xsrfCookieName: 'XSRF-TOKEN',
xsrfHeaderName: 'X-XSRF-TOKEN',
maxContentLength: -1,
maxBodyLength: -1,
validateStatus: [Function: validateStatus],
headers: {
@timothycarambat commented on GitHub (Dec 11, 2023):
@Justin-KM Indeed, looks like a 429 error. Are you using the "free" version of the OpenAI API. A long time ago we had a question like this on Discord and that was the issue.
We have a couple known instances using OpenAi with ~20-30 concurrent users chatting during regular hours and never getting a 429. So wondering if this is an OpenAI account-level thing. Also if you are using the
gpt-4-turbomodel preview I believe the rate limits are lower on that model than others!@35develr commented on GitHub (Dec 19, 2023):
I am not shure if this is a rate limit issue. I run localai in docker with Mistral7B. I dont know where to set a rate-limit and there is no documentation where to set it. Localai gets the input, as follows:
The File i try to vectorize is named test.txt, the Content is: "Dies ist ein Text."
Output of localai Debug Log:
`[127.0.0.1]:44650 200 - GET /readyz
9:40AM DBG Request received:
9:40AM DBG Parameter Config: &{PredictionOptions:{Model:mistral-7b-openorca.Q6_K.gguf Language: N:0 TopP:0.95 TopK:40 Temperature:0.2 Maxtokens:0 Echo:false Batch:0 F16:false IgnoreEOS:false RepeatPenalty:0 Keep:0 MirostatETA:0 MirostatTAU:0 Mirostat:0 FrequencyPenalty:0 TFZ:0 TypicalP:0 Seed:0 NegativePrompt: RopeFreqBase:0 RopeFreqScale:0 NegativePromptScale:0 UseFastTokenizer:false ClipSkip:0 Tokenizer:} Name:mistral F16:true Threads:4 Debug:true Roles:map[] Embeddings:false Backend: TemplateConfig:{Chat:chatml-block ChatMessage:chatml Completion:completion Edit: Functions:} PromptStrings:[] InputStrings:[Dies ist ein Text.] InputToken:[] functionCallString: functionCallNameString: FunctionsConfig:{DisableNoAction:false NoActionFunctionName: NoActionDescriptionName:} FeatureFlag:map[] LLMConfig:{SystemPrompt: TensorSplit: MainGPU: RMSNormEps:0 NGQA:0 PromptCachePath: PromptCacheAll:false PromptCacheRO:false MirostatETA:0 MirostatTAU:0 Mirostat:0 NGPULayers:0 MMap:true MMlock:false LowVRAM:false Grammar: StopWords:[<|im_end|>] Cutstrings:[] TrimSpace:[] ContextSize:4096 NUMA:false LoraAdapter: LoraBase: LoraScale:0 NoMulMatQ:false DraftModel: NDraft:0 Quantization: MMProj: RopeScaling: YarnExtFactor:0 YarnAttnFactor:0 YarnBetaFast:0 YarnBetaSlow:0} AutoGPTQ:{ModelBaseName: Device: Triton:false UseFastTokenizer:false} Diffusers:{PipelineType: SchedulerType: CUDA:false EnableParameters: CFGScale:0 IMG2IMG:false ClipSkip:0 ClipModel: ClipSubFolder:} Step:0 GRPC:{Attempts:0 AttemptsSleepTime:0} VallE:{AudioPath:}}
[172.23.0.1]:57178 500 - POST /v1/embeddings
HERE IS SEE A 500 ERROR AT POST DATA
[127.0.0.1]:44682 200 - GET /readyz`
How can i fix this?
@timothycarambat commented on GitHub (Dec 19, 2023):
You should put this as a bug in LocalAI, this is not a bug from AnythingLLM and appears to be related to how the model is loaded or run on the LocalAI side. AnythingLLM is just reporting back there was an error.
issue: Error: 1 documents could not be embeddedto [GH-ISSUE #406] issue: Error: 1 documents could not be embedded