mirror of
https://github.com/Mintplex-Labs/anything-llm.git
synced 2026-07-19 14:13:35 -04:00
[GH-ISSUE #4572] [BUG]: LMStudio provided LLM stopped working with Anything LLM after upgrading to 1.9.0 #2905
Closed
opened 2026-02-22 18:31:45 -05:00 by yindo
·
7 comments
No Branch/Tag Specified
master
5846-bug-when-scrolling-up-scrolling-jumps
refactor-remove-workspace-pfp
5969-bug-stopgenerationbutton-disappears-on-the-first-prompt-that-initiates-an-agent-session
2235-bug-how-to-upload-a-folder-with-subfolders-with-files-to-anythingllm
5990-workspace-update-fails-with-unknown-argument-router_id-v1130v1150-intel-mac
opencomputer-examples
pg
feat/image-generation-translations
feat/image-generation
5924-bug-meeting-summary-fails-with-sincludes-is-not-a-function-when-default-llm-is-anthropic-claude
render
feat/uniform-modal-component
5883-bug-unescaped-content-in-json-strings-being-passed-to-document-generator-tools
5901-bug-api-update-embeddings-fails-prisma-argument-filename-is-missing-on-workspace_documentscreate-desktop-windows-v1141
hybrid-search
1981-translations
5752-bug-prompts-to-local-jan-endpoint-unresponsive
feat-disable-native-tool-calling-env-var
5676-bug-non-ollama-agent-providers-do-not-parse-and-present-reasoning-content
feat/markdown-web-scraping
5717-bug-apiv1documentupload-silently-drops-metadata-field-in-desktop-1130-arg-count-mismatch-nested-payload-key
fix/aibitat-context-overflow
5711-bug-erratic-deepseek-v4-flash-the-agent-model-failed-to-respond-400-the-reasoning_content
5631-feat-custom-api-request-timeouts-for-ai-providers
feat-reasoning-control
feat-agent-clarifying-questions-translations
5583-bug-lm-studio-provider-does-not-present-reasoning-output
5313-normalize-translations
feat/memory-translations
5305-lemonade-embedding-engine-swallows-errors-falsely-reports-documents-as-embedded
5060-bug-agent-interactions-agent-are-not-persisted-to-thread-history-via-api
pptx-subagent
feat-render-images-from-mcp-tool-results
stt-provider-expansion-openai-api-compatible
feat-file-search-agent-tool
feat-native-embedder-job-queue
5189-normalize-translations
feat-file-search-agent-tool-translations
i18n-eslint
5140-auto-migration
5112-bug-openrouter-failed-message-bug
3506-feat-parameters-for-openrouter-models
4992-feat-preserve-scroll-position
4973-bug-markdown-numbered-list-display-in-reasoning-pane
desktop
4938-bug-pending-chat-rerendering-ui-bug
quickstart-env
node-llama-cpp-in-container-cuda
node-llama-cpp-in-container
ollama-in-container
4817-feat-set-cooldown-per-mcp-server
4845-keyboard-shortcuts-to-navigate-in-chat
4844-feat-reorder-threads-by-latest-interaction
standardize-username-constraints-normalize-translations
4792-feat-refactor-workspacepfp-image
1382-embed-ip-improvements
1382-bug-embed-api-improvements
refactor-eslint-frontend
4687-feat-refactor-vector-db-providers
4615-feat-disable-apidocs-with-environment-variable
4559-feat-agent-web-search-enable-ordering-of-results
4599-bug-ollama-race-condition-bug
4572-bug-lmstudio-provided-llm-stopped-working-with-anything-llm-after-upgrading-to-190
4508-agent-youtube-transcript-analysis
4497-feat-workspace-names
frontend-eslint
ollama-lmstudio-auto-context-window
4431-validate-vector-database-connectioN
2019-slash-command-keyboard-selection
microsoft-foundry-provider
4431-validate-vector-database-connection
4325-sys-prompt-var-improvements
3209-feat-apiv1workspacestream-chat-sources-citations
4210-bug-voice-to-text-overwrite
4136-feat-jan-as-a-backend-server-option
4172-feat-openai-o3-support
1.8.3-rerelease
web-push-notifications-service
tasks
3955-feat-jinaai-embedder-provider-support
3921-feat-agent-skills-uiux-improvements
3901-bug-validfunccall-checks-optional-arguments
keyboard-dev
1787-custom-roles-and-permissions
add-jira-slack-data-connector
office-extension-wip
lightmode-dropdown-color-update
3586-bug-agent-flow-function-description-provided-by-user-is-not-seen-in-the-llm-query
3463-bug-agent-continues-to-run-if-request-failed-even-after-exit
3439-feat-call-variables-within-the-flow-api-block-url-field
3282-manager-view-models-workspace
3280-token-counting-server-side-truncation-improvements
3147-bug-embedded-chat-widget---not-considering-query-mode-option-always-working-in-chat-mode
2995-feat-disable-temperature-setting-for-deepseek-r1-deepseek-reasoner-model
2827-feat-perplexity-citations
2866-feat-finally-a-gemini-models-endpoint
2647-feat-hpp-header-for-a-c++-code-file-mime-addition
lancedb-revert
1656-feat-implement-tooltip-ui-designs
2011-feat-bump-perplexity-models
1873-feat-auto-add-and-watch-folder-for-document-uploads
1297-feat-gemini-agent-support
1759-bug-ui-bug-fixes
1686-feat-implement-winston-for-logging
1536-bug-toggling-on-users-can-delete-workspaces-does-not-take-effect
agent-ui-mobile-styles
1522-feat-chromadb-support
1595-bug-unable-to-get-live-web-search-and-browsing-agent-working-using-google-custom-search-engine-error-getaddrinfo-enotfound-http-errno-3008
1582-bug-lm-studio-does-not-allow-for-different-model-selection
1312-bug-usernames-should-not-be-case-sensitive-when-logging-in
1029-feat-hf-serverless-inference-api
1086-feat-implement-normalized-input-fields
knowledge-graph-support
644-bug-uploaded-file-name-does-not-match-the-displayed-file-name-after-the-upload
v1.15.0
v1.14.2
v1.14.1
v1.14.0
v1.13.0
v1.12.1
v1.12.0
v1.11.2
v1.11.1
v1.11.0
v1.10.0
v1.9.1
v1.9.0
v1.8.5
v1.8.4
v1.8.3
v1.8.2
v1.8.1
v1.8.0
v1.7.8
v1.7.6
v1.7.5
v1.7.4
v1.4.0
v1.3.0
v1.2.4
v1.2.3
v1.2.2
v1.2.1
v1.2.0
v1.1.1
v1.1.0
v1.0.0
Labels
Clear labels
Desktop
Docker
Integration Request
Integration Request
OS: Linux
OS: Mobile
OS: Windows
UI/UX
blocked
bug
bug
core-team-only
documentation
duplicate
embed-widget
enhancement
feature request
github_actions
good first issue
investigating
needs info / can't replicate
possible bug
pull-request
question
stage: specifications
wontfix
Mirrored from GitHub Pull Request
Milestone
No items
No Milestone
Projects
Clear projects
No project
Notifications
Due Date
No due date set.
Dependencies
No dependencies set.
Reference: Mintplex-Labs/anything-llm#2905
Reference in New Issue
Block a user
Blocking a user prevents them from interacting with repositories, such as opening or commenting on pull requests or issues. Learn more about blocking a user.
Delete Branch "%!s()"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
Originally created by @Ackerka on GitHub (Oct 21, 2025).
Original GitHub issue: https://github.com/Mintplex-Labs/anything-llm/issues/4572
Originally assigned to: @shatfield4, @timothycarambat on GitHub.
How are you running AnythingLLM?
Desktop
What happened?
I had 1.8.5-r2 Anything LLM desktop installed. It worked properly with LM Studio as LLM and embedding model provider simultaneously. During my attempt to add an MCP server to Anything LLM I upgraded to the latest version (1.9.0). Although my configuration remained intact the system stopped working: it was possible to select an LLM, so there was some sort of communication between LMStudio and Anything LLM but once I started to chat I got the following mysterious error message: "An error occurred while streaming response. network error". Once this message appeared Anything LLM stopped being able to list LM Studio models at all until the application was restarted. After I managed to fix my MCP configuration issue (full path was required in the command field for proper operation) I had to reinstall the older 1.8.5-r2 version of Anything LLM to get it working properly with LLM Studio as LLM and embedder provider as well. So now the old version works properly and it is able to use the MCP server as well in agent mode (@agent).
Please, check the reason of functional degradation with the upgrade before too many users run into the same issue.
My platform is a Mac Studio M3 Ultra if it counts.
Are there known steps to reproduce?
No response
@shatfield4 commented on GitHub (Oct 22, 2025):
I have tested this on the latest version of AnythingLLM (1.9.0) and am unable to replicate the issue. If you are able to consistently replicate the issue, please provide me with exact instructions on how you're able to get it to do this on your setup.
Also, please provide me with what models you're using for the embedder and LLM inside of LMStudio since this may help us narrow down possibilities.
@nhaneezy commented on GitHub (Oct 23, 2025):
Update: Downgraded to 1.8.5-r2 and don't see the problem anymore
Also running macOS Tahoe 26.1 on a Mac Studio M3 Ultra.
I'm having a similar problem. After upgrading to 1.9.0, using AnythingLLM will sometimes work and sometimes not work. Here's a log of LM Studio of it working and then I continue the chat on AnythingLLM and it fails:
2025-10-23 14:05:18 [INFO]
Returning {
"data": [
{
"id": "qwen3-235b-a22b-mlx",
"object": "model",
"type": "llm",
"publisher": "LibraxisAI",
"arch": "qwen3_moe",
"compatibility_type": "mlx",
"quantization": "5bit",
"state": "loaded",
"max_context_length": 40960,
"loaded_context_length": 40960,
"capabilities": [
"tool_use"
]
},
{
"id": "text-embedding-nomic-embed-text-v1.5",
"object": "model",
"type": "embeddings",
"publisher": "nomic-ai",
"arch": "nomic-bert",
"compatibility_type": "gguf",
"quantization": "Q4_K_M",
"state": "not-loaded",
"max_context_length": 2048
},
{
"id": "qwen/qwen3-235b-a22b-2507",
"object": "model",
"type": "llm",
"publisher": "qwen",
"arch": "qwen3_moe",
"compatibility_type": "mlx",
"quantization": "6bit",
"state": "not-loaded",
"max_context_length": 262144,
"capabilities": [
"tool_use"
]
},
{
"id": "qwen/qwen3-1.7b",
"object": "model",
"type": "llm",
"publisher": "qwen",
"arch": "qwen3",
"compatibility_type": "mlx",
"quantization": "4bit",
"state": "not-loaded",
"max_context_length": 40960,
"capabilities": [
"tool_use"
]
},
{
"id": "qwen3-235b-a22b",
"object": "model",
"type": "llm",
"publisher": "unsloth",
"arch": "qwen3moe",
"compatibility_type": "gguf",
"quantization": "Q5_K_S",
"state": "not-loaded",
"max_context_length": 40960,
"capabilities": [
"tool_use"
]
}
],
"object": "list"
}
2025-10-23 14:05:19 [INFO]
A value was passed to the 'Authorization' header, but the server is not configured to authenticate via API token. The 'Authorization' header will be ignored.
2025-10-23 14:05:19 [INFO]
A value was passed to the 'Authorization' header, but the server is not configured to authenticate via API token. The 'Authorization' header will be ignored.
2025-10-23 14:05:19 [INFO]
[LM STUDIO SERVER] Running chat completion on conversation with 8 messages.
2025-10-23 14:05:19 [INFO]
[LM STUDIO SERVER] Streaming response...
2025-10-23 14:06:02 [INFO]
Finished streaming response
2025-10-23 14:06:22 [INFO]
Returning {
"data": [
{
"id": "qwen3-235b-a22b-mlx",
"object": "model",
"type": "llm",
"publisher": "LibraxisAI",
"arch": "qwen3_moe",
"compatibility_type": "mlx",
"quantization": "5bit",
"state": "loaded",
"max_context_length": 40960,
"loaded_context_length": 40960,
"capabilities": [
"tool_use"
]
},
{
"id": "text-embedding-nomic-embed-text-v1.5",
"object": "model",
"type": "embeddings",
"publisher": "nomic-ai",
"arch": "nomic-bert",
"compatibility_type": "gguf",
"quantization": "Q4_K_M",
"state": "not-loaded",
"max_context_length": 2048
},
{
"id": "qwen/qwen3-235b-a22b-2507",
"object": "model",
"type": "llm",
"publisher": "qwen",
"arch": "qwen3_moe",
"compatibility_type": "mlx",
"quantization": "6bit",
"state": "not-loaded",
"max_context_length": 262144,
"capabilities": [
"tool_use"
]
},
{
"id": "qwen/qwen3-1.7b",
"object": "model",
"type": "llm",
"publisher": "qwen",
"arch": "qwen3",
"compatibility_type": "mlx",
"quantization": "4bit",
"state": "not-loaded",
"max_context_length": 40960,
"capabilities": [
"tool_use"
]
},
{
"id": "qwen3-235b-a22b",
"object": "model",
"type": "llm",
"publisher": "unsloth",
"arch": "qwen3moe",
"compatibility_type": "gguf",
"quantization": "Q5_K_S",
"state": "not-loaded",
"max_context_length": 40960,
"capabilities": [
"tool_use"
]
}
],
"object": "list"
}
2025-10-23 14:06:27 [INFO]
A value was passed to the 'Authorization' header, but the server is not configured to authenticate via API token. The 'Authorization' header will be ignored.
2025-10-23 14:06:27 [INFO]
A value was passed to the 'Authorization' header, but the server is not configured to authenticate via API token. The 'Authorization' header will be ignored.
2025-10-23 14:06:27 [INFO]
[LM STUDIO SERVER] Running chat completion on conversation with 10 messages.
2025-10-23 14:06:27 [INFO]
[LM STUDIO SERVER] Streaming response...
2025-10-23 14:07:02 [INFO]
Finished streaming response
2025-10-23 14:08:19 [INFO]
Returning {
"data": [
{
"id": "qwen3-235b-a22b-mlx",
"object": "model",
"type": "llm",
"publisher": "LibraxisAI",
"arch": "qwen3_moe",
"compatibility_type": "mlx",
"quantization": "5bit",
"state": "loaded",
"max_context_length": 40960,
"loaded_context_length": 40960,
"capabilities": [
"tool_use"
]
},
{
"id": "text-embedding-nomic-embed-text-v1.5",
"object": "model",
"type": "embeddings",
"publisher": "nomic-ai",
"arch": "nomic-bert",
"compatibility_type": "gguf",
"quantization": "Q4_K_M",
"state": "not-loaded",
"max_context_length": 2048
},
{
"id": "qwen/qwen3-235b-a22b-2507",
"object": "model",
"type": "llm",
"publisher": "qwen",
"arch": "qwen3_moe",
"compatibility_type": "mlx",
"quantization": "6bit",
"state": "not-loaded",
"max_context_length": 262144,
"capabilities": [
"tool_use"
]
},
{
"id": "qwen/qwen3-1.7b",
"object": "model",
"type": "llm",
"publisher": "qwen",
"arch": "qwen3",
"compatibility_type": "mlx",
"quantization": "4bit",
"state": "not-loaded",
"max_context_length": 40960,
"capabilities": [
"tool_use"
]
},
{
"id": "qwen3-235b-a22b",
"object": "model",
"type": "llm",
"publisher": "unsloth",
"arch": "qwen3moe",
"compatibility_type": "gguf",
"quantization": "Q5_K_S",
"state": "not-loaded",
"max_context_length": 40960,
"capabilities": [
"tool_use"
]
}
],
"object": "list"
}
@shatfield4 commented on GitHub (Oct 23, 2025):
Thank you for this info, we are investigating this :)
@timothycarambat commented on GitHub (Oct 24, 2025):
what version of LMStudio is in use here?
@Ackerka commented on GitHub (Oct 25, 2025):
I use
Even if I switched back to the LLM shipped with Anything LLM it was not responding.
Platform:
Mac Studio M3 Ultra 512GB
OS: macOS Tahoe 26.0.1
It looks like something broke after the first attempt to chat after upgrade, so even model selection becomes impossible until restarting Anything LLM 1.9.0.
@shatfield4 commented on GitHub (Oct 29, 2025):
@Ackerka @nhaneezy When you get this generic error inside AnythingLLM, can you please provide us with the LM Studio developer logs?
I have been testing this with various mcp servers and found that some mcp severs like Desktop Commander have many tools which causes the system prompt to be massive and sometimes causes issues with context overflow in specific models like the
mistralai/magistral-small-2509that was mentioned.I just want to make sure that this bug I am getting is the same thing that you are experiencing.
I also discovered that there may be a race condition bug that was introduced in #4468 when we try to get context window sizes. It's possible that if LMStudio isn't running or responds slowly, the provider fails to initialize properly and crashes the model loading. Any logs you can provide from LM Studio would be very helpful to ensure we get this fixed.
@nhaneezy commented on GitHub (Oct 29, 2025):
in my initial response above I pasted the LM Studio developer logs. I realize it may not be too illuminating but that's what's in the logs after the initial successful chat entered and then the following chat entered which ends in the generic error on AnythingLLM
[BUG]: LMStudio provided LLM stopped working with Anything LLM after upgrading to 1.9.0to [GH-ISSUE #4572] [BUG]: LMStudio provided LLM stopped working with Anything LLM after upgrading to 1.9.0