mirror of
https://github.com/Mintplex-Labs/anything-llm.git
synced 2026-07-21 01:25:28 -04:00
Open
opened 2026-02-22 18:32:07 -05:00 by yindo
·
4 comments
No Branch/Tag Specified
master
5846-bug-when-scrolling-up-scrolling-jumps
2235-translations
2235-bug-how-to-upload-a-folder-with-subfolders-with-files-to-anythingllm
refactor-remove-workspace-pfp
5969-bug-stopgenerationbutton-disappears-on-the-first-prompt-that-initiates-an-agent-session
5968-chore-update-default-system-prompt-for-better-tool-use
5990-workspace-update-fails-with-unknown-argument-router_id-v1130v1150-intel-mac
opencomputer-examples
pg
feat/image-generation-translations
feat/image-generation
5924-bug-meeting-summary-fails-with-sincludes-is-not-a-function-when-default-llm-is-anthropic-claude
render
feat/uniform-modal-component
5883-bug-unescaped-content-in-json-strings-being-passed-to-document-generator-tools
5901-bug-api-update-embeddings-fails-prisma-argument-filename-is-missing-on-workspace_documentscreate-desktop-windows-v1141
hybrid-search
1981-translations
5752-bug-prompts-to-local-jan-endpoint-unresponsive
feat-disable-native-tool-calling-env-var
5676-bug-non-ollama-agent-providers-do-not-parse-and-present-reasoning-content
feat/markdown-web-scraping
5717-bug-apiv1documentupload-silently-drops-metadata-field-in-desktop-1130-arg-count-mismatch-nested-payload-key
fix/aibitat-context-overflow
5711-bug-erratic-deepseek-v4-flash-the-agent-model-failed-to-respond-400-the-reasoning_content
5631-feat-custom-api-request-timeouts-for-ai-providers
feat-reasoning-control
feat-agent-clarifying-questions-translations
5583-bug-lm-studio-provider-does-not-present-reasoning-output
5313-normalize-translations
feat/memory-translations
5305-lemonade-embedding-engine-swallows-errors-falsely-reports-documents-as-embedded
5060-bug-agent-interactions-agent-are-not-persisted-to-thread-history-via-api
pptx-subagent
feat-render-images-from-mcp-tool-results
stt-provider-expansion-openai-api-compatible
feat-file-search-agent-tool
feat-native-embedder-job-queue
5189-normalize-translations
feat-file-search-agent-tool-translations
i18n-eslint
5140-auto-migration
5112-bug-openrouter-failed-message-bug
3506-feat-parameters-for-openrouter-models
4992-feat-preserve-scroll-position
4973-bug-markdown-numbered-list-display-in-reasoning-pane
desktop
4938-bug-pending-chat-rerendering-ui-bug
quickstart-env
node-llama-cpp-in-container-cuda
node-llama-cpp-in-container
ollama-in-container
4817-feat-set-cooldown-per-mcp-server
4845-keyboard-shortcuts-to-navigate-in-chat
4844-feat-reorder-threads-by-latest-interaction
standardize-username-constraints-normalize-translations
4792-feat-refactor-workspacepfp-image
1382-embed-ip-improvements
1382-bug-embed-api-improvements
refactor-eslint-frontend
4687-feat-refactor-vector-db-providers
4615-feat-disable-apidocs-with-environment-variable
4559-feat-agent-web-search-enable-ordering-of-results
4599-bug-ollama-race-condition-bug
4572-bug-lmstudio-provided-llm-stopped-working-with-anything-llm-after-upgrading-to-190
4508-agent-youtube-transcript-analysis
4497-feat-workspace-names
frontend-eslint
ollama-lmstudio-auto-context-window
4431-validate-vector-database-connectioN
2019-slash-command-keyboard-selection
microsoft-foundry-provider
4431-validate-vector-database-connection
4325-sys-prompt-var-improvements
3209-feat-apiv1workspacestream-chat-sources-citations
4210-bug-voice-to-text-overwrite
4136-feat-jan-as-a-backend-server-option
4172-feat-openai-o3-support
1.8.3-rerelease
web-push-notifications-service
tasks
3955-feat-jinaai-embedder-provider-support
3921-feat-agent-skills-uiux-improvements
3901-bug-validfunccall-checks-optional-arguments
keyboard-dev
1787-custom-roles-and-permissions
add-jira-slack-data-connector
office-extension-wip
lightmode-dropdown-color-update
3586-bug-agent-flow-function-description-provided-by-user-is-not-seen-in-the-llm-query
3463-bug-agent-continues-to-run-if-request-failed-even-after-exit
3439-feat-call-variables-within-the-flow-api-block-url-field
3282-manager-view-models-workspace
3280-token-counting-server-side-truncation-improvements
3147-bug-embedded-chat-widget---not-considering-query-mode-option-always-working-in-chat-mode
2995-feat-disable-temperature-setting-for-deepseek-r1-deepseek-reasoner-model
2827-feat-perplexity-citations
2866-feat-finally-a-gemini-models-endpoint
2647-feat-hpp-header-for-a-c++-code-file-mime-addition
lancedb-revert
1656-feat-implement-tooltip-ui-designs
2011-feat-bump-perplexity-models
1873-feat-auto-add-and-watch-folder-for-document-uploads
1297-feat-gemini-agent-support
1759-bug-ui-bug-fixes
1686-feat-implement-winston-for-logging
1536-bug-toggling-on-users-can-delete-workspaces-does-not-take-effect
agent-ui-mobile-styles
1522-feat-chromadb-support
1595-bug-unable-to-get-live-web-search-and-browsing-agent-working-using-google-custom-search-engine-error-getaddrinfo-enotfound-http-errno-3008
1582-bug-lm-studio-does-not-allow-for-different-model-selection
1312-bug-usernames-should-not-be-case-sensitive-when-logging-in
1029-feat-hf-serverless-inference-api
1086-feat-implement-normalized-input-fields
knowledge-graph-support
644-bug-uploaded-file-name-does-not-match-the-displayed-file-name-after-the-upload
v1.15.0
v1.14.2
v1.14.1
v1.14.0
v1.13.0
v1.12.1
v1.12.0
v1.11.2
v1.11.1
v1.11.0
v1.10.0
v1.9.1
v1.9.0
v1.8.5
v1.8.4
v1.8.3
v1.8.2
v1.8.1
v1.8.0
v1.7.8
v1.7.6
v1.7.5
v1.7.4
v1.4.0
v1.3.0
v1.2.4
v1.2.3
v1.2.2
v1.2.1
v1.2.0
v1.1.1
v1.1.0
v1.0.0
Labels
Clear labels
Desktop
Docker
Integration Request
Integration Request
OS: Linux
OS: Mobile
OS: Windows
UI/UX
blocked
bug
bug
core-team-only
documentation
duplicate
embed-widget
enhancement
feature request
github_actions
good first issue
investigating
needs info / can't replicate
possible bug
pull-request
question
stage: specifications
wontfix
Mirrored from GitHub Pull Request
Milestone
No items
No Milestone
Projects
Clear projects
No project
Notifications
Due Date
No due date set.
Dependencies
No dependencies set.
Reference: Mintplex-Labs/anything-llm#2979
Reference in New Issue
Block a user
Blocking a user prevents them from interacting with repositories, such as opening or commenting on pull requests or issues. Learn more about blocking a user.
Delete Branch "%!s()"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
Originally created by @mgraffam on GitHub (Dec 1, 2025).
Original GitHub issue: https://github.com/Mintplex-Labs/anything-llm/issues/4700
What would you like to see?
It's crucial for mobile, less so for desktop: allow configuring the time-out on the model response.
When I have bad service on mobile, models time out just because of connectivity issues. If AnythingLLM were just more patient, there is a high degree of probability it would all work out.
It would be nice to have this on desktop too.. because why not? It's a simple field.
Seems it would also be good for people running local models that might be a bit on the slow side for performance.
The mobile app times out so fast its dumb as hell.
@Sinkingdev commented on GitHub (Dec 27, 2025):
I don't know about the phone instance. But as for desktop, at least the Docker instance.
You can set the timeout limit, although it's placed in a fairly hidden spot.
For workspace settings:
For system settings:
This is something that I think should be changed; the timeout is both too hidden and too short by default. For me, many models just wouldn’t work at all. @timothycarambat, what do you think? Could we surface this control directly on the main LLM preference and workspace preference pages instead of hiding it behind a dropdown? It’s a clever spot, but I don't think most users interact with the provider dropdown often enough to stumble on it.
I also searched the docs for “timeout” and “provider timeout” without finding any guidance on how to do this. I might do a write-up for the docs, but I also think the UI change would be the better fix. Which solution would you prefer? UI changes are probably beyond my scope, so I’d be happy to document the current workflow if that helps in the meantime.
[FEAT]: Allow configuring time-outto [GH-ISSUE #4700] [FEAT]: Allow configuring time-out@MikeDuncan1 commented on GitHub (Mar 25, 2026):
This is a huge issue with a seemingly simple fix, when I get the model pre loaded it works so well but when it unloads it's another trip to the desktop to load it up again. I would keep it loaded in memory on the desktop but other models need to load periodically for tasks and a man can only afford so much vram
@timothycarambat commented on GitHub (Mar 26, 2026):
What provider are you using? If the VRAM is competitive and other models are pushing it out we cannot keep it in memory, but we can (and do) have an option for
keepAlive- at least for Ollama. Other providers have different ways of handling or specifying this.The other comments are talking about the overall timeout, like for the HTTP request, not the model unload
@elevatingcreativity commented on GitHub (Apr 2, 2026):
I am having the same issue as the OP, and have a proposed fix to submit for @timothycarambat's review.
More details:
Provider: LM Studio
Symptom: Chat returns "Could not respond to message / Socket Timeout" (or similar)
when running large local models (e.g. Qwen2.5-72B+ or similar multi-hundred GB models)
on long prompts. The error occurs before the model finishes generating.
Root cause: The LM Studio provider uses the OpenAI SDK with no custom timeout
configured, so it inherits the SDK's hardcoded default of 10 minutes (timeout:
600000). For large local models, prompt processing alone (before the first token is
produced) can easily exceed this — there is no data flowing on the connection during
this phase, so there's nothing to listen to; the timeout simply fires.
Comparison with other providers: Ollama already has OLLAMA_RESPONSE_TIMEOUT,
OpenRouter has OPENROUTER_TIMEOUT_MS, and Novita has NOVITA_LLM_TIMEOUT_MS. LM Studio
has none.
Proposed fix: Add LMSTUDIO_RESPONSE_TIMEOUT (in ms) env var support, defaulting to 2
hours.
Working branch: fix/lmstudio-response-timeout on https://github.com/elevatingcreativity/anything-llm. Will submit a PR once tested.