mirror of
https://github.com/Mintplex-Labs/anything-llm.git
synced 2026-07-21 01:25:28 -04:00
Closed
opened 2026-06-05 15:21:34 -04:00 by yindo
·
0 comments
No Branch/Tag Specified
master
5846-bug-when-scrolling-up-scrolling-jumps
2235-translations
2235-bug-how-to-upload-a-folder-with-subfolders-with-files-to-anythingllm
refactor-remove-workspace-pfp
5969-bug-stopgenerationbutton-disappears-on-the-first-prompt-that-initiates-an-agent-session
5968-chore-update-default-system-prompt-for-better-tool-use
5990-workspace-update-fails-with-unknown-argument-router_id-v1130v1150-intel-mac
opencomputer-examples
pg
feat/image-generation-translations
feat/image-generation
5924-bug-meeting-summary-fails-with-sincludes-is-not-a-function-when-default-llm-is-anthropic-claude
render
feat/uniform-modal-component
5883-bug-unescaped-content-in-json-strings-being-passed-to-document-generator-tools
5901-bug-api-update-embeddings-fails-prisma-argument-filename-is-missing-on-workspace_documentscreate-desktop-windows-v1141
hybrid-search
1981-translations
5752-bug-prompts-to-local-jan-endpoint-unresponsive
feat-disable-native-tool-calling-env-var
5676-bug-non-ollama-agent-providers-do-not-parse-and-present-reasoning-content
feat/markdown-web-scraping
5717-bug-apiv1documentupload-silently-drops-metadata-field-in-desktop-1130-arg-count-mismatch-nested-payload-key
fix/aibitat-context-overflow
5711-bug-erratic-deepseek-v4-flash-the-agent-model-failed-to-respond-400-the-reasoning_content
5631-feat-custom-api-request-timeouts-for-ai-providers
feat-reasoning-control
feat-agent-clarifying-questions-translations
5583-bug-lm-studio-provider-does-not-present-reasoning-output
5313-normalize-translations
feat/memory-translations
5305-lemonade-embedding-engine-swallows-errors-falsely-reports-documents-as-embedded
5060-bug-agent-interactions-agent-are-not-persisted-to-thread-history-via-api
pptx-subagent
feat-render-images-from-mcp-tool-results
stt-provider-expansion-openai-api-compatible
feat-file-search-agent-tool
feat-native-embedder-job-queue
5189-normalize-translations
feat-file-search-agent-tool-translations
i18n-eslint
5140-auto-migration
5112-bug-openrouter-failed-message-bug
3506-feat-parameters-for-openrouter-models
4992-feat-preserve-scroll-position
4973-bug-markdown-numbered-list-display-in-reasoning-pane
desktop
4938-bug-pending-chat-rerendering-ui-bug
quickstart-env
node-llama-cpp-in-container-cuda
node-llama-cpp-in-container
ollama-in-container
4817-feat-set-cooldown-per-mcp-server
4845-keyboard-shortcuts-to-navigate-in-chat
4844-feat-reorder-threads-by-latest-interaction
standardize-username-constraints-normalize-translations
4792-feat-refactor-workspacepfp-image
1382-embed-ip-improvements
1382-bug-embed-api-improvements
refactor-eslint-frontend
4687-feat-refactor-vector-db-providers
4615-feat-disable-apidocs-with-environment-variable
4559-feat-agent-web-search-enable-ordering-of-results
4599-bug-ollama-race-condition-bug
4572-bug-lmstudio-provided-llm-stopped-working-with-anything-llm-after-upgrading-to-190
4508-agent-youtube-transcript-analysis
4497-feat-workspace-names
frontend-eslint
ollama-lmstudio-auto-context-window
4431-validate-vector-database-connectioN
2019-slash-command-keyboard-selection
microsoft-foundry-provider
4431-validate-vector-database-connection
4325-sys-prompt-var-improvements
3209-feat-apiv1workspacestream-chat-sources-citations
4210-bug-voice-to-text-overwrite
4136-feat-jan-as-a-backend-server-option
4172-feat-openai-o3-support
1.8.3-rerelease
web-push-notifications-service
tasks
3955-feat-jinaai-embedder-provider-support
3921-feat-agent-skills-uiux-improvements
3901-bug-validfunccall-checks-optional-arguments
keyboard-dev
1787-custom-roles-and-permissions
add-jira-slack-data-connector
office-extension-wip
lightmode-dropdown-color-update
3586-bug-agent-flow-function-description-provided-by-user-is-not-seen-in-the-llm-query
3463-bug-agent-continues-to-run-if-request-failed-even-after-exit
3439-feat-call-variables-within-the-flow-api-block-url-field
3282-manager-view-models-workspace
3280-token-counting-server-side-truncation-improvements
3147-bug-embedded-chat-widget---not-considering-query-mode-option-always-working-in-chat-mode
2995-feat-disable-temperature-setting-for-deepseek-r1-deepseek-reasoner-model
2827-feat-perplexity-citations
2866-feat-finally-a-gemini-models-endpoint
2647-feat-hpp-header-for-a-c++-code-file-mime-addition
lancedb-revert
1656-feat-implement-tooltip-ui-designs
2011-feat-bump-perplexity-models
1873-feat-auto-add-and-watch-folder-for-document-uploads
1297-feat-gemini-agent-support
1759-bug-ui-bug-fixes
1686-feat-implement-winston-for-logging
1536-bug-toggling-on-users-can-delete-workspaces-does-not-take-effect
agent-ui-mobile-styles
1522-feat-chromadb-support
1595-bug-unable-to-get-live-web-search-and-browsing-agent-working-using-google-custom-search-engine-error-getaddrinfo-enotfound-http-errno-3008
1582-bug-lm-studio-does-not-allow-for-different-model-selection
1312-bug-usernames-should-not-be-case-sensitive-when-logging-in
1029-feat-hf-serverless-inference-api
1086-feat-implement-normalized-input-fields
knowledge-graph-support
644-bug-uploaded-file-name-does-not-match-the-displayed-file-name-after-the-upload
v1.15.0
v1.14.2
v1.14.1
v1.14.0
v1.13.0
v1.12.1
v1.12.0
v1.11.2
v1.11.1
v1.11.0
v1.10.0
v1.9.1
v1.9.0
v1.8.5
v1.8.4
v1.8.3
v1.8.2
v1.8.1
v1.8.0
v1.7.8
v1.7.6
v1.7.5
v1.7.4
v1.4.0
v1.3.0
v1.2.4
v1.2.3
v1.2.2
v1.2.1
v1.2.0
v1.1.1
v1.1.0
v1.0.0
Labels
Clear labels
Desktop
Docker
Integration Request
Integration Request
OS: Linux
OS: Mobile
OS: Windows
UI/UX
blocked
bug
bug
core-team-only
documentation
duplicate
embed-widget
enhancement
feature request
github_actions
good first issue
investigating
needs info / can't replicate
possible bug
pull-request
question
stage: specifications
wontfix
Mirrored from GitHub Pull Request
No Label
pull-request
Milestone
No items
No Milestone
Projects
Clear projects
No project
Notifications
Due Date
No due date set.
Dependencies
No dependencies set.
Reference: Mintplex-Labs/anything-llm#5506
Reference in New Issue
Block a user
Blocking a user prevents them from interacting with repositories, such as opening or commenting on pull requests or issues. Learn more about blocking a user.
Delete Branch "%!s()"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
📋 Pull Request Information
Original PR: https://github.com/Mintplex-Labs/anything-llm/pull/5637
Author: @angelplusultra
Created: 5/15/2026
Status: ❌ Closed
Base:
master← Head:5631-feat-custom-api-request-timeouts-for-ai-providers📝 Commits (6)
2205117add getCustomFetchWithTimeout helper function and apply to all llm providers, agent providers and embedding providers02a4892revise .env.example commentsc7f32b8add anthropic6e9928bfix default ollama socket timeout value8afad0fhoist agent instantiation out of custom fetch functionb24c8faMerge branch 'master' into 5631-feat-custom-api-request-timeouts-for-ai-providers📊 Changes
81 files changed (+1022 additions, -60 deletions)
View changed files
📝
server/.env.example(+33 -1)➕
server/__tests__/utils/AiProviders/helpers/index.test.js(+185 -0)📝
server/utils/AiProviders/anthropic/index.js(+9 -0)📝
server/utils/AiProviders/apipie/index.js(+9 -0)📝
server/utils/AiProviders/azureOpenAi/index.js(+9 -0)📝
server/utils/AiProviders/cometapi/index.js(+9 -0)📝
server/utils/AiProviders/deepseek/index.js(+9 -0)📝
server/utils/AiProviders/dellProAiStudio/index.js(+9 -0)📝
server/utils/AiProviders/dockerModelRunner/index.js(+5 -0)📝
server/utils/AiProviders/fireworksAi/index.js(+9 -0)📝
server/utils/AiProviders/foundry/index.js(+10 -5)📝
server/utils/AiProviders/gemini/index.js(+9 -0)📝
server/utils/AiProviders/genericOpenAi/index.js(+9 -0)📝
server/utils/AiProviders/giteeai/index.js(+9 -0)📝
server/utils/AiProviders/groq/index.js(+9 -0)➕
server/utils/AiProviders/helpers/index.js(+86 -0)📝
server/utils/AiProviders/huggingface/index.js(+9 -0)📝
server/utils/AiProviders/koboldCPP/index.js(+9 -0)📝
server/utils/AiProviders/lemonade/index.js(+11 -0)📝
server/utils/AiProviders/liteLLM/index.js(+9 -0)...and 61 more files
📄 Description
Pull Request Type
Relevant Issues
resolves #5631
resolves #5603
resolves #5681
Description
Adds a per-provider, env-configurable socket-idle timeout for every AI provider whose SDK exposes a fetch injection point (the OpenAI SDK, Ollama's client, and the Anthropic SDK), so slow local models stop tripping the SDK's hardcoded 5-minute socket idle timeout and surfacing ERR_SOCKET_TIMEOUT.
New helper
server/utils/AiProviders/helpers/index.js — getFetchWithCustomTimeout(rawTimeoutValue, providerSlog, fallbackTimeoutValue?). Given an env value, returns either the global fetch (when no timeout is resolved) or a wrapped fetch whose dispatcher is an undici.Agent with headersTimeout and bodyTimeout set to the resolved value. This bypasses the OpenAI SDK's node-fetch + agentkeepalive stack — that's where the 5-minute idle cap lives — and applies to gaps between streamed chunks as well as the wait for response headers.
Resolution order: valid positive integer → use it; invalid value → log via the provider's slog and fall through; fallback (when provided) → use it; otherwise → unwrapped fetch.
Note on retries: the OpenAI SDK retries failed requests (default maxRetries: 2), so a timeout-aborted attempt can still be retried up to twice before surfacing. Not changed here, but documented in the helper JSDoc for callers who want a hard wall-clock cap.
Applied to
The helper is wired into every provider that constructs an OpenAI SDK client, an Ollama client, or an Anthropic SDK client across three layers:
Each constructor now passes fetch: getFetchWithCustomTimeout(process.env._RESPONSE_TIMEOUT, .slog). Several #slog private static loggers were promoted to public slog so the helper could surface its setup/parse logs through the right provider prefix.
Ollama refactor
Ollama previously had its own applyOllamaFetch() that only honored timeouts greater than 5 minutes. That bespoke implementation is removed in favor of the shared helper, and a 15-minute fallback (DEFAULT_OLLAMA_SOCKET_TIMEOUT = 900000) is passed in so existing setups keep working without env changes.
Out of scope
Cohere and AWS Bedrock are not covered. Their SDKs do not expose a standard fetch injection point compatible with this helper:
These can be added in follow-ups; they're intentionally deferred to keep this PR focused on the providers the same helper can cover unmodified.
Docs
server/.env.example documents _RESPONSE_TIMEOUT for each affected provider. The comment phrasing makes clear this is a socket-idle timeout (waiting for headers, or a gap between streamed chunks) — not a total request timeout, since a long-running streaming response can legitimately exceed this value.
Visuals (if applicable)
N/A — backend change with no UI surface.
Additional Information
Developer Validations
🔄 This issue represents a GitHub Pull Request. It cannot be merged through Gitea due to API limitations.