mirror of
https://github.com/Mintplex-Labs/anything-llm.git
synced 2026-07-19 14:13:35 -04:00
[GH-ISSUE #2243] [BUG]: GitHub Repo Fetch Failed message appears but then proceeds without notification #1463
Closed
opened 2026-02-22 18:24:56 -05:00 by yindo
·
4 comments
No Branch/Tag Specified
master
5846-bug-when-scrolling-up-scrolling-jumps
refactor-remove-workspace-pfp
5969-bug-stopgenerationbutton-disappears-on-the-first-prompt-that-initiates-an-agent-session
2235-bug-how-to-upload-a-folder-with-subfolders-with-files-to-anythingllm
5990-workspace-update-fails-with-unknown-argument-router_id-v1130v1150-intel-mac
opencomputer-examples
pg
feat/image-generation-translations
feat/image-generation
5924-bug-meeting-summary-fails-with-sincludes-is-not-a-function-when-default-llm-is-anthropic-claude
render
feat/uniform-modal-component
5883-bug-unescaped-content-in-json-strings-being-passed-to-document-generator-tools
5901-bug-api-update-embeddings-fails-prisma-argument-filename-is-missing-on-workspace_documentscreate-desktop-windows-v1141
hybrid-search
1981-translations
5752-bug-prompts-to-local-jan-endpoint-unresponsive
feat-disable-native-tool-calling-env-var
5676-bug-non-ollama-agent-providers-do-not-parse-and-present-reasoning-content
feat/markdown-web-scraping
5717-bug-apiv1documentupload-silently-drops-metadata-field-in-desktop-1130-arg-count-mismatch-nested-payload-key
fix/aibitat-context-overflow
5711-bug-erratic-deepseek-v4-flash-the-agent-model-failed-to-respond-400-the-reasoning_content
5631-feat-custom-api-request-timeouts-for-ai-providers
feat-reasoning-control
feat-agent-clarifying-questions-translations
5583-bug-lm-studio-provider-does-not-present-reasoning-output
5313-normalize-translations
feat/memory-translations
5305-lemonade-embedding-engine-swallows-errors-falsely-reports-documents-as-embedded
5060-bug-agent-interactions-agent-are-not-persisted-to-thread-history-via-api
pptx-subagent
feat-render-images-from-mcp-tool-results
stt-provider-expansion-openai-api-compatible
feat-file-search-agent-tool
feat-native-embedder-job-queue
5189-normalize-translations
feat-file-search-agent-tool-translations
i18n-eslint
5140-auto-migration
5112-bug-openrouter-failed-message-bug
3506-feat-parameters-for-openrouter-models
4992-feat-preserve-scroll-position
4973-bug-markdown-numbered-list-display-in-reasoning-pane
desktop
4938-bug-pending-chat-rerendering-ui-bug
quickstart-env
node-llama-cpp-in-container-cuda
node-llama-cpp-in-container
ollama-in-container
4817-feat-set-cooldown-per-mcp-server
4845-keyboard-shortcuts-to-navigate-in-chat
4844-feat-reorder-threads-by-latest-interaction
standardize-username-constraints-normalize-translations
4792-feat-refactor-workspacepfp-image
1382-embed-ip-improvements
1382-bug-embed-api-improvements
refactor-eslint-frontend
4687-feat-refactor-vector-db-providers
4615-feat-disable-apidocs-with-environment-variable
4559-feat-agent-web-search-enable-ordering-of-results
4599-bug-ollama-race-condition-bug
4572-bug-lmstudio-provided-llm-stopped-working-with-anything-llm-after-upgrading-to-190
4508-agent-youtube-transcript-analysis
4497-feat-workspace-names
frontend-eslint
ollama-lmstudio-auto-context-window
4431-validate-vector-database-connectioN
2019-slash-command-keyboard-selection
microsoft-foundry-provider
4431-validate-vector-database-connection
4325-sys-prompt-var-improvements
3209-feat-apiv1workspacestream-chat-sources-citations
4210-bug-voice-to-text-overwrite
4136-feat-jan-as-a-backend-server-option
4172-feat-openai-o3-support
1.8.3-rerelease
web-push-notifications-service
tasks
3955-feat-jinaai-embedder-provider-support
3921-feat-agent-skills-uiux-improvements
3901-bug-validfunccall-checks-optional-arguments
keyboard-dev
1787-custom-roles-and-permissions
add-jira-slack-data-connector
office-extension-wip
lightmode-dropdown-color-update
3586-bug-agent-flow-function-description-provided-by-user-is-not-seen-in-the-llm-query
3463-bug-agent-continues-to-run-if-request-failed-even-after-exit
3439-feat-call-variables-within-the-flow-api-block-url-field
3282-manager-view-models-workspace
3280-token-counting-server-side-truncation-improvements
3147-bug-embedded-chat-widget---not-considering-query-mode-option-always-working-in-chat-mode
2995-feat-disable-temperature-setting-for-deepseek-r1-deepseek-reasoner-model
2827-feat-perplexity-citations
2866-feat-finally-a-gemini-models-endpoint
2647-feat-hpp-header-for-a-c++-code-file-mime-addition
lancedb-revert
1656-feat-implement-tooltip-ui-designs
2011-feat-bump-perplexity-models
1873-feat-auto-add-and-watch-folder-for-document-uploads
1297-feat-gemini-agent-support
1759-bug-ui-bug-fixes
1686-feat-implement-winston-for-logging
1536-bug-toggling-on-users-can-delete-workspaces-does-not-take-effect
agent-ui-mobile-styles
1522-feat-chromadb-support
1595-bug-unable-to-get-live-web-search-and-browsing-agent-working-using-google-custom-search-engine-error-getaddrinfo-enotfound-http-errno-3008
1582-bug-lm-studio-does-not-allow-for-different-model-selection
1312-bug-usernames-should-not-be-case-sensitive-when-logging-in
1029-feat-hf-serverless-inference-api
1086-feat-implement-normalized-input-fields
knowledge-graph-support
644-bug-uploaded-file-name-does-not-match-the-displayed-file-name-after-the-upload
v1.15.0
v1.14.2
v1.14.1
v1.14.0
v1.13.0
v1.12.1
v1.12.0
v1.11.2
v1.11.1
v1.11.0
v1.10.0
v1.9.1
v1.9.0
v1.8.5
v1.8.4
v1.8.3
v1.8.2
v1.8.1
v1.8.0
v1.7.8
v1.7.6
v1.7.5
v1.7.4
v1.4.0
v1.3.0
v1.2.4
v1.2.3
v1.2.2
v1.2.1
v1.2.0
v1.1.1
v1.1.0
v1.0.0
Labels
Clear labels
Desktop
Docker
Integration Request
Integration Request
OS: Linux
OS: Mobile
OS: Windows
UI/UX
blocked
bug
bug
core-team-only
documentation
duplicate
embed-widget
enhancement
feature request
github_actions
good first issue
investigating
needs info / can't replicate
possible bug
pull-request
question
stage: specifications
wontfix
Mirrored from GitHub Pull Request
No Label
needs info / can't replicate
Milestone
No items
No Milestone
Projects
Clear projects
No project
Notifications
Due Date
No due date set.
Dependencies
No dependencies set.
Reference: Mintplex-Labs/anything-llm#1463
Reference in New Issue
Block a user
Blocking a user prevents them from interacting with repositories, such as opening or commenting on pull requests or issues. Learn more about blocking a user.
Delete Branch "%!s()"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
Originally created by @Sch0field9 on GitHub (Sep 8, 2024).
Original GitHub issue: https://github.com/Mintplex-Labs/anything-llm/issues/2243
How are you running AnythingLLM?
AnythingLLM desktop app
What happened?
When fetching a specific Github repo it hangs a long time after Access token set! Recursive loading enabled!
after a minute or two the message "Fetch failed" appears. Neither the GUI or Logs indicate any further processes. However after doing nothing for another minute the fetch continues and downloads the files from the repo. This does not happen on all repos
[collector] info: [Github Loader]: Access token set! Recursive loading enabled! [backend] info: [CollectorApi] fetch failed [backend] info: [TELEMETRY SENT] {"event":"extension_invoked","distinctId":"372a5230-9bfd-4c05-a653-ae4d165a0c67","properties":{"type":"github_repo","runtime":"desktop"}} [collector] info: [Github Loader]: Found 3952 source files. Saving...Are there known steps to reproduce?
No response
@timothycarambat commented on GitHub (Sep 9, 2024):
Is this repo public for repro?
Additionally, is there any more content to the logs or does is simply only output
[backend] info: [CollectorApi] fetch failedwith nothing else in that line?Given the source have nearly 4K files, I am willing to bet this is related to downloading 4K files via the PAT and it hitting rate limits.
@Sch0field9 commented on GitHub (Sep 9, 2024):
Those are all the logs I have. I couldn't figure out how to change the log level.
The repo I had the troubles with was Github home-assistant/core:master --
(Not sure if it was core)
However, I had something similar happen when I tried it with the AnythinLLM repo. It startet way faster, but it somewhere in the middle was a fetch failed log entry, only one and it continued downloading. The Frontend had the same behaviour tough. Message came up fetch failed and no other indications that it is still downloading files.
I also suspect it has to do with large repos. My bug reports focus is more about that there is no indication to the user, less that why exactly it failed. And possibly a feature request to be able to change the log level? Since there is also no indication in the info logs, that its still processing.
@timothycarambat commented on GitHub (Sep 9, 2024):
Adding verbose logging to the tool at least so the process can be seen in logs even if the operator/user cannot see it in UI currently. Did pull in AnythingLLM repo without errors at this time
[BUG]: GitHub Repo Fetch Failed message appears but then proceeds without notificationto [GH-ISSUE #2243] [BUG]: GitHub Repo Fetch Failed message appears but then proceeds without notification@edddeduck commented on GitHub (Feb 26, 2026):
I have reopened the bug as I have found a 100% reproducible test case using a public documentation website that is not too large. I have been able to reproduce this using v1.11.0.
I told it to capture the website for https://docs.mod.io The site has 669 pages.
After about 10 minutes I got a message in the UI saying that the fetch failed. I checked the logs and could see this.
However I noted the CPU usage was using just over 100% (where 100% indicates one core being fully utilised) so something was obviously happening.
NOTE: I can using another tool check you can load all the pages on the site in under a minute so it feels strange it took close to 30 mins to review the full site, but without extra logging I am not sure what was taking the time here.
After another approximately 20 minutes the logs started exporting the following:
It would be nice to get some progress information in the logs (and the UI) as it seems if you have a large capture requested it'll always time out and say the fetch failed even though it's actually just a large site to request.
The tool is great but having to monitor CPU usage in the terminal, the log file and the UI to try and work out what is happening makes it a little hard. The biggest issue is the large delay between the failed message and the found links to scrape message. Based on logs it indicates things have failed but if you check the CPU usage it's clear something was happening so if you're patient it'll work in the end.
Given it would be useful to ingest larger data sets than this one a more robust UX would be a great improvement. Happy to provide more information as needed.