mirror of
https://github.com/Mintplex-Labs/anything-llm.git
synced 2026-07-21 01:25:28 -04:00
Open
opened 2026-02-22 18:20:14 -05:00 by yindo
·
17 comments
No Branch/Tag Specified
master
5846-bug-when-scrolling-up-scrolling-jumps
2235-translations
2235-bug-how-to-upload-a-folder-with-subfolders-with-files-to-anythingllm
refactor-remove-workspace-pfp
5969-bug-stopgenerationbutton-disappears-on-the-first-prompt-that-initiates-an-agent-session
5968-chore-update-default-system-prompt-for-better-tool-use
5990-workspace-update-fails-with-unknown-argument-router_id-v1130v1150-intel-mac
opencomputer-examples
pg
feat/image-generation-translations
feat/image-generation
5924-bug-meeting-summary-fails-with-sincludes-is-not-a-function-when-default-llm-is-anthropic-claude
render
feat/uniform-modal-component
5883-bug-unescaped-content-in-json-strings-being-passed-to-document-generator-tools
5901-bug-api-update-embeddings-fails-prisma-argument-filename-is-missing-on-workspace_documentscreate-desktop-windows-v1141
hybrid-search
1981-translations
5752-bug-prompts-to-local-jan-endpoint-unresponsive
feat-disable-native-tool-calling-env-var
5676-bug-non-ollama-agent-providers-do-not-parse-and-present-reasoning-content
feat/markdown-web-scraping
5717-bug-apiv1documentupload-silently-drops-metadata-field-in-desktop-1130-arg-count-mismatch-nested-payload-key
fix/aibitat-context-overflow
5711-bug-erratic-deepseek-v4-flash-the-agent-model-failed-to-respond-400-the-reasoning_content
5631-feat-custom-api-request-timeouts-for-ai-providers
feat-reasoning-control
feat-agent-clarifying-questions-translations
5583-bug-lm-studio-provider-does-not-present-reasoning-output
5313-normalize-translations
feat/memory-translations
5305-lemonade-embedding-engine-swallows-errors-falsely-reports-documents-as-embedded
5060-bug-agent-interactions-agent-are-not-persisted-to-thread-history-via-api
pptx-subagent
feat-render-images-from-mcp-tool-results
stt-provider-expansion-openai-api-compatible
feat-file-search-agent-tool
feat-native-embedder-job-queue
5189-normalize-translations
feat-file-search-agent-tool-translations
i18n-eslint
5140-auto-migration
5112-bug-openrouter-failed-message-bug
3506-feat-parameters-for-openrouter-models
4992-feat-preserve-scroll-position
4973-bug-markdown-numbered-list-display-in-reasoning-pane
desktop
4938-bug-pending-chat-rerendering-ui-bug
quickstart-env
node-llama-cpp-in-container-cuda
node-llama-cpp-in-container
ollama-in-container
4817-feat-set-cooldown-per-mcp-server
4845-keyboard-shortcuts-to-navigate-in-chat
4844-feat-reorder-threads-by-latest-interaction
standardize-username-constraints-normalize-translations
4792-feat-refactor-workspacepfp-image
1382-embed-ip-improvements
1382-bug-embed-api-improvements
refactor-eslint-frontend
4687-feat-refactor-vector-db-providers
4615-feat-disable-apidocs-with-environment-variable
4559-feat-agent-web-search-enable-ordering-of-results
4599-bug-ollama-race-condition-bug
4572-bug-lmstudio-provided-llm-stopped-working-with-anything-llm-after-upgrading-to-190
4508-agent-youtube-transcript-analysis
4497-feat-workspace-names
frontend-eslint
ollama-lmstudio-auto-context-window
4431-validate-vector-database-connectioN
2019-slash-command-keyboard-selection
microsoft-foundry-provider
4431-validate-vector-database-connection
4325-sys-prompt-var-improvements
3209-feat-apiv1workspacestream-chat-sources-citations
4210-bug-voice-to-text-overwrite
4136-feat-jan-as-a-backend-server-option
4172-feat-openai-o3-support
1.8.3-rerelease
web-push-notifications-service
tasks
3955-feat-jinaai-embedder-provider-support
3921-feat-agent-skills-uiux-improvements
3901-bug-validfunccall-checks-optional-arguments
keyboard-dev
1787-custom-roles-and-permissions
add-jira-slack-data-connector
office-extension-wip
lightmode-dropdown-color-update
3586-bug-agent-flow-function-description-provided-by-user-is-not-seen-in-the-llm-query
3463-bug-agent-continues-to-run-if-request-failed-even-after-exit
3439-feat-call-variables-within-the-flow-api-block-url-field
3282-manager-view-models-workspace
3280-token-counting-server-side-truncation-improvements
3147-bug-embedded-chat-widget---not-considering-query-mode-option-always-working-in-chat-mode
2995-feat-disable-temperature-setting-for-deepseek-r1-deepseek-reasoner-model
2827-feat-perplexity-citations
2866-feat-finally-a-gemini-models-endpoint
2647-feat-hpp-header-for-a-c++-code-file-mime-addition
lancedb-revert
1656-feat-implement-tooltip-ui-designs
2011-feat-bump-perplexity-models
1873-feat-auto-add-and-watch-folder-for-document-uploads
1297-feat-gemini-agent-support
1759-bug-ui-bug-fixes
1686-feat-implement-winston-for-logging
1536-bug-toggling-on-users-can-delete-workspaces-does-not-take-effect
agent-ui-mobile-styles
1522-feat-chromadb-support
1595-bug-unable-to-get-live-web-search-and-browsing-agent-working-using-google-custom-search-engine-error-getaddrinfo-enotfound-http-errno-3008
1582-bug-lm-studio-does-not-allow-for-different-model-selection
1312-bug-usernames-should-not-be-case-sensitive-when-logging-in
1029-feat-hf-serverless-inference-api
1086-feat-implement-normalized-input-fields
knowledge-graph-support
644-bug-uploaded-file-name-does-not-match-the-displayed-file-name-after-the-upload
v1.15.0
v1.14.2
v1.14.1
v1.14.0
v1.13.0
v1.12.1
v1.12.0
v1.11.2
v1.11.1
v1.11.0
v1.10.0
v1.9.1
v1.9.0
v1.8.5
v1.8.4
v1.8.3
v1.8.2
v1.8.1
v1.8.0
v1.7.8
v1.7.6
v1.7.5
v1.7.4
v1.4.0
v1.3.0
v1.2.4
v1.2.3
v1.2.2
v1.2.1
v1.2.0
v1.1.1
v1.1.0
v1.0.0
Labels
Clear labels
Desktop
Docker
Integration Request
Integration Request
OS: Linux
OS: Mobile
OS: Windows
UI/UX
blocked
bug
bug
core-team-only
documentation
duplicate
embed-widget
enhancement
feature request
github_actions
good first issue
investigating
needs info / can't replicate
possible bug
pull-request
question
stage: specifications
wontfix
Mirrored from GitHub Pull Request
Milestone
No items
No Milestone
Projects
Clear projects
No project
Notifications
Due Date
No due date set.
Dependencies
No dependencies set.
Reference: Mintplex-Labs/anything-llm#573
Reference in New Issue
Block a user
Blocking a user prevents them from interacting with repositories, such as opening or commenting on pull requests or issues. Learn more about blocking a user.
Delete Branch "%!s()"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
Originally created by @Mirgiacomo on GitHub (Mar 21, 2024).
Original GitHub issue: https://github.com/Mintplex-Labs/anything-llm/issues/943
Originally assigned to: @timothycarambat on GitHub.
What would you like to see?
In addition to text generation, is it possible to allow anythingLLM to generate images?
Via the image generation API of OpenAI, or Gemini, etc.?
PS: i mean text to image and not image to text
@ytjhai commented on GitHub (Sep 5, 2024):
I'm also interested in this, specifically around the way images can be "retrieved" from existing documentation rather than generated from scratch using something like DALL-E. Embedding images is currently possible with some LLMs but some things can only be communicated through the original image. So, retrieving that image in its original form would help a lot. The "Show Citations" modal would be the part where this can be visible.
@appdevhm88mt commented on GitHub (Oct 23, 2024):
Hi, regarding the function of generating images from text, is there any estimate of when or which version it will be supported by?
Thx
@severfire commented on GitHub (Jan 28, 2025):
Its something very much needed. Embedding images and generating images or links to them.
@keramblock commented on GitHub (Feb 14, 2025):
also Flux.1 would be cool
@morbificagent commented on GitHub (Mar 21, 2025):
Oh yes, would be very great!
@timothycarambat in #3386 you say that there isnt a model that can be run local and gives good quality on low end devices... i understand that but using an online model would be great anyway :-)
@psimm commented on GitHub (Apr 4, 2025):
My team and I would appreciate image generation via an external API, e.g. a FLUX or Stable Diffusion model.
@Arp1tMoga commented on GitHub (Apr 21, 2025):
+1
@timothycarambat commented on GitHub (Apr 21, 2025):
please dont comment +1
We cannot track comments of
+1- leave a reaction so that it gets bumped!Filter we use: https://github.com/Mintplex-Labs/anything-llm/issues?q=is%3Aissue+is%3Aopen+sort%3Areactions-desc+
@acuentascanarias commented on GitHub (Apr 22, 2025):
We have the same need as @Mirgiacomo . Is it possible to allow anythingLLM to generate images?
I mean text to image (image to text already available)
@IronBeardKnight commented on GitHub (May 19, 2025):
Having ability to use even localai's ability to do image generating, tts etc would be very cool
@dimdieo commented on GitHub (Jun 24, 2025):
It would be great if they made support for Comfyui
@cannin commented on GitHub (Sep 30, 2025):
@timothycarambat Below is a simple MCP server to get an image from an external source that could be used for image generation. Using @agent will recognize the commands. The problem is that the tiny image (star.png) is returned and shown, but not the large one (star_big.png); larger images always seem to take longer and get severely cutoff. Not sure if it is a AnythingLLM or Python FastMCP problem.
MCP Config
From: anythingllm_mcp_servers.json
MCP Server
Related to: https://github.com/Mintplex-Labs/anything-llm/issues/1526
@manuelkamp commented on GitHub (Oct 19, 2025):
any progress or update on this feature?
I'd really like to use image generation with my localai in anythingllm too (since I use anythingllm as my primary ai tool). now I have to use nextcloud assistant for image generations :)
@ChupaUps commented on GitHub (Feb 2, 2026):
Following up on previous requests about image generation support. Any progress on multi-modal outputs (images/files) from LLMs? Currently using external tools for this while Anything LLM remains my primary AI tool. Looking for updates or workarounds.
[FEAT]: Image generationto [GH-ISSUE #943] [FEAT]: Image generation@jjnxpct commented on GitHub (Mar 4, 2026):
I recently started using AnythingLLM and I love the platform, but I was surprised to find that image generation is not yet supported. This is something I was really hoping to offer my clients.
I tried working around this by building a custom agent skill that calls the BFL Flux API, and the skill itself works correctly — but I can't get it to function reliably because I exclusively want to use EU-based models (Mistral). Mistral models consistently fail to pass arguments to the tool call, so the skill is invoked but receives empty parameters every time.
It would be great to see native image generation support added — ideally with the ability to connect models like Flux (BFL) for generation and editing. This would be a natural fit for a platform with such strong agent capabilities.
Hoping to see this on the roadmap soon.
@DDesmond95 commented on GitHub (Apr 23, 2026):
Still not supported?
@LeighBicknell commented on GitHub (May 3, 2026):
Also wanting support for this.