Request to Add "Doubao-embedding-vision" to Volcano Engine's Text Embedding Models #783

Closed
opened 2026-02-16 10:20:29 -05:00 by yindo · 2 comments
Owner

Originally created by @jiachenglin-creator on GitHub (Nov 6, 2025).

Self Checks

  • I have read the Contributing Guide and Language Policy.
  • I have searched for existing issues search for existing issues, including closed ones.
  • I confirm that I am using English to submit this report, otherwise it will be closed.
  • Please do not modify this template :) and fill in all the required fields.

1. Is this request related to a challenge you're experiencing? Tell me about your story.

I’ve been working on tasks that involve embedding text content associated with visual information—like captions for images, product descriptions with visual details, or mixed media texts that reference visual elements. To handle these, I tried using the Volcano Engine provider in Dify, specifically selecting the "Text Embedding" feature to generate embeddings that could better align with both textual and underlying visual contexts.
However, when I accessed the model options under Volcano Engine’s Text Embedding, I found only "Doubao-embedding" and "Doubao-embedding-large" available. The "Doubao-embedding-vision" model, which I understand is designed to handle text with visual associations more effectively, is missing from the list.
This was frustrating because it forced me to use less suitable models that don’t account for the visual components in my text, leading to less accurate embeddings. It disrupted my workflow, as I had to either compromise on results or find workaround tools outside Dify to get the job done. Having "Doubao-embedding-vision" available would make Dify a one-stop solution for my mixed text-visual embedding needs with Volcano Engine.

2. Additional context or comments

No response

3. Can you help us with this feature?

  • I am interested in contributing to this feature.
Originally created by @jiachenglin-creator on GitHub (Nov 6, 2025). ### Self Checks - [x] I have read the [Contributing Guide](https://github.com/langgenius/dify/blob/main/CONTRIBUTING.md) and [Language Policy](https://github.com/langgenius/dify/issues/1542). - [x] I have searched for existing issues [search for existing issues](https://github.com/langgenius/dify/issues), including closed ones. - [x] I confirm that I am using English to submit this report, otherwise it will be closed. - [x] Please do not modify this template :) and fill in all the required fields. ### 1. Is this request related to a challenge you're experiencing? Tell me about your story. I’ve been working on tasks that involve embedding text content associated with visual information—like captions for images, product descriptions with visual details, or mixed media texts that reference visual elements. To handle these, I tried using the Volcano Engine provider in Dify, specifically selecting the "Text Embedding" feature to generate embeddings that could better align with both textual and underlying visual contexts. However, when I accessed the model options under Volcano Engine’s Text Embedding, I found only "Doubao-embedding" and "Doubao-embedding-large" available. The "Doubao-embedding-vision" model, which I understand is designed to handle text with visual associations more effectively, is missing from the list. This was frustrating because it forced me to use less suitable models that don’t account for the visual components in my text, leading to less accurate embeddings. It disrupted my workflow, as I had to either compromise on results or find workaround tools outside Dify to get the job done. Having "Doubao-embedding-vision" available would make Dify a one-stop solution for my mixed text-visual embedding needs with Volcano Engine. ### 2. Additional context or comments _No response_ ### 3. Can you help us with this feature? - [x] I am interested in contributing to this feature.
yindo added the enhancement label 2026-02-16 10:20:29 -05:00
yindo closed this issue 2026-02-16 10:20:29 -05:00
Author
Owner

@jiachenglin-creator commented on GitHub (Nov 6, 2025):

Image
@jiachenglin-creator commented on GitHub (Nov 6, 2025): <img width="641" height="173" alt="Image" src="https://github.com/user-attachments/assets/ba49bab2-9a98-471f-ab5a-0a825796b73c" />
Author
Owner

@dosubot[bot] commented on GitHub (Nov 22, 2025):

Hi, @jiachenglin-creator. I'm Dosu, and I'm helping the dify-official-plugins team manage their backlog and am marking this issue as stale.

Issue Summary:

  • You requested adding the "Doubao-embedding-vision" model to Volcano Engine's Text Embedding options in Dify.
  • This model is preferred for better handling of text with visual context, improving embedding accuracy for mixed text-visual tasks.
  • You expressed willingness to assist with the implementation.
  • No responses or updates from maintainers have been recorded so far.

Next Steps:

  • Please let me know if this feature request is still relevant to the latest version of dify-official-plugins by commenting on this issue.
  • If I do not hear back within 5 days, I will automatically close the issue.

Thank you for your understanding and contribution!

@dosubot[bot] commented on GitHub (Nov 22, 2025): Hi, @jiachenglin-creator. I'm [Dosu](https://dosu.dev), and I'm helping the dify-official-plugins team manage their backlog and am marking this issue as stale. **Issue Summary:** - You requested adding the "Doubao-embedding-vision" model to Volcano Engine's Text Embedding options in Dify. - This model is preferred for better handling of text with visual context, improving embedding accuracy for mixed text-visual tasks. - You expressed willingness to assist with the implementation. - No responses or updates from maintainers have been recorded so far. **Next Steps:** - Please let me know if this feature request is still relevant to the latest version of dify-official-plugins by commenting on this issue. - If I do not hear back within 5 days, I will automatically close the issue. Thank you for your understanding and contribution!
Sign in to join this conversation.
1 Participants
Notifications
Due Date
No due date set.
Dependencies

No dependencies set.

Reference: langgenius/dify-official-plugins#783