[PR #7132] Feat: Add model provider Text Embedding Inference for embedding and rerank #25579

Closed
opened 2026-02-21 20:25:14 -05:00 by yindo · 0 comments
Owner

Original Pull Request: https://github.com/langgenius/dify/pull/7132

State: closed
Merged: Yes


Checklist:

Important

Please review the checklist below before submitting your pull request.

  • Please open an issue before creating a PR or link to an existing issue
  • I have performed a self-review of my own code
  • I have commented my code, particularly in hard-to-understand areas
  • I ran dev/reformat(backend) and cd web && npx lint-staged(frontend) to appease the lint gods

Description

Describe the big picture of your changes here to communicate to the maintainers why we should accept this pull request. If it fixes a bug or resolves a feature request, be sure to link to that issue. Close issue syntax: Fixes #<issue number>, see documentation for more details.

Huggingface Text Embeddings inference (TEI) (https://github.com/huggingface/text-embeddings-inference) is a blazing fast inference solution for text embeddings models. Which include embedding and rerank model support.

For Embedding model, TEI provides OpenAI compatible api, which can be directly used in current dify.
In this pr, the gpt2 tokenizer is replaced by TEI's tokenize api, which provides better token estimation and also supports batch input (currently MAX_CHUNKS is set to 1 by default in the
OpenAI compatible provider)

This pr also add support for TEI rerank model.

During the implementation, also found a small problem in OpenAI compatible provider, probably related to pr #6807, I will fix that in another pr.

Type of Change

  • Bug fix (non-breaking change which fixes an issue)
  • New feature (non-breaking change which adds functionality)
  • Breaking change (fix or feature that would cause existing functionality to not work as expected)
  • This change requires a documentation update, included: Dify Document
  • Improvement, including but not limited to code refactoring, performance optimization, and UI/UX improvement
  • Dependency upgrade

Testing Instructions

Please describe the tests that you ran to verify your changes. Provide instructions so we can reproduce. Please also list any relevant details for your test configuration

  • Add TEI provider test
**Original Pull Request:** https://github.com/langgenius/dify/pull/7132 **State:** closed **Merged:** Yes --- # Checklist: > [!IMPORTANT] > Please review the checklist below before submitting your pull request. - [ ] Please open an issue before creating a PR or link to an existing issue - [x] I have performed a self-review of my own code - [x] I have commented my code, particularly in hard-to-understand areas - [x] I ran `dev/reformat`(backend) and `cd web && npx lint-staged`(frontend) to appease the lint gods # Description Describe the big picture of your changes here to communicate to the maintainers why we should accept this pull request. If it fixes a bug or resolves a feature request, be sure to link to that issue. Close issue syntax: `Fixes #<issue number>`, see [documentation](https://docs.github.com/en/issues/tracking-your-work-with-issues/linking-a-pull-request-to-an-issue#linking-a-pull-request-to-an-issue-using-a-keyword) for more details. Huggingface Text Embeddings inference (TEI) (https://github.com/huggingface/text-embeddings-inference) is a blazing fast inference solution for text embeddings models. Which include embedding and rerank model support. For Embedding model, TEI provides OpenAI compatible api, which can be directly used in current dify. In this pr, the gpt2 tokenizer is replaced by TEI's tokenize api, which provides better token estimation and also supports batch input (currently MAX_CHUNKS is set to 1 by default in the OpenAI compatible provider) This pr also add support for TEI rerank model. During the implementation, also found a small problem in OpenAI compatible provider, probably related to pr #6807, I will fix that in another pr. ## Type of Change - [ ] Bug fix (non-breaking change which fixes an issue) - [x] New feature (non-breaking change which adds functionality) - [ ] Breaking change (fix or feature that would cause existing functionality to not work as expected) - [ ] This change requires a documentation update, included: [Dify Document](https://github.com/langgenius/dify-docs) - [ ] Improvement, including but not limited to code refactoring, performance optimization, and UI/UX improvement - [ ] Dependency upgrade # Testing Instructions Please describe the tests that you ran to verify your changes. Provide instructions so we can reproduce. Please also list any relevant details for your test configuration - [x] Add TEI provider test
yindo added the pull-request label 2026-02-21 20:25:14 -05:00
yindo closed this issue 2026-02-21 20:25:14 -05:00
Sign in to join this conversation.
1 Participants
Notifications
Due Date
No due date set.
Dependencies

No dependencies set.

Reference: langgenius/dify#25579