Prompt Caching - GPT/Gemini/Anthropic Models #6600

Closed
opened 2026-02-21 18:16:29 -05:00 by yindo · 1 comment
Owner

Originally created by @LateralusAtCrezam on GitHub (Nov 6, 2024).

Self Checks

  • I have searched for existing issues search for existing issues, including closed ones.
  • I confirm that I am using English to submit this report (我已阅读并同意 Language Policy).
  • [FOR CHINESE USERS] 请务必使用英文提交 Issue,否则会被关闭。谢谢!:)
  • Please do not modify this template :) and fill in all the required fields.

1. Is this request related to a challenge you're experiencing? Tell me about your story.

I am prototyping with Dify Self-Hosted for one of my projects, and I am excited about the possibility of doing context caching within Dify itself. There is a whole load of applications to this feature, including more efficient and faster document query, as well as (for smaller documents) completely RAG-free query.

2. Additional context or comments

No response

3. Can you help us with this feature?

  • I am interested in contributing to this feature.
Originally created by @LateralusAtCrezam on GitHub (Nov 6, 2024). ### Self Checks - [X] I have searched for existing issues [search for existing issues](https://github.com/langgenius/dify/issues), including closed ones. - [X] I confirm that I am using English to submit this report (我已阅读并同意 [Language Policy](https://github.com/langgenius/dify/issues/1542)). - [X] [FOR CHINESE USERS] 请务必使用英文提交 Issue,否则会被关闭。谢谢!:) - [X] Please do not modify this template :) and fill in all the required fields. ### 1. Is this request related to a challenge you're experiencing? Tell me about your story. I am prototyping with Dify Self-Hosted for one of my projects, and I am excited about the possibility of doing context caching within Dify itself. There is a whole load of applications to this feature, including more efficient and faster document query, as well as (for smaller documents) completely RAG-free query. ### 2. Additional context or comments _No response_ ### 3. Can you help us with this feature? - [X] I am interested in contributing to this feature.
yindo added the 💪 enhancement label 2026-02-21 18:16:29 -05:00
yindo closed this issue 2026-02-21 18:16:29 -05:00
Author
Owner

@dosubot[bot] commented on GitHub (Dec 7, 2024):

Hi, @LateralusAtCrezam. I'm Dosu, and I'm helping the Dify team manage their backlog. I'm marking this issue as stale.

Issue Summary:

  • Feature request for implementing prompt caching in Dify Self-Hosted.
  • Targets GPT, Gemini, and Anthropic models to enhance efficiency and speed.
  • You have expressed willingness to contribute to the development.
  • No comments or further activity on the issue yet.

Next Steps:

  • Is this issue still relevant to the latest version of the Dify repository? If so, please comment to keep the discussion open.
  • Otherwise, the issue will be automatically closed in 15 days.

Thank you for your understanding and contribution!

@dosubot[bot] commented on GitHub (Dec 7, 2024): Hi, @LateralusAtCrezam. I'm [Dosu](https://dosu.dev), and I'm helping the Dify team manage their backlog. I'm marking this issue as stale. **Issue Summary:** - Feature request for implementing prompt caching in Dify Self-Hosted. - Targets GPT, Gemini, and Anthropic models to enhance efficiency and speed. - You have expressed willingness to contribute to the development. - No comments or further activity on the issue yet. **Next Steps:** - Is this issue still relevant to the latest version of the Dify repository? If so, please comment to keep the discussion open. - Otherwise, the issue will be automatically closed in 15 days. Thank you for your understanding and contribution!
Sign in to join this conversation.
1 Participants
Notifications
Due Date
No due date set.
Dependencies

No dependencies set.

Reference: langgenius/dify#6600