support returning structured output when using LLM API non streaming invocation #18558

Closed
opened 2026-02-21 19:49:08 -05:00 by yindo · 0 comments
Owner

Originally created by @goofy-z on GitHub (Sep 29, 2025).

Self Checks

  • I have read the Contributing Guide and Language Policy.
  • I have searched for existing issues search for existing issues, including closed ones.
  • I confirm that I am using English to submit this report, otherwise it will be closed.
  • Please do not modify this template :) and fill in all the required fields.

1. Is this request related to a challenge you're experiencing? Tell me about your story.

In our use case, we have a feature for non-streaming LLM invocations. After the recent structured_output refactoring, the LLM node no longer returns the model's structured output results. We believe there could be future non-streaming scenarios that would require structured outputs. For example, when using Dify's external blocking APIs, the current design relies on vendors' streaming APIs even for blocking-mode calls that utilize an LLM node. For tasks that are not latency-sensitive, using streaming APIs for LLM calls primarily introduces increased difficulty in debugging and troubleshooting.

2. Additional context or comments

No response

3. Can you help us with this feature?

  • I am interested in contributing to this feature.
Originally created by @goofy-z on GitHub (Sep 29, 2025). ### Self Checks - [x] I have read the [Contributing Guide](https://github.com/langgenius/dify/blob/main/CONTRIBUTING.md) and [Language Policy](https://github.com/langgenius/dify/issues/1542). - [x] I have searched for existing issues [search for existing issues](https://github.com/langgenius/dify/issues), including closed ones. - [x] I confirm that I am using English to submit this report, otherwise it will be closed. - [x] Please do not modify this template :) and fill in all the required fields. ### 1. Is this request related to a challenge you're experiencing? Tell me about your story. In our use case, we have a feature for non-streaming LLM invocations. After the recent structured_output refactoring, the LLM node no longer returns the model's structured output results. We believe there could be future non-streaming scenarios that would require structured outputs. For example, when using Dify's external blocking APIs, the current design relies on vendors' streaming APIs even for blocking-mode calls that utilize an LLM node. For tasks that are not latency-sensitive, using streaming APIs for LLM calls primarily introduces increased difficulty in debugging and troubleshooting. ### 2. Additional context or comments _No response_ ### 3. Can you help us with this feature? - [x] I am interested in contributing to this feature.
yindo added the 💪 enhancement label 2026-02-21 19:49:08 -05:00
yindo closed this issue 2026-02-21 19:49:08 -05:00
Sign in to join this conversation.
1 Participants
Notifications
Due Date
No due date set.
Dependencies

No dependencies set.

Reference: langgenius/dify#18558