[GH-ISSUE #2023] [BUG]: I need help,When requesting/api/v1/openai/chat/completions for anything-llm 1.6.0 desktop version, the max_token does not take effect, and the connected model is a local model(xinference) #1320

Closed
opened 2026-02-22 18:24:15 -05:00 by yindo · 3 comments
Owner

Originally created by @c935289832 on GitHub (Aug 1, 2024).
Original GitHub issue: https://github.com/Mintplex-Labs/anything-llm/issues/2023

How are you running AnythingLLM?

AnythingLLM desktop app

What happened?

When requesting/api/v1/openai/chat/completions for anything-llm 1.6.0 desktop version, the max_token does not take effect, and the connected model is a local model(xinference)

Are there known steps to reproduce?

this is xinference:
image
image
this is anything-llm desktop:
image
1722561266193

Originally created by @c935289832 on GitHub (Aug 1, 2024). Original GitHub issue: https://github.com/Mintplex-Labs/anything-llm/issues/2023 ### How are you running AnythingLLM? AnythingLLM desktop app ### What happened? When requesting/api/v1/openai/chat/completions for anything-llm 1.6.0 desktop version, the max_token does not take effect, and the connected model is a local model(xinference) ### Are there known steps to reproduce? this is xinference: ![image](https://github.com/user-attachments/assets/99a32ca3-94e2-41d7-bca3-4551a30f22b6) ![image](https://github.com/user-attachments/assets/99251c9e-e178-48ec-93c9-f982c9df2d50) this is anything-llm desktop: ![image](https://github.com/user-attachments/assets/8d3929ad-697f-4308-98da-1d7b6e336693) ![1722561266193](https://github.com/user-attachments/assets/df1ab8e8-0c0b-47a1-8d28-94afa54c0705)
yindo added the possible bug label 2026-02-22 18:24:15 -05:00
yindo closed this issue 2026-02-22 18:24:15 -05:00
Author
Owner

@timothycarambat commented on GitHub (Aug 1, 2024):

What LLM provider is this?

@timothycarambat commented on GitHub (Aug 1, 2024): What LLM provider is this?
Author
Owner

@c935289832 commented on GitHub (Aug 4, 2024):

What LLM provider is this?

xinference+qwen2-instruct

@c935289832 commented on GitHub (Aug 4, 2024): > What LLM provider is this? xinference+qwen2-instruct
Author
Owner

@timothycarambat commented on GitHub (Aug 5, 2024):

@c935289832 The provider, (ollama, LMStudio, etc) not the model specifically as that is independent of the model

@timothycarambat commented on GitHub (Aug 5, 2024): @c935289832 The provider, (ollama, LMStudio, etc) not the model specifically as that is independent of the model
yindo changed title from [BUG]: I need help,When requesting/api/v1/openai/chat/completions for anything-llm 1.6.0 desktop version, the max_token does not take effect, and the connected model is a local model(xinference) to [GH-ISSUE #2023] [BUG]: I need help,When requesting/api/v1/openai/chat/completions for anything-llm 1.6.0 desktop version, the max_token does not take effect, and the connected model is a local model(xinference) 2026-06-05 14:40:08 -04:00
Sign in to join this conversation.
1 Participants
Notifications
Due Date
No due date set.
Dependencies

No dependencies set.

Reference: Mintplex-Labs/anything-llm#1320