Files
LocalAI/backend/python/vllm
lei_lei a760a7ab4b fix(backends): honor enable_thinking=false in sglang and vllm (#11715)
Those backends only forwarded the flag when it was "true", so "false"
never reached apply_chat_template and Qwen3 kept thinking on.

Signed-off-by: lei_lei <96427312+leilei3167@users.noreply.github.com>
2026-08-25 12:52:57 +02:00
..
2025-08-22 08:42:29 +02:00
2025-06-15 14:56:52 +02:00

Creating a separate environment for the vllm project

make vllm