LocalAI/Makefile at 1f313cfdb0defbaf1b2fef505941c170f0cf6ce2

mirror of https://github.com/mudler/LocalAI.git synced 2026-05-17 13:10:23 -04:00

Files

Ettore Di Giacinto 3bc5ae8da6 fix(tests/e2e-backends): bump ctx_size for llama-cpp transcription

Qwen3-ASR-0.6B encodes the jfk.wav fixture into 777 audio tokens via
its mmproj, but the test harness defaulted BACKEND_TEST_CTX_SIZE to
512, so llama.cpp server rejected every transcription request with
"request (777 tokens) exceeds the available context size (512 tokens)".

Set BACKEND_TEST_CTX_SIZE=2048 on the llama-cpp transcription target
only — sherpa-onnx and vibevoice transcription targets don't go
through llama.cpp's slot/n_ctx and weren't failing.

Signed-off-by: Ettore Di Giacinto <mudler@localai.io>
Assisted-by: Claude:claude-opus-4-7 [Claude Code]

2026-05-07 22:31:08 +00:00

60 KiB

Raw Blame History

View Raw

60 KiB Raw Blame History

60 KiB

Raw Blame History