mirror of
https://github.com/mudler/LocalAI.git
synced 2026-09-20 05:07:07 -04:00
Add Q4_K_M, Q6_K, and Q8_0 builds with embedded Jinja templates. The pinned llama.cpp revision supports the Spark-X2.5 architecture. Document installation and use a 32K context default to limit memory. Assisted-by: Codex:gpt-6