mirror of
https://github.com/mudler/LocalAI.git
synced 2026-09-28 00:55:06 -04:00
Add Q4_K_M, Q6_K, and Q8_0 text builds for llama.cpp with pinned shards. Use the embedded chat template and document installation and sizing. Assisted-by: Codex:gpt-6