mirror of
https://github.com/mudler/LocalAI.git
synced 2026-10-05 04:24:39 -04:00
* ⬆️ Update ggml-org/llama.cpp Signed-off-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com> * fix(llama-cpp): migrate score batches to the new API The pinned engine removes common_batch_add and the raw batch view. Use common_batch entries and llama_process for score suffix decoding. Read shared-prefix scores from the current common_batch view. Validation: reproduce both compiler errors on the original patch. The patched server context and complete grpc-server translation unit pass g++ -std=c++17 -fsyntax-only with generated protobuf headers. Assisted-by: Codex:gpt-6 --------- Signed-off-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com> Co-authored-by: mudler <2420543+mudler@users.noreply.github.com> Co-authored-by: localai-org-maint-bot <306269227+localai-org-maint-bot@users.noreply.github.com>