mirror of
https://github.com/mudler/LocalAI.git
synced 2026-07-31 02:18:50 -04:00
fix(ci): build Bonsai backend images Register the Bonsai C++ source path with the backend matrix filter so changes select its image jobs. Also make shared llama.cpp changes rebuild the Bonsai and Turboquant fork images in the actual matrix, not only their test flags.\n\nAssisted-by: Codex:gpt-5 [Codex] Co-authored-by: localai-org-maint-bot <306269227+localai-org-maint-bot@users.noreply.github.com>
20 lines
1.1 KiB
Markdown
20 lines
1.1 KiB
Markdown
# bonsai fork skew patches
|
|
|
|
The `bonsai` backend reuses `backend/cpp/llama-cpp/grpc-server.cpp` (written against
|
|
LocalAI's pinned *upstream* llama.cpp) but compiles it against the PrismML `prism` fork,
|
|
which branched from upstream some commits earlier. Any upstream API change that the shared
|
|
gRPC server depends on, but that the fork does not yet carry, is back-ported here as a
|
|
`*.patch` file and applied to the cloned fork checkout by `../apply-patches.sh`.
|
|
|
|
CI treats both this directory and `backend/cpp/llama-cpp/` as Bonsai inputs, since
|
|
the wrapper copies and builds the shared llama.cpp backend sources.
|
|
|
|
Rules:
|
|
|
|
- One upstream commit (or minimal hunk) per patch, named `NNNN-short-description.patch`.
|
|
- Patches are applied with `git apply` from the fork's checkout root.
|
|
- `apply-patches.sh` fails fast if a patch stops applying cleanly — that is the signal the
|
|
fork has caught up (or diverged), so re-cut or drop the patch.
|
|
- Keep this set as small as possible; the long-term fix is the fork rebasing onto a newer
|
|
upstream (or Q1_0/Q2_0 landing in mainline llama.cpp, retiring this backend entirely).
|