LocalAI/.github at df86e8d6d4b14518d6cc79948845afe495649101 - LocalAI - Gitea: Git with a cup of tea

mirror/LocalAI

mirror of https://github.com/mudler/LocalAI.git synced 2026-06-07 00:06:51 -04:00

Files

History

Ettore Di Giacinto df86e8d6d4 ci(turboquant): drop the ROCm/hipblas build flavor

The TheTom/llama-cpp-turboquant fork is not ROCm-clean at the current pin:
beyond the CUDA-API gaps already patched (3D-peer copy, cudaEventCreate),
its llama.cpp base fails to compile the flash-attention MMA f16 kernels for
head-dim 640 under HIP (cols_per_warp evaluates to 0 -> division-by-zero /
non-constant static asserts in fattn-mma-f16.cuh). That is a deep
ggml-on-ROCm kernel issue, not something a small fork patch can paper over.

Drop -gpu-rocm-hipblas-turboquant from the build matrix so turboquant still
ships for cpu / cublas / vulkan / sycl. Re-add it once the fork's HIP path
compiles (or upstream ggml fixes the large-head-dim MMA kernels for ROCm).

Signed-off-by: Ettore Di Giacinto <mudler@localai.io>
Assisted-by: Claude:claude-opus-4-8 [Claude Code]

2026-06-06 22:43:47 +00:00

..

ci: phase 1-3 of GHA free tier migration (path filter, multi-arch split prep, /mnt disk relief) (#9726 )

2026-05-08 23:43:41 +02:00

fix: roll out bluemonday Sanitize more widely (#3794 )

2024-10-12 09:45:47 +02:00

Harden gallery-agent Hugging Face fetches against transient rate limiting (#10187 )

2026-06-05 23:43:06 +02:00

docs/examples: enhancements (#1572 )

2024-01-18 19:41:08 +01:00

ci: close GC race + cascade-skip + darwin grpc gaps from v4.2.1 (#9781 )

2026-05-12 17:22:09 +02:00

chore(deps): bump securego/gosec from 2.22.9 to 2.27.1 (#10147 )

2026-06-03 10:38:07 +02:00

backend-matrix.yml

ci(turboquant): drop the ROCm/hipblas build flavor

2026-06-06 22:43:47 +00:00

bump_deps.sh

feat: do not bundle llama-cpp anymore (#5790 )

2025-07-18 13:24:12 +02:00

bump_docs.sh

fix: github bump_docs.sh regex to drop emoji and other text (#2180 )

2024-04-29 03:55:29 +00:00

bump_vllm_wheel.sh

feat(vllm): expose AsyncEngineArgs via generic engine_args YAML map (#9563 )

2026-04-29 00:49:28 +02:00

check_and_update.py

fix(ci): fixup checksum scanning pipeline (#3631 )

2024-09-23 10:56:10 +02:00

checksum_checker.sh

fix(ci): fixup correct path for check_and_update.py (#2777 )

2024-07-11 23:05:43 +02:00

dependabot.yml

feat: Add backend gallery (#5607 )

2025-06-15 14:56:52 +02:00

FUNDING.yml

Create FUNDING.yml (#725 )

2023-07-09 13:39:00 +02:00

labeler.yml

chore(ci): update labels

2025-02-13 09:58:19 +01:00

PULL_REQUEST_TEMPLATE.md

feat(vllm): Allow to set quantization (#1094 )

2023-09-22 15:52:38 +02:00

release.yml

feat(p2p): Federation and AI swarms (#2723 )

2024-07-08 22:04:06 +02:00

stale.yml

feat: add PR template and stale configuration (#316 )

2023-05-20 09:10:20 +02:00