mirror of
https://github.com/mudler/LocalAI.git
synced 2026-09-13 06:45:26 -04:00
475dc254be9bfdcd20986cb4e333dc0d60815421
9
Commits
| Author | SHA1 | Message | Date | |
|---|---|---|---|---|
|
|
3684a534bb |
docs(website): simplify installation paths (#11631)
Keep the homepage focused on runtime capabilities and move engine details to their canonical directory. Make installation choices stable and explicit for users across supported hardware. Assisted-by: Codex:gpt-5 Co-authored-by: localai-org-bot <306113404+localai-org-bot@users.noreply.github.com> |
||
|
|
d2588b9177 |
docs(blog): add the 4.9 release post and its demo clips (#11629)
* docs(blog): add the 4.9 release post and its demo clips The 4.9 cycle changed how you authenticate, how chat handles a history that no longer fits, and where models and backends live in the UI. The release notes list every pull request; this post covers the three changes that alter day-to-day use, and leads with the auth one because it needs an action before upgrading. Two clips are recorded from a real session against a local-ai built from master with the live gallery loaded: model-lifecycle.mp4 walks the unified models and backends pages, import-model.mp4 shows the rebuilt import form. Both follow the clip conventions in .agents/preparing-a-release.md: h264, no audio track, 1000x562, under 30 seconds, and named after the feature so they stay reusable. Assisted-by: Claude Code:claude-opus-5 [Bash] [Playwright] Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * docs(blog): anti-slop pass over the 4.9 post Ran the post through the humanizer and no-ai-slop rules, calibrated against what-landed-in-localai-4-8.md. That post is the one #11324 left unchanged, so it sets the rate for this series. The draft ran denser than it on two constructions: "rather than" at 5.1 per thousand words against 3.5, and "instead of" at 3.1 against 1.6. Both are now at or below the 4.8 rate, 2.7 and 1.5, by rewriting seven of them as plain statements. Also cut: "keeping both cost a mode switch", a ledger metaphor for something that is not money, which is the tell #11324 removed eight times from the APEX post. "A follow-up fixed the thing that made that awkward", an unearned framing plus a colon reveal. "This release adds a different one: compress them", a second colon reveal. And "byte-structurally identical", a second exactness idiom in a post that already uses "byte-identical" where the precision carries weight. Five paragraphs opened with "Two things" or "Two details", so three of them start differently now. The summary listed three items, which is the rule of three; it lists four, like the 4.8 summary. Every figure, PR number, link and media reference is unchanged, checked by diffing them out of both revisions. Hugo builds clean and the rendered HTML has no em dashes. Assisted-by: Claude Code:claude-opus-5 [Bash] Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Co-authored-by: Ettore Di Giacinto <mudler@localai.io> |
||
|
|
2c0e7c584d |
website: re-record the hero and gallery clips for the 4.8 UI
The two landing-page clips predated the v4.8.0 interface work (#11288, #11305, #11307): the gallery clip showed the retired light-theme Install Models table, and the hero clip toured the Nodes pages in a full browser window while its caption promised a chat completion on CPU. Both are re-recorded from a real local-ai built from v4.8.0, dark theme, app chrome only: - hero-ui.mp4: a chat completion on lfm2.5-1.2b-instruct streaming on CPU with the live tok/s meter, so the caption now matches the footage. The poster frame is regenerated from the new clip. - gallery.mp4: the Discover rail and detail pane, the hardware recommendation lanes, the VRAM-by-context chart, and a real install with the live progress banner. The hand-typed model count moves from 1,585 to 1,255 in the three places it appears, matching the distinct-model count the recorded UI shows on screen. The 3d-generation clip is untouched: the post-capture UI changes do not show in its footage. Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Assisted-by: Claude Code:claude-fable-5 |
||
|
|
c61b6f2286 |
docs(blog): new DeepSeek and Laguna numbers, visuals, humanizer pass (#11369)
* docs(blog): new DeepSeek and Laguna numbers, visuals, humanizer pass vllm.cpp master moved 26 commits past what the post was written against, and two results changed enough to matter. Both came from the same lever: staging weights device-resident at load instead of reading them from the GGUF mmap over unified memory, which the GB10 reads about 20% slower per GEMV than device memory. - DeepSeek-V4-Flash against DwarfStar: 0.997x parity becomes 1.144x ahead, 18.69 vs 16.33 tok/s decode, same generated tokens. - Laguna-XS-2.1 against vLLM: 87% becomes 1.03x, 44.46 vs 43.10 tok/s. New row in the scoreboard. Adds three visuals. A chart of throughput against every reference engine, which is worth having now that the spread is 0.976 to 1.144 rather than a flat line at parity. The Activity page with four installs running, and the model detail pane with all four pocket-35b variants. Both screenshots were recaptured on 2026-08-04 because #11288, #11305, #11307 and #11222 had all changed those pages since the earlier set. llama.cpp is deliberately absent from the chart: its 1.18x is a prefill ratio, and putting it on the same axis as throughput ratios would be comparing two different measurements. Also carries the media the release notes embed, since a GitHub release body needs URLs that survive publishing and drag-and-drop has no CLI. Supersedes #11364. Humanizer pass on the prose. The post had collected five exactness idioms in one section (token-for-token, byte-exact twice, byte-identical, token-identical). One is precision, five is a tic, so the 27B row keeps its "token-for-token identical" where identical output is the actual claim and the rest say what they mean. That also fixed a hyphen in predicate position ("is token-identical"). Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Assisted-by: Claude Code:claude-opus-5 [Read] [Edit] [Bash] * docs(blog): redraw the benchmark chart as a branded card The Flint bar chart was generic: default palette, no brand, and drawn from zero, which made five ratios between 0.976 and 1.144 look like five bars of roughly equal length. Redrawn in the style of recorder-for-agents' render-card.sh cards, the same shape as the vllm.cpp README GIF. Palette taken from the two logos rather than invented (LocalAI navy #0E2632 and teal #469AAF, vllm.cpp teal #3AB4CA), SVG generated by a small JS loop so the geometry is exact at any scale, headless Chrome to PNG at 2x. The substantive change is that bars now run from the 1.00 parity line instead of from zero. Deviation is what the data is about, so DeepSeek's +14.4% and MLX-LM's -2.4% are both legible, and the one row that is behind is the one row in amber. Each bar carries its ratio and the raw measurement under it. Keeps the .html source next to the .png so the chart is editable later: change a number, re-run render-card.sh. Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Assisted-by: Claude Code:claude-opus-5 [Read] [Edit] [Bash] --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Co-authored-by: Ettore Di Giacinto <mudler@localai.io> |
||
|
|
359bd4850d |
docs(blog): bring the 4.8 release post up to the final changelog (#11287)
The post was written against the first draft of the release notes, when the cycle stood at 214 PRs over thirteen days. It closed at 321 PRs over eighteen days, and three of the larger user-facing changes landed after it was written. - Correct the counts throughout: 321 PRs, eighteen days, 24 contributors (11 first-time), gallery 1,221 to 1,505. - Add sections for the three new capabilities: 3D generation as a modality (Generate3D, FLAG_3D, /v1/3d/generations, trellis2cpp), audio.cpp serving six audio endpoints from one process, and the operations bar becoming the Activity page. - Cover the two further hardening fixes (tar hardlink escape, cyclic $ref stack overflow) alongside the TRL one. - Note the Valkey store, systemd socket activation, persistent trace history, in-place chat edits, the self-contained SYCL backend and the site split. - Group the new-engine sections together rather than splitting them across the operational ones. Embeds the existing vllm-race and magpie clips, and adds a 3D generation clip cut from the demo recording to the conventions in .agents/preparing-a-release.md (no audio track, 14s, named for the feature). blog.css styled figure img but not figure video, so a clip in a post rendered outside the card; both selectors now share the rule. Assisted-by: Claude Code:claude-opus-5 [Read] [Edit] [Bash] Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Co-authored-by: Ettore Di Giacinto <mudler@localai.io> |
||
|
|
e6b235baf2 |
fix(website): wrap the timeline, run integrations as a reel, fix blog cards
The timeline set six 15rem columns in a flex row with overflow-x:auto, which needs 90rem and so scrolled sideways on any normal laptop. It is a wrapping grid now, and the rule that carries the dots moves from the container onto each item so a wrapped row still gets a line above it. Column gap is zero and the items carry their own right padding, so the rule stays continuous. Integrations move from a card grid to a reel. Any single integration is a weak signal and the whole moving line is the strong one, so the count is doing the argument. It pauses on hover and on keyboard focus, since the names are links. The list grows from 8 to 26: Open WebUI, Dify, LibreChat, RAGFlow, Continue, big-AGI, Nextcloud, Frigate, promptfoo, Mods, TypingMind, baibot, k8sgpt- operator and others. Each was admitted only after opening that project's own repository or docs and reading the line that names LocalAI. The ones that failed that test are listed in the data file so nobody re-adds them. The blog cards were hand-written, which is how one of them came to advertise "Porting vLLM to C++", a post that does not exist, and how all three linked to the blog index instead of an article. They range over the posts now. The section intro used the "a changelog tells you what moved, these posts show you what it does" shape, which is the standard machine-written antithesis. It states what the posts contain instead, including the perplexity regression that APEX costs, because publishing the price is the actual claim. Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Assisted-by: Claude Code:claude-opus-5[1m] |
||
|
|
0bedc75921 |
fix(website): rewrite the ecosystem band, drop two false coverage links
The band led with three sentences of hedging and printed a commit count next to each employer, so a one-commit entry beside a large name read as weakness rather than as the modest, true claim it was. It now opens on the contributor count, sets the employers as a sentence instead of a pill wall, and keeps the caveat to one line. The counts stay in ecosystem.yaml, since they are the provenance for the list and anyone re-checking it needs them. Two "coverage" cards were not about this project. The modelslab.com piece reviews Frikallo/parakeet.cpp, an unrelated project of the same name, and the snailtext.app benchmark measures Parakeet through ONNX Runtime without mentioning LocalAI at all. Both are removed, along with the contributor card that duplicated the band's opening line. Press was four posts from one vendor, which read as the whole of the coverage rather than one enthusiastic outlet. SUSE collapses to a single series entry, and Pulumi, Semaphore and Spectro Cloud join it. Each was opened and checked against the project before being added. K8sGPT and LlamaIndex join the integrations; both document LocalAI as a backend. The quotes move above the lists so the section opens on its strongest line, which is somebody else's. The hero gains a GitHub call to action, the APEX collection link was returning 404 and is corrected, and the footer no longer describes the site as a design mock. Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Assisted-by: Claude Code:claude-opus-5[1m] |
||
|
|
daab94134c |
feat(website): add an ecosystem band, and an ADOPTERS file to back it (#11248)
Adds the "who turns up around this project" section, split into three lists
because the evidence behind each one is a different strength and collapsing
them into a single logo wall would overclaim.
Contributors 21 companies whose engineers have commits here. Evidence is
the commit history plus the employer on that person's public
GitHub profile, so it is a claim about the person. Commit
counts are shown next to each name, including the ones that
are a single patch, because hiding that would be the whole
problem.
Integrations six projects that reference LocalAI in their own repository
or documentation, which anyone can verify without asking us.
Press four SUSE Communities articles about running LocalAI.
Names are set in type rather than fetched as logos. A logo reads as
endorsement, and a one-line typo fix from somebody who happens to work at a
large company does not support that, quite apart from what their trademark
policy says about it.
ADOPTERS.md is the mechanism for the stronger claim. An organisation that
wants to be listed as a user opens a pull request adding itself, which is both
the evidence and the permission, and is publicly auditable afterwards. The
file says plainly what the website does and does not claim, so the next person
to ask "can we add some big names" has the answer in the repository.
Assisted-by: Claude Code:claude-opus-5 [Bash] [Edit] [Write] [WebSearch]
Signed-off-by: Ettore Di Giacinto <mudler@localai.io>
Co-authored-by: Ettore Di Giacinto <mudler@localai.io>
|
||
|
|
94d5affcea |
feat(website): split the site, move docs to /docs, add a landing page (#11243)
* feat(website): split the site, move docs to /docs, add a landing page The Hugo docs site has always been localai.io itself, which left nowhere to explain what LocalAI is or show what the team builds. This adds a separate marketing site at the root and moves the documentation under /docs/. Docs: The existing site keeps its content tree and its Relearn theme, and now builds with baseURL <root>/docs/. Its _index.md, which held a hand written landing page, becomes a real documentation home. Every previously published URL keeps working. GitHub Pages has no server side rewrites, so .github/ci/gen-redirects.sh walks the built docs output and leaves a meta refresh plus a canonical link at each old root path. It covers bare .html files too, which is what keeps /gallery.html alive, and it never overwrites a path the marketing site already owns. Website: A second Hugo site under website/ with its own layouts and no external theme, so the marketing side does not have to fight Relearn's home rooted menu and asset pipeline. CI builds both and merges them into one Pages artifact. The design is derived from the project logo rather than invented: the navy of the triangle, the cyan of the llama, the purple of the speed bars. Those offset bars became the motion signature. The background renders a real depth-anything.cpp depth map as contour lines and switches to a locate-anything.cpp style detection overlay over the engines section. Also included: an /engines/ index driven entirely by data/engines.yaml, a /blog/ section with five posts written from the release notes and the engine benchmark suites, install.sh and a Kubernetes manifest since the site advertises both, and a rule in .agents/ that release preparation now includes a blog post and demo clips. Every figure on the site is derived from the repository or the GitHub API, not from memory. Correcting them against their sources found one error in README.md: voxtral-tts.c is text to speech, not speech to text. Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Assisted-by: Claude Code:claude-opus-5 [Bash] [Edit] [Write] [Agent] * feat(website): add a star history chart, rewrite the history post in first person The history post read like a changelog written by a committee. It is now in Ettore's voice, first person, with the admissions left in. The numbers paragraph in particular read like a directory listing. It now says what the figures mean rather than which file they came from. Adds an interactive star history chart, built from the GitHub stargazers API rather than embedded from a third party, so the page makes no external request and cannot break when someone else's service is down. The four releases the post is organised around are marked on the curve, and the labels stack into rows because three of them land within two months of each other. The API stops paginating at 40,000 items, so the curve is measured up to December 2025 and the segment from there to today's total is drawn dashed, labelled as an estimate in the caption and in the tooltip. It is a straight line between two known points, and the chart says so rather than implying it is data. Also drops "marketing site" from the README heading and everywhere else it appeared, and calls it the main site instead. Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Assisted-by: Claude Code:claude-opus-5 [Bash] [Edit] [Write] --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Co-authored-by: Ettore Di Giacinto <mudler@localai.io> |