LocalAI

mirror of https://github.com/mudler/LocalAI.git synced 2026-01-19 03:40:46 -05:00

Author	SHA1	Message	Date
Ettore Di Giacinto	21c84f432f	feat(function): Add tool streaming, XML Tool Call Parsing Support (#7865 ) * feat(function): Add XML Tool Call Parsing Support Extend the function parsing system in LocalAI to support XML-style tool calls, similar to how JSON tool calls are currently parsed. This will allow models that return XML format (like <tool_call><function=name><parameter=key>value</parameter></function></tool_call>) to be properly parsed alongside text content. Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * thinking before tool calls, more strict support for corner cases with no tools Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Support streaming tools Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Iterative JSON Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Iterative parsing Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Consume JSON marker Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Fixup Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * add tests Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Fix pending TODOs Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Don't run other parsing with ParseRegex Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2026-01-05 18:25:40 +01:00
Ettore Di Giacinto	33cc0b8e13	fix(chat/ui): record model name in history for consistency (#7845 ) Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2026-01-03 18:05:33 +01:00
lif	4cd95b8a9d	fix: Highly inconsistent agent response to cogito agent calling MCP server - Body "Invalid http method" (#7790 ) * fix: resolve duplicate MCP route registration causing 50% failure rate Fixes #7772 The issue was caused by duplicate registration of the MCP endpoint /mcp/v1/chat/completions in both openai.go and localai.go, leading to a race condition where requests would randomly hit different handlers with incompatible behaviors. Changes: - Removed duplicate MCP route registration from openai.go - Kept the localai.MCPStreamEndpoint as the canonical handler - Added all three MCP route patterns for backward compatibility: * /v1/mcp/chat/completions * /mcp/v1/chat/completions * /mcp/chat/completions - Added comments to clarify route ownership and prevent future conflicts - Fixed formatting in ui_api.go The localai.MCPStreamEndpoint handler is more feature-complete as it supports both streaming and non-streaming modes, while the removed openai.MCPCompletionEndpoint only supported synchronous requests. This eliminates the ~50% failure rate where the cogito library would receive "Invalid http method" errors when internal HTTP requests were routed to the wrong handler. 🤖 Generated with [Claude Code](https://claude.com/claude-code) Co-Authored-By: Claude Sonnet 4.5 <noreply@anthropic.com> Signed-off-by: majiayu000 <1835304752@qq.com> * Address feedback from review Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: majiayu000 <1835304752@qq.com> Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Co-authored-by: Claude Sonnet 4.5 <noreply@anthropic.com> Co-authored-by: Ettore Di Giacinto <mudler@localai.io>	2026-01-03 15:43:23 +01:00
Ettore Di Giacinto	5f6c941399	fix(llama.cpp/mmproj): fix loading mmproj in nested sub-dirs different from model path (#7832 ) fix(mmproj): fix loading mmproj in nested sub-dirs Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2026-01-02 20:17:30 +01:00
Ettore Di Giacinto	841e8f6d47	fix(image-gen): fix scrolling issues (#7829 ) Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2026-01-02 09:05:49 +01:00
Ettore Di Giacinto	76cfe1f367	feat(image-gen/UI): move controls to the left, make the page more compact (#7823 ) Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2026-01-01 22:07:42 +01:00
Ettore Di Giacinto	797f27f09f	feat(UI): image generation improvements (#7804 ) * chore: drop mode from image generation(unused) Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * feat(UI): improve image generation front-end Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * feat(UI): only ref images. files is to be deprecated Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * do not override default steps Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-31 21:59:46 +01:00
lif	8bd7143a44	fix: propagate validation errors (#7787 ) fix: validate MCP configuration in model config Fixes #7334 The Validate() function was not checking if MCP configuration (mcp.stdio and mcp.remote) contains valid JSON. This caused malformed JSON with missing commas to be silently accepted. Changes: - Add MCP configuration validation to ModelConfig.Validate() - Properly report validation errors instead of discarding them - Add test cases for valid and invalid MCP configurations The fix ensures that malformed JSON in MCP config sections will now be caught and reported during validation. 🤖 Generated with [Claude Code](https://claude.com/claude-code) Signed-off-by: majiayu000 <1835304752@qq.com> Co-authored-by: Claude Sonnet 4.5 <noreply@anthropic.com>	2025-12-30 09:54:27 +01:00
lif	0d0ef0121c	fix: Usage for image generation is incorrect (and causes error in LiteLLM) (#7786 ) * fix: Add usage fields to image generation response for OpenAI API compatibility Fixes #7354 Added input_tokens, output_tokens, and input_tokens_details fields to the image generation API response to comply with OpenAI's image generation API specification. This resolves validation errors in LiteLLM and the OpenAI SDK. Changes: - Added InputTokensDetails struct with text_tokens and image_tokens fields - Extended OpenAIUsage struct with input_tokens, output_tokens, and input_tokens_details - Updated ImageEndpoint to populate usage object with required fields - Updated InpaintingEndpoint to populate usage object with required fields - All fields initialized to 0 as per current behavior 🤖 Generated with [Claude Code](https://claude.com/claude-code) Co-Authored-By: Claude Sonnet 4.5 <noreply@anthropic.com> Signed-off-by: majiayu000 <1835304752@qq.com> * fix: Correct usage field types for image generation API compatibility Changed InputTokens and OutputTokens from pointer types (*int) to regular int types to match OpenAI API specification. This fixes validation errors with LiteLLM and OpenAI SDK when parsing image generation responses. Fixes #7354 🤖 Generated with [Claude Code](https://claude.com/claude-code) Co-Authored-By: Claude Sonnet 4.5 <noreply@anthropic.com> Signed-off-by: majiayu000 <1835304752@qq.com> --------- Signed-off-by: majiayu000 <1835304752@qq.com> Co-authored-by: Claude Sonnet 4.5 <noreply@anthropic.com>	2025-12-30 09:53:05 +01:00
lif	d7b2eee08f	fix: add nil checks before mergo.Merge to prevent panic in gallery model installation (#7785 ) Fixes #7420 Added nil checks before calling mergo.Merge in InstallModelFromGallery and InstallModel functions to prevent panic when req.Overrides or configOverrides are nil. The panic was occurring at models.go:248 during Qwen-Image-Edit gallery model download. Changes: - Added nil check for req.Overrides before merging in InstallModelFromGallery (line 126) - Added nil check for configOverrides before merging in InstallModel (line 248) - Added test case to verify nil configOverrides are handled without panic Signed-off-by: majiayu000 <1835304752@qq.com>	2025-12-30 09:51:45 +01:00
Richard Palethorpe	99b5c5f156	feat(api): Allow tracing of requests and responses (#7609 ) * feat(api): Allow tracing of requests and responses Signed-off-by: Richard Palethorpe <io@richiejp.com> * feat(traces): Add traces UI Signed-off-by: Richard Palethorpe <io@richiejp.com> --------- Signed-off-by: Richard Palethorpe <io@richiejp.com>	2025-12-29 11:06:06 +01:00
Ettore Di Giacinto	21c464c34f	fix(cli): import via CLI needs system state (#7746 ) pass system state to application config to avoid nil pointer exception during import. Fixes: https://github.com/mudler/LocalAI/issues/7728 Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-27 11:10:28 +01:00
Ettore Di Giacinto	c844b7ac58	feat: disable force eviction (#7725 ) * feat: allow to set forcing backends eviction while requests are in flight Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * feat: try to make the request sit and retry if eviction couldn't be done Otherwise calls that in order to pass would need to shutdown other backends would just fail. In this way instead we make the request sit and retry eviction until it succeeds. The thresholds can be configured by the user. Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * add tests Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * expose settings to CLI Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Update docs Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-25 14:26:18 +01:00
Ettore Di Giacinto	bb459e671f	fix(ui): correctly parse import errors (#7726 ) errors are nested Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-25 10:43:12 +01:00
Ettore Di Giacinto	35d71cf25e	fix: remove duplicate logging line Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-24 09:35:18 +01:00
Ettore Di Giacinto	83ed16f325	chore(logging): be consistent and do not emit logs from echo (#7710 ) Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-24 09:22:27 +01:00
Ettore Di Giacinto	8b3e0ebf8a	chore: allow to set local-ai log format, default to custom one (#7679 ) Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-21 21:21:59 +01:00
Ettore Di Giacinto	c37785b78c	chore(refactor): move logging to common package based on slog (#7668 ) Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-21 19:33:13 +01:00
Ettore Di Giacinto	8b6f443cd5	chore(deps): bump cogito to latest and adapt API changes (#7655 ) Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-19 22:50:18 +01:00
Richard Palethorpe	716dba94b4	feat(whisper): Add prompt to condition transcription output (#7624 ) * chore(makefile): Add buildargs for sd and cuda when building backend Signed-off-by: Richard Palethorpe <io@richiejp.com> * feat(whisper): Add prompt to condition transcription output Signed-off-by: Richard Palethorpe <io@richiejp.com> --------- Signed-off-by: Richard Palethorpe <io@richiejp.com>	2025-12-18 14:40:45 +01:00
Ettore Di Giacinto	d8ee02e607	chore(tests): simplify tests and run intensive ones only once Signed-off-by: Ettore Di Giacinto <mudler@users.noreply.github.com>	2025-12-18 09:05:58 +01:00
Ettore Di Giacinto	f3c70a96ba	chore(memory-reclaimer): use saner defaults Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-16 16:25:09 +01:00
Ettore Di Giacinto	878c9d46d5	fix: improve ram estimation (#7603 ) * fix: default to 10seconds of watchdog if runtime setting is malformed Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * fix: use gosigar for RAM estimation Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-16 10:18:36 +01:00
Ettore Di Giacinto	50f9c9a058	feat(watchdog): add Memory resource reclaimer (#7583 ) * feat(watchdog): add GPU reclaimer Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Handle vram calculation for unified memory devices Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Support RAM eviction, set watchdog interval from runtime settings Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-16 09:15:18 +01:00
Ettore Di Giacinto	8ac7e8c299	fix(chat-ui): model selection toggle and new chat (#7574 ) Fixes a minor glitch that happens when switching model in from the chat pane where the header was not getting updated. Besides, it allows to create new chat directly when clicking from the management pane to the model. Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-14 22:29:11 +01:00
Ettore Di Giacinto	e1874cdb54	feat(ui): add mask to install custom backends (#7559 ) * feat: allow to install backends from URL in the WebUI and API Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * tests Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * trace backends installations Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-13 19:11:32 +01:00
Ettore Di Giacinto	fc5b9ebfcc	feat(loader): enhance single active backend to support LRU eviction (#7535 ) * feat(loader): refactor single active backend support to LRU This changeset introduces LRU management of loaded backends. Users can set now a maximum number of models to be loaded concurrently, and, when setting LocalAI in single active backend mode we set LRU to 1 for backward compatibility. Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * chore: add tests Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Update docs Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Fixups Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-12 12:28:38 +01:00
Ettore Di Giacinto	3b5c2ea633	feat(ui): allow to order search results (#7507 ) * feat(ui): improve table view and let items to be sorted Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * refactorings Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * chore: add tests Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * chore: use constants Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-11 00:11:33 +01:00
Ettore Di Giacinto	f51d3e380b	fix(config): make syncKnownUsecasesFromString idempotent (#7493 ) fix(config): correctly parse usecases from strings Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-09 21:08:22 +01:00
Ettore Di Giacinto	6cc5cac7b0	fix(downloader): do not download model files if not necessary (#7492 ) Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-09 19:08:10 +01:00
Ettore Di Giacinto	03a17a2986	fix(paths): remove trailing slash from requests (#7451 ) This removes any ambiguity from how paths are handled, and at the same time it uniforms the ui paths with the other paths that don't have a trailing slash Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-07 21:45:09 +01:00
Ettore Di Giacinto	8ca98c90ea	chore(importers/llama.cpp): add models to 'llama-cpp' subfolder (#7450 ) This makes paths predictable, and avoids multiple model files to show in the main view Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-07 21:44:57 +01:00
rampa3	6aee29d18f	fix(ui): Update few links in web UI from 'browse' to '/browse/' (#7445 ) * Update few links in web UI from 'browse' to '/browse/' Signed-off-by: rampa3 <68955305+rampa3@users.noreply.github.com> * Update core/http/views/404.html Signed-off-by: Ettore Di Giacinto <mudler@users.noreply.github.com> * Update core/http/views/error.html Signed-off-by: Ettore Di Giacinto <mudler@users.noreply.github.com> * Update core/http/views/manage.html Signed-off-by: Ettore Di Giacinto <mudler@users.noreply.github.com> --------- Signed-off-by: rampa3 <68955305+rampa3@users.noreply.github.com> Signed-off-by: Ettore Di Giacinto <mudler@users.noreply.github.com> Co-authored-by: Ettore Di Giacinto <mudler@users.noreply.github.com>	2025-12-06 22:40:26 +01:00
Ettore Di Giacinto	92ee8c2256	fix(ui): prevent box overflow in chat view (#7430 ) Otherwise tool call and result might overflow the box Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-04 17:21:17 +01:00
Ettore Di Giacinto	78105e6b20	chore(ui): uniform buttons (#7429 ) Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-04 17:18:51 +01:00
Ettore Di Giacinto	100ebdfa2c	chore(ci): do not overload the apple tests Skip tests that are already run on other jobs and not really adding anything here. We have already functional tests that cover apple. Signed-off-by: Ettore Di Giacinto <mudler@users.noreply.github.com>	2025-12-04 14:15:15 +01:00
Ettore Di Giacinto	045baf7fd2	fix(ui): navbar ordering and login icon (#7407 ) Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-01 21:20:11 +01:00
Ettore Di Giacinto	8a54ffa668	fix: do not require auth for readyz/healthz endpoints (#7403 ) Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-12-01 10:35:28 +01:00
Ettore Di Giacinto	a3423f33e1	feat(agent-jobs): add multimedia support (#7398 ) * feat(agent-jobs): add multimedia support Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Refactoring Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-11-30 14:09:25 +01:00
Ettore Di Giacinto	54b5dfa8e1	chore: refactor css, restyle to be slightly minimalistic (#7397 ) restyle Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-11-29 22:11:44 +01:00
Ettore Di Giacinto	53e5b2d6be	feat: agent jobs panel (#7390 ) * feat(agent): agent jobs Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Multiple webhooks, simplify Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Do not use cron with seconds Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Create separate pages for details Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Detect if no models have MCP configuration, show wizard Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Make services test to run Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-11-28 23:05:39 +01:00
Igor B. Poretsky	ab022172a9	chore: switch from /usr/share to /var/lib for data storage (#7361 ) * More appropriate place for data storing The /usr/share subtree in Linux is used for data that generally are not supposed to change. Conventional places for changeable data are usually located under /var, so /var/lib seems to be a reasonable default here. * Data paths consistency fix * Directory name consistency fix	2025-11-27 09:18:28 +01:00
Gregory Mariani	745c31e013	feat(inpainting): add inpainting endpoint, wire ImageGenerationFunc and return generated image URL (#7328 ) feat(inpainting): add inpainting endpoint with automatic model selection Signed-off-by: Greg <marianigregory@pm.me>	2025-11-24 21:13:54 +01:00
Ettore Di Giacinto	aceebf81d6	chore(ui): fix slider overflow Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-11-24 14:43:38 +01:00
Ettore Di Giacinto	71ed03102f	feat(ui): add chat history (#7325 ) * feat(chat): add history and management Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Display in progress chats Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Fetch available context size as we switch chat Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Add search Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Display MCP toggle correctly Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Re-ordering Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Re-style Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Stable ordering Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Display token/sec correctly Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Visual changes Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Display chat time Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-11-24 11:48:24 +01:00
Ettore Di Giacinto	fa00aa0085	chore(ci): add OS check to skip test if not on Linux Skip test on non-Linux operating systems. Signed-off-by: Ettore Di Giacinto <mudler@users.noreply.github.com>	2025-11-21 15:01:04 +01:00
Ettore Di Giacinto	0e53ce60b4	chore(ci): remove context size configuration from application Signed-off-by: Ettore Di Giacinto <mudler@users.noreply.github.com>	2025-11-21 14:57:32 +01:00
Ettore Di Giacinto	8aba078439	chore(tests): add context size option to application initialization Signed-off-by: Ettore Di Giacinto <mudler@users.noreply.github.com>	2025-11-21 09:50:05 +01:00
Copilot	16e5689162	feat(importers): Add diffuser backend importer with ginkgo tests and UI support (#7316 ) * Initial plan * Add diffuser backend importer with ginkgo tests Co-authored-by: mudler <2420543+mudler@users.noreply.github.com> * Finalize diffuser backend importer implementation Co-authored-by: mudler <2420543+mudler@users.noreply.github.com> * Add diffuser preferences to model-editor import section Co-authored-by: mudler <2420543+mudler@users.noreply.github.com> * Use gopkg.in/yaml.v3 for consistency in diffuser importer Co-authored-by: mudler <2420543+mudler@users.noreply.github.com> --------- Co-authored-by: copilot-swe-agent[bot] <198982749+Copilot@users.noreply.github.com> Co-authored-by: mudler <2420543+mudler@users.noreply.github.com>	2025-11-20 22:38:30 +01:00
Ettore Di Giacinto	2dd42292dc	feat(ui): runtime settings (#7320 ) * feat(ui): add watchdog settings Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Do not re-read env Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Some refactor, move other settings to runtime (p2p) Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Add API Keys handling Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Allow to disable runtime settings Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Documentation Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Small fixups Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * show MCP toggle in index Signed-off-by: Ettore Di Giacinto <mudler@localai.io> * Drop context default Signed-off-by: Ettore Di Giacinto <mudler@localai.io> --------- Signed-off-by: Ettore Di Giacinto <mudler@localai.io>	2025-11-20 22:37:20 +01:00

1 2 3 4 5 ...

472 Commits