mirror of
https://github.com/mudler/LocalAI.git
synced 2026-10-01 18:44:35 -04:00
The proxy now forwards TTS, streaming TTS, sound generation, transcription (plain and streaming), diarization, VAD, sound detection and audio transforms to the upstream LocalAI. Streaming TTS passes the upstream WAV bytes through unchanged. A streaming transcription that stops before its final frame, or sends an error frame, fails with Unavailable instead of ending as a short success. Transcription always sends diarize, because the upstream treats a missing field as true. Audio transforms also download the separation stems the upstream names and write them beside Dst. Sound generation from a source clip returns Unimplemented, because the REST endpoint has no field for the clip. The multipart helper now takes repeated fields and several files. Assisted-by: Claude:claude-opus-5-5 Signed-off-by: Ettore Di Giacinto <mudler@localai.io>