mirror of
https://github.com/mudler/LocalAI
synced 2026-05-24 09:28:23 +00:00
Use pb.Reply instead of []byte with Reply.GetMessage() in llama grpc to get the proper usage data in reply streaming mode at the last [DONE] frame Co-authored-by: Ettore Di Giacinto <mudler@users.noreply.github.com> |
||
|---|---|---|
| .. | ||
| elevenlabs | ||
| explorer | ||
| jina | ||
| localai | ||
| openai | ||