LocalAI

mirror of https://github.com/mudler/LocalAI.git synced 2025-04-23 18:33:49 +00:00

History

* Use pb.Reply instead of []byte with Reply.GetMessage() in llama grpc to get the proper usage data in reply streaming mode at the last [DONE] frame

* Fix 'hang' on empty message from the start

Seems like that empty message marker trick was unnecessary

---------

Co-authored-by: Ettore Di Giacinto <mudler@users.noreply.github.com>

2024-12-18 09:48:50 +01:00

application

feat(template): read jinja templates from gguf files (#4332 )

2024-12-08 13:50:33 +01:00

backend

feat: stream tokens usage (#4415 )

2024-12-18 09:48:50 +01:00

cli

feat(template): read jinja templates from gguf files (#4332 )

2024-12-08 13:50:33 +01:00

clients

feat(store): add Golang client (#1977 )