The llama.cpp server now exposes a new /v1/systemone API endpoint, adding support for the Laya, Julia-1, Lev, OpenJev, and Kev models in GGUF format. Hugging Face links for each model are provided, along with a reference to an additional Bespoke-Nimble-9B-v3-GGUF model and pull request #29844.
Read original
reddit/r/LocalLLaMA