AINewsnow

llama, server: add /v1/systemone API (models: laya, julia-1, lev, openjev, kev) by ngxson · Pull Request #29818 · ggml-org/llama.cpp

article: https://huggingface.co/blog/ggml-org/decision-models-in-llamacpp Now you can jev without jev https://huggingface.co/ggml-org/OpenJev-GGUF https://huggingface.co/ggml-org/lev-GGUF https://huggingface.co/ggml-org/Julia-1-GGUF https://huggingface.co/ggml-org/Laya-GGUF https://huggingface.co/g…

Read the full story at r/LocalLLaMA ↗

Timeline · 3 reports

  1. 2026-10-03 06:18 · r/LocalLLaMA
    qwen4exp : halve the indexer score memory by ServeurpersoCom · Pull Request #29825 · ggml-org/llama.cpp
  2. 2026-10-02 18:47 · r/LocalLLaMA
    CUDA: fuse shared experts into MMVQ by am17an · Pull Request #29184 · ggml-org/llama.cpp
  3. 2026-10-02 10:29 · r/LocalLLaMA
    llama, server: add /v1/systemone API (models: laya, julia-1, lev, openjev, kev) by ngxson · Pull Request #29818 · ggml-org/llama.cpp

More stories

  1. Pi extension: Skip reasoning with local Qwen 27B and proceed to answer right now — r/LocalLLaMA
  2. Claude and Grok built me a local monitoring setup for my AI box: two dashboards, one for the machine, one for model training — r/LocalLLM
  3. I benchmarked Jev 1.13 against 4 local LLMs on RTX5070 12GB - amazing. — r/singularity
  4. Who’s the current “king” of local LLMs for you — Qwen, Gemma, Llama, something else? — r/LocalLLM
  5. Which LLM is best for coding/agents if I have dual R9700 GPUs? — r/LocalLLM
  6. microsoft/FrogNano-4B-2609 · Hugging Face — r/LocalLLaMA
  7. Viggle/Qwen-Image-2.1-viggle-turbo · Hugging Face — r/StableDiffusion
  8. Jev for beginners: how to use it and what to build — How I AI

Get the daily brief of stories like this at 6:30 every morning →