I made Ignis: an open-source engine that only runs Qwen3.8-27B on one RTX 5090, and that's the point. 512K context, Very Fast and JEV-live Support for Images and long text/logs retrival
Hi r/LocalLLM , Everything started from Ninfer... But it was not enough. So for the last few weeks I've been building from scratch (except for some kernels) Ignis , an Apache-2.0 inference engine in rust that deliberately gives up generality: one model family (Qwen3.8-27B NVFP4), one class of card…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-04 18:07 · r/LocalLLM
I made Ignis: an open-source engine that only runs Qwen3.8-27B on one RTX 5090, and that's the point. 512K context, Very Fast and JEV-live Support for Images and long text/logs retrival
More stories
- llama, server: add /v1/systemone API (models: laya, julia-1, lev, openjev, kev) by ngxson · Pull Request #29818 · ggml-org/llama.cpp — r/LocalLLaMA
- How do you keep up with AI when new things happen every day? I used JEV to help me in this: (used ChatGPT for structure only) — r/AI_Agents
- New frontier AI models, TypeSafe’s Jev AI, & NASA’s IBM collab — Mixture of Experts (IBM)
- Startup TypeSafe AI’s Jev Model Sparks Copycats, Talk of LLM Alternatives — Wall Street Journal Technology
- I built a PvP on-hold simulator where you race Jev through phone menus — r/learnmachinelearning
- Benchmarking decision models is fun - Clef Q8 vs Jev — r/LocalLLaMA
- MoralityBench.ai: Morality Leaderboard for AI — r/artificial
- Can someone explain how JEV is different from a simple embeddings model? — r/LocalLLaMA
Get the daily brief of stories like this at 6:30 every morning →