Maybe a noob question, but why aren't there safetensor quants of inclusionAI/Ling-3.0-flash-Fin?
This story is from 2026-09-06. It is preserved in the archive; the latest stories are on the live feed.
Usually, everyone and their dog jumps on releasing different quants for new models, but when I check for inclusionAI/Ling-3.0-flash-Fin , I see quants only for llama.cpp . So I'm just wondering, is it architectural?
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-06 18:26 · r/LocalLLaMA
Maybe a noob question, but why aren't there safetensor quants of inclusionAI/Ling-3.0-flash-Fin?