Open-Source Community Releases Multiple Uncensored AI Models in GGUF Format
This story is from 2026-08-30. It is preserved in the archive; the latest stories are on the live feed.
A developer released several uncensored AI models including LongCat-Flash-Lite-Sparse, Qwen3.8-27B, Qwen3.5-122B-A10B, Qwen3-Coder-Next, and Laguna-S2.1, all in GGUF format, with custom llama.cpp forks for support.
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-08-30 14:16 · r/LocalLLaMA
Uncensored Multi-Model Releases, LongCat-Flash-Lite-Sparse with MTPs and LSAs, Qwen3.8-27B with MTPs, Qwen3.5-122B-A10B with MTPs, Qwen3-Coder-Next and Laguna-S2.1 with Vision, All Available in GGUF Format! Bonus: Links to my llama.cpp Fork for LongCat-Flash-Lite Support and J-Wash Enhanced Fork!
More stories
- Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
- Meta Launches Muse Mac App With File, Messages, and Calendar Access — Unite.AI
- M2 Mac ultra128gb Qwen flash next — r/LocalLLM
- Qwen3.8-Flash-Next-Heretic2-IQ4XS on Halogen Flash Server vs llama-server on Strix Halo: 2.3-7.7x prefill speedup with half the VRAM (+ vision works on BYO GGUF) — r/LocalLLM
- A New AI Model Has Emerged — OPTES AI. — r/GeminiAI
- Multi-hour llama.cpp optimization experiments on Qwen MoE models, patches, benchmarks, and reproduction guides — r/LocalLLM
- Pacing the AI frontier, IBM Granite 4.2 & Meta’s Muse assistant — Mixture of Experts (IBM)
- Meta’s Muse TV ad is the latest sign that AI agents are going mainstream — Business Insider AI
Get the daily brief of stories like this at 6:30 every morning →