AINewsnow

More stories

  1. Qwen 3.8 Flash Next - doubled Strata throughput on 3090+5070 Ti, IQ3_S 2466 pp/167 tps, UD-Q4_K_XL 2341 pp / 126 tps (yes, really) — r/LocalLLM
  2. We just open-sourced the world's fastest WebGPU kernels for local AI on Hugging Face — r/LocalLLaMA
  3. add GLM-5.3-Flash (GLM5-Next) support (#27773) · ggml-org/llama.cpp@649dcb1 — r/LocalLLaMA
  4. Anyone tried Swift 1.5 Flash Next GSQ-RCO IQ2_XS + Strata? — r/LocalLLM
  5. Bytedance release 4-step for Minimax-h3; DMAD: Distribution Matching as Adversarial Distillation — r/StableDiffusion
  6. Update on my free open-source local image app: LoRA support (up to 4 stacked) and Krea 2 are in — r/StableDiffusion
  7. California attorney general subpoenas OpenAI over cyber incidents — The Hill Technology
  8. Perplexity Releases pplx-embed-v2-context-9b-preview: A Contextual Embedding Model That Retrieves Answers and Their Supporting Evidence — r/machinelearningnews

Get the daily brief of stories like this at 6:30 every morning →