AINewsnow

More stories

  1. We just open-sourced the world's fastest WebGPU kernels for local AI on Hugging Face — r/LocalLLaMA
  2. Benchmarks: Best engine for Qwen 3.8-Flash-Next on Strix Halo — r/LocalLLM
  3. Perplexity Releases pplx-embed-v2-context-9b-preview: A Contextual Embedding Model That Retrieves Answers and Their Supporting Evidence — MarkTechPost
  4. OpenAI Gets Sued Over the Hugging Face Hack — Wired AI
  5. add GLM-5.3-Flash (GLM5-Next) support by timkhronos · Pull Request #27773 · ggml-org/llama.cpp — r/LocalLLaMA
  6. Nova v3: 148M model, 1% of the data, same benchmark scores as SmolLM2-135M — r/huggingface
  7. Minimax H3 new comfy Model — r/StableDiffusion
  8. Update on my free open-source local image app: LoRA support (up to 4 stacked) and Krea 2 are in — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →