AINewsnow

2x Tesla p100s, q6_k quant, Qwen 3.8 27B ~60tps V3.0

Hey guys! I have been excited to share this here. This is a project consisting of kernel optimizations for the Tesla p100 series graphics card ($80). I want to start by saying I am 17 years old and do not have a formal degree. I used Ai for a lot of this and while I understand some, I do not unders…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-09-27 20:42 · r/LocalLLaMA
    2x Tesla p100s, q6_k quant, Qwen 3.8 27B ~60tps V3.0

More stories

  1. Qwen 3.8 27B vs Qwen 3.8 Flash Next and time to complete a coding task. — r/LocalLLaMA
  2. Help me plan a Qwen 3.8 Flash Next install on a 5090 + 64gb DDR5 system — r/LocalLLM
  3. Layer Extract & Layer Remove Loras For Qwen Image 2.1 — r/StableDiffusion
  4. I built Slopus, a free, open-source desktop app for generating and editing AI videos locally (Minimax H3) — r/StableDiffusion
  5. Deepseek V4 Flash 0731 on m5 max 128gb — r/LocalLLM
  6. Community reports say the first samples of Qwen 4 are already approaching Fable / Opus-level quality. — r/singularity
  7. Qwen-Image 2.1 Inpainting with LanPaint — alpha channel included — r/StableDiffusion
  8. A LoRA I made: AnyAngle LoRA for Qwen Image 2.1. Style-Aligned Arbitrary Camera Angles — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →