AINewsnow

A Jev-style model fine-tuned on Qwen3.5 4B

This weekend, I did a fun experiment to create something similar to Jev. I LoRA fine-tuned Qwen3.5 4B using a mix of publicly available datasets and synthetic data. For the synthetic data, I used DeepSeek V4.1 Flash, around 25M tokens. I trained the model for about 2 hours on a rented RTX 3090. So…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-09-20 16:39 · r/LocalLLaMA
    A Jev-style model fine-tuned on Qwen3.5 4B

More stories

  1. Bolt Adds DeepSeek V4.1 Flash at 10x Cheaper Than V4 Pro — AlphaSignal
  2. DeepSeek’s Insane New Architecture — Two Minute Papers
  3. Jina AI Releases jina-ocr-v1: A 3.4B MoE Document Parser With Built-In Speculative Decoding for Low-Budget GPUs — MarkTechPost
  4. Gemini 4 is good enough - JUST RELEASE IT — r/GeminiAI
  5. Ai used for chatbots — r/artificial
  6. I gave 6 different AIs the same 5 questions — r/AI_Agents
  7. PromptDeck v1.1.0 – open-source desktop app to benchmark local AND cloud LLMs side-by-side (Ollama, LM Studio + OpenRouter, Groq, DeepSeek…) — r/LocalLLaMA
  8. What’s your favorite AI model for coding right now? — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →