A Jev-style model fine-tuned on Qwen3.5 4B
This weekend, I did a fun experiment to create something similar to Jev. I LoRA fine-tuned Qwen3.5 4B using a mix of publicly available datasets and synthetic data. For the synthetic data, I used DeepSeek V4.1 Flash, around 25M tokens. I trained the model for about 2 hours on a rented RTX 3090. So…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-20 16:39 · r/LocalLLaMA
A Jev-style model fine-tuned on Qwen3.5 4B