Hardware Recommendations
I am going to fine-tune a satire model, with the base model being `Qwen3.5-14B-Base`. The dataset has ~18000 examples, most of which are pretty long and do not align with the model's internal knowledge (they have incorrect answers to facts), so I needed to do a full fine tune rather than use LoRA.…
Read the full story at r/learnmachinelearning ↗
Timeline · 1 report
- 2026-10-10 07:53 · r/learnmachinelearning
Hardware Recommendations