Fine-tuning Qwen2.5-1.5B on a Mac to fix a parroting chatbot (and halve the prompt)
By Georgii Kharlampiiev, Mindscend Our company website is a single chat page backed by a model we host ourselves. This post walks through how we fine-tuned that model with LoRA on a MacBook, how we checked it was better, and how we shipped it to a CPU-only server. All the numbers are from our own r…
Read the full story at DEV Community — Machine Learning ↗
Timeline · 1 report
- 2026-10-08 02:48 · DEV Community — Machine Learning
Fine-tuning Qwen2.5-1.5B on a Mac to fix a parroting chatbot (and halve the prompt)