Fine-tuning a legal LLM: combine tasks into one adapter or train separate LoRAs? (+ roadmap questions for a domain "lawyerGPT")
This story is from 2026-08-31. It is preserved in the archive; the latest stories are on the live feed.
Working on Polish legal AI as a founder and I've hit a fork in the road. Would love input and also it's kind of my hobby so i would like to learn. Where I am: Task 1 (done, v2): legal reasoning summarizer. Qwen3.8-27B + LoRA, trained on CoT distilled from opus. Result: the student now beats its own…
Read the full story at r/MLQuestions ↗
Timeline · 1 report
- 2026-08-31 21:15 · r/MLQuestions
Fine-tuning a legal LLM: combine tasks into one adapter or train separate LoRAs? (+ roadmap questions for a domain "lawyerGPT")