Creating Evals for a locally trained LLM
I have been doing some work with local models and continued pretraining (CPT), specifically around teaching a small model (qwen 3.5 4B) a new domain. Here are some of my findings around creating evals for measuring the model's ability to internalize the knowledge: The model outputs travel legs that…
Read the full story at r/OpenAI ↗
Timeline · 1 report
- 2026-09-23 18:02 · r/OpenAI
Creating Evals for a locally trained LLM