Abliterated models
I saw something called obliterated models that use llms and change them by calculating a direction vector of harm and orthogonalizing the weights. I think this is a very very cool thing!!!!!! I am so surprised this can be done. What is the tradeoff in terms of model accuracy, I am surprised it does…
Read the full story at r/deeplearning ↗
Timeline · 1 report
- 2026-09-23 04:27 · r/deeplearning
Abliterated models