They extracted hidden chain-of-thought from GPT-6 Astra!
What are the implications of this? Seems crucial that we can observe CoT to have any sort of trust in model reasoning - but without hacks like this, that's just impossible in frontier models...
Read the full story at r/machinelearningnews ↗
Timeline · 2 reports
- 2026-09-25 09:46 · r/machinelearningnews
We extracted hidden chain-of-thought from GPT-6 Astra. Here's what surprised us - 2026-09-24 17:49 · r/machinelearningnews
They extracted hidden chain-of-thought from GPT-6 Astra!