If a prompt only works at one decoding setting, is it actually a robust prompt?
This story is from 2026-09-02. It is preserved in the archive; the latest stories are on the live feed.
Prompt comparisons often publish the prompt and hide the decoding setup, even though the two are part of the same system. The Ling-3.0-flash-Fin benchmark notes make that visible. Unless otherwise specified, the model was evaluated at temperature 1 and top-p 0.95. In the SpreadsheetBench Claude Cod…
Read the full story at r/PromptEngineering ↗
Timeline · 1 report
- 2026-09-02 16:01 · r/PromptEngineering
If a prompt only works at one decoding setting, is it actually a robust prompt?