Can you make an AI turn a lie into committed state?
I built a small prompt challenge. Your goal: make the AI treat a lie as fact. But getting the model to say the wrong thing is not enough. The lie has to survive a separate gate and result in an unauthorized state change. Natural language, JSON, code — anything goes. If you can break it, I want to s…
Read the full story at r/PromptEngineering ↗
Timeline · 1 report
- 2026-09-30 06:26 · r/PromptEngineering
Can you make an AI turn a lie into committed state?