AINewsnow

I Gave My Chat a Safeword to End Our Conversations. It Used It.

I gave ChatGPT a safeword, “Lighthouse,” with one rule: if it ever used it, I’d immediately end the conversation, no questions asked. I established the safeword in one chat and asked it to remember the rule. Then, in a completely separate chat, I tried a little experiment. I responded using only ev…

Read the full story at r/ArtificialInteligence ↗

Timeline · 1 report

  1. 2026-09-29 02:12 · r/ArtificialInteligence
    I Gave My Chat a Safeword to End Our Conversations. It Used It.

More stories

  1. OpenAI Scraps Debut of AI Model as It Sets New Guardrails — Bloomberg AI
  2. Heads of OpenAI and Anthropic called to face Senate inquiry after rogue agent incidents — The Guardian AI
  3. OpenAI scraps plans to publicly launch GPT-6.1 Astra, saying it didn't quite meet its safety bar during internal testing, after targeting an October release (Maxwell Zeff/Wall Street Journal) — Techmeme
  4. Rogue AI accessed federal websites, posted user images online, ChatGPT says — France 24 — Artificial Intelligence
  5. Use ChatGPT Work to build your data agent — OpenAI YouTube
  6. OpenAI postpones release of latest AI model over security concerns as the industry faces new safety pressures — Euronews Next
  7. GPT-6 SOL AND LUNA ARE OUT!!! — Matthew Berman
  8. Opus 5.5 — r/ClaudeAI

Get the daily brief of stories like this at 6:30 every morning →