AINewsnow

We had a blind judge guess human or AI on our support agent's replies. 40 out of 40 said AI

Our agents answer customers on web chat, voice, WhatsApp and phone. They were accurate and still felt like bots, so we tested it. The setup A made-up car dealership as the knowledge base, so any fact not in it was provably invented. Two models, gemini flash-lite and a self-hosted qwen3-14b. A blind…

Read the full story at r/PromptEngineering ↗

Timeline · 1 report

  1. 2026-10-05 11:28 · r/PromptEngineering
    We had a blind judge guess human or AI on our support agent's replies. 40 out of 40 said AI

More stories

  1. I noticed this happens to nearly all Google models even Nanobana on web now is nerfed — r/GeminiAI
  2. Upcoming changes to Gemini model access starting October 9th — r/GeminiAI
  3. Gemini Plus removing Pro model nerfed into oblivion — r/GeminiAI
  4. Create your own voices with Gemini 3.8 text-to-speech — Google DeepMind YouTube
  5. Day 3 No Gemini 4 — r/GeminiAI
  6. WTF Google getting rid of free Gemini flash and pro — r/GeminiAI
  7. The new nanobanana pro is coming. — r/GeminiAI
  8. How do I stop feeling left behind with Gemini Pro? (Trying to replicate Claude Code / homelab setups) — r/Bard

Get the daily brief of stories like this at 6:30 every morning →