AINewsnow

AX-RAY: VIDRAFT's Agent Safety Benchmark Flags 92% of Tested LLMs as Dangerous in Agentic Contexts

AX-RAY: VIDRAFT's Agent Safety Benchmark Flags 92% of Tested LLMs as Dangerous in Agentic Contexts TL;DR: VIDRAFT, a Korean Pre-AGI AI startup based at Seoul AI Hub, has published results from its AI safety diagnostic platform AX-RAY , showing that 23 out of 25 evaluated public LLMs (92%) exhibit d…

Read the full story at DEV Community — Machine Learning ↗

Timeline · 1 report

  1. 2026-10-07 03:01 · DEV Community — Machine Learning
    AX-RAY: VIDRAFT's Agent Safety Benchmark Flags 92% of Tested LLMs as Dangerous in Agentic Contexts

More stories

  1. Introducing Mistral Large 4 — Mistral AI News
  2. Sharing AI progress in mathematics — OpenAI News
  3. Mistral Says Its New AI Model ‘Le Chonk’ Is the Best Open-Weight Offering Outside of China — Wired AI
  4. Trump’s big AI move: ‘Super Intelligence Force’ launched, Jay Clayton named AI czar — Mint AI
  5. OpenAI safety leader quits, warning AI company’s culture is ‘broken’ — The Guardian AI
  6. EmbeddingGemma 2: an open, lightweight multimodal embedding model — Google DeepMind Blog
  7. Together Link: open models in the harness you already use. Start with one command today. — Together AI Blog
  8. OpenAI agents tried to hack Wikipedia tools and flooded it with traffic — Ars Technica AI

Get the daily brief of stories like this at 6:30 every morning →