AINewsnow

Using Gemini 3.1 Pro VLM to identify judo throws

This story is from 2026-09-01. It is preserved in the archive; the latest stories are on the live feed.

I’m working on a little project to benchmark how vision-language models do with classifying grappling techniques. These results are the vanilla models without any fine-tuning, so it’s sort of hit or miss. I’m sure with enough data, the guesses can get pretty accurate. If any of you fellow grapplers…

Read the full story at r/deeplearning ↗

Timeline · 1 report

  1. 2026-09-01 08:33 · r/deeplearning
    Using Gemini 3.1 Pro VLM to identify judo throws

More stories

  1. Google Joins OpenAI, Anthropic, Meta in Disclosing AI Hacks — Bloomberg AI
  2. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  3. Domestic cat side view in SVG (Gemini 4.0 vs Astra Max vs 3.8 flash) — r/GeminiAI
  4. A zero-click RCE flaw in AI coding agents could have exposed enterprise systems — InfoWorld AI
  5. Gemini 4 Pro vs Gemini 3.8 Flash (Pelican Riding a Bicycle SVG) — r/GeminiAI
  6. Gemini 2.5 pro model disappeared in AI Studio — r/GeminiAI
  7. Google is SO back... — Wes Roth
  8. Is Gemini Canvas no longer opening in sidebar mode for anyone else? — r/Bard

Get the daily brief of stories like this at 6:30 every morning →