AINewsnow

JetBrains Ranked AI Agents on Real Kotlin Projects. The Token Column Is the Real Story.

This story is from 2026-09-15. It is preserved in the archive; the latest stories are on the live feed.

JetBrains released the Kotlin Benchmark , its official benchmark for grading AI coding agents on real Kotlin engineering work. Claude Code with Opus 4.7 xhigh leads the first leaderboard at 85.7 percent , with JetBrains Junie and OpenAI's Codex right behind at 81.9 percent. But the resolution rate…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-15 16:06 · DEV Community — AI
    JetBrains Ranked AI Agents on Real Kotlin Projects. The Token Column Is the Real Story.

More stories

  1. Anthropic selects Accenture as first embedded evaluator to help implement Amodei's slowdown proposal — CNBC Technology
  2. OpenAI ‘ethically hacked’ with help of Anthropic’s Claude chatbot — The Guardian AI
  3. Researchers used Claude to hack OpenAI — Ars Technica AI
  4. I ran Claude code and Codex in parallel for 15 days. Here's what I found. — r/AI_Agents
  5. I built an iOS app with Claude code to break out of my usual chord habits and unlock new progressions. — r/ClaudeAI
  6. Pay $39.99 once to put ChatGPT, Claude, Gemini, and more in a single workspace for life — Mashable AI
  7. Independent Security Researchers Used Anthropic’s Claude to Break Into OpenAI — r/singularity
  8. This Ford exec put her family's Claude assistant on a PIP. ChatGPT has taken over. — Business Insider AI

Get the daily brief of stories like this at 6:30 every morning →