i built a benchmark to test whether LLMs can understand and create jokes
This story is from 2026-09-10. It is preserved in the archive; the latest stories are on the live feed.
hey guys! i built lolbench - a benchmark for LLM humor LLMs take three tests: - explain why jokes work (or don't) - write jokes under shared premises - predict which jokes humans prefer The finding so far that surprised me: every model aces explaining real jokes (95%+) but drops hard on explaining…
Read the full story at r/artificial ↗
Timeline · 1 report
- 2026-09-10 17:00 · r/artificial
i built a benchmark to test whether LLMs can understand and create jokes