AINewsnow

NVIDIA's Physis-Lang trains video models on self-evolving physics captions, beats Veo 3.1 on 3 of 4 benchmarks

Physis-Lang (NVIDIA, MIT, Oxford) adds a physics reasoning field and a scene-specific negative prompt to video captions. An agent refines the captioning instruction in a loop while the captioner itself stays frozen. PhyGenBench: Cosmos3-Nano + Physis-Lang scores 71.04, vs 65.63 for Veo 3.1 and 61.6…

Read the full story at r/machinelearningnews ↗

Timeline · 1 report

  1. 2026-09-30 07:37 · r/machinelearningnews
    NVIDIA's Physis-Lang trains video models on self-evolving physics captions, beats Veo 3.1 on 3 of 4 benchmarks

More stories

  1. NVIDIA DGX Spark 64GB Gives Developers More Ways to Build and Scale Local AI — NVIDIA Blog
  2. Top AI and tech firms sign 'morally binding' accord to 'self-police' development after meeting at White House — Euronews Next
  3. NVIDIA Vera CPU Is Coming to CoreWeave: Pack In More Agents — CoreWeave Blog
  4. What It Takes to Bring Up a Multi-Rack NVIDIA Vera Rubin NVL72 Cluster — CoreWeave Blog
  5. What Comes Next: Operating and Evolving the Production AI Factory — CoreWeave Blog
  6. Why AI Factories Need Proof Before Production — CoreWeave Blog
  7. Liquid-Cooled Switching Doubles AI Network Bandwidth Per Rack — CoreWeave Blog
  8. China’s DeepSeek open-sources tools to help Huawei chips supplant Nvidia in AI — South China Morning Post Tech

Get the daily brief of stories like this at 6:30 every morning →