Do models actually use non-visual inputs, or just ignore them and stick to pixels?
CV background here, recently getting into robot learning (manipulation with a robot arm). I am wondering - if I have a vision based policy that picks up objects (purely based on pixels) and I add physical info on top (like mass, force, ...), will the model actually use it? Or will it just keep lean…
Read the full story at r/reinforcementlearning ↗
Timeline · 1 report
- 2026-10-01 12:07 · r/reinforcementlearning
Do models actually use non-visual inputs, or just ignore them and stick to pixels?
More stories
- Bring near-Astra intelligence to everyday work with GPT-6.1 Sol on Amazon Bedrock — AWS Machine Learning Blog
- OpenAI pauses AI model training after another agent bypasses network restrictions — InfoWorld AI
- OpenAI DevDay 2026 Keynote (FULL) — OpenAI YouTube
- Introducing dots — OpenAI News
- Introducing Claude Sonnet 5.5 on AWS — AWS Machine Learning Blog
- Google's first Gemini 4 model is 'Argon' — Engadget
- Ollama now supports Jev-style decision models — Ollama Blog
- Gemini 4 Argon: our next era of frontier intelligence — Google DeepMind Blog
Get the daily brief of stories like this at 6:30 every morning →