AINewsnow

Inductive Visual Logic for Few-Shot Out-Of-Distribution Adaptation in VLMs

arXiv:2609.38362v1 Announce Type: new Abstract: Generative vision-language models (VLMs) such as Qwen-VL and LLaVA achieve strong zero-shot performance on tasks overlapping with their pretraining distribution, yet fail on specialized domains where the required discriminative features were never lea…

Read the full story at arXiv cs.CV ↗

Timeline · 1 report

  1. 2026-10-01 04:00 · arXiv cs.CV
    Inductive Visual Logic for Few-Shot Out-Of-Distribution Adaptation in VLMs

More stories

  1. Two open-weights releases: Victoria (Qwen3.8-Flash-Next with 44% of experts cut, 70% Terminal-Bench 2.1, GGUF included) and Maple (a Canada-first fine-tune) — r/LocalLLM
  2. If you are running Qwen 3.8 Flash Next on Strix Halo, use this software for inference. It's so much faster than llama.cpp especially at high context. — r/LocalLLaMA
  3. Layer Extract & Layer Remove Loras For Qwen Image 2.1 — r/StableDiffusion
  4. I tested Qwen Image 2.1, FLUX Klein 2, Krea 2, and Z Image in ComfyUI using the same prompts. Some interesting differences came out 👀 Full comparison here if you want to check it out! — r/comfyui
  5. What model sits between Qwen 3.8 27b and Flash next for coding? — r/LocalLLaMA
  6. vulkan: fuse qwen4exp's SCALE -> SIGMOID -> SCALE -> hc_post chain by fxgsell · Pull Request #29520 · ggml-org/llama.cpp — r/LocalLLaMA
  7. Qwen-family LLMs are quietly becoming the backbone of modern audio models; One chart for the architectures of 100+ audio models — r/LocalLLaMA
  8. Best Open source TTS models to run locally, I´ve been using Qwen tts 1.7b. Is there something small and better? — r/comfyui

Get the daily brief of stories like this at 6:30 every morning →