GPEC: Efficient Pre-LLM Gaussian Process Embedding Correction for Cardiac Video Caption Generation
arXiv:2610.00196v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown strong potential for video understanding and caption generation, but their performance may decline in specialized medical imaging domains such as echocardiography. This work introduces Gaussian Proce…
Read the full story at arXiv cs.CV ↗
Timeline · 1 report
- 2026-10-02 04:00 · arXiv cs.CV
GPEC: Efficient Pre-LLM Gaussian Process Embedding Correction for Cardiac Video Caption Generation