AV-Reasoner: Improving and Benchmarking Clue-Grounded Audio-Visual Counting for MLLMs2026年5月5日·Lidong Lu,Guo Chen,Zhu Wei,Zhiqi Li,Yicheng LiuTong Lu· 0 分钟阅读时长 引用 URL类型会议文章出版物Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)最近更新于 2026年5月5日AuthorsTong Lu南京大学← Understanding New-Knowledge-Induced Factual Hallucinations in LLMs: Analysis and Interpretation 2026年5月8日Bayesian Decomposition and Semantic Completion for Few-shot Semantic Segmentation 2026年5月5日 →