多模态检索与生成任务竞赛结果出炉,多个系统超越去年最佳表现。
Findings of the MAGMaR 2026 Shared Task

- 参赛团队分别聚焦视频检索与基于检索视频的文本生成
- 17个检索系统全部优于去年冠军基线,16个生成系统均产出人类评选最佳报告
- 适合关注多模态生成与信息检索交叉研究的研究者参考
本文综述了第二届多模态增强生成与多模态检索研讨会(MAGMaR)共享任务的成果。本次任务分为视频检索与基于检索视频的文本生成两个方向,参赛团队可任选其一提交。在检索任务中,共有2支团队提交了17个系统,全部超越了去年冠军基准模型的表现。在生成任务中,4支团队提交了16个系统,每个团队至少有一个生成报告被人工标注为最优。结果表明,当前方法在多模态生成与检索结合方面已取得显著进展。
原文摘要 · Abstract (English)
This overview paper presents the results of the shared task for the second workshop on Multimodal Augmented Generation via Multimodal Retrieval (MAGMaR). In this shared task participants submitted systems focused on either (i) video retrieval or (ii) grounded generation of articles given retrieved videos. Teams could submit to either task. For the retrieval task, we had 2 participating teams that submitted a total of 17 systems -- all of which beat a baseline derived from the winner of last year's shared task. On the generation side, we had 4 teams submit 16 systems. All teams had at least one generated report that was labeled the best by a human annotator.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。