ملف الباحث

Chenhao Cui

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Modeling Paragraph-Level Vision-Language Semantic Alignment for Multi-Modal Summarization

    2022 · arXiv (Cornell University)

    Most current multi-modal summarization methods follow a cascaded manner, where an off-the-shelf object detector is first used to extract visual features, then these features are fused with language representations to generate the summary with an …