ملف الباحث

Chaoyou Fu

ورقتان في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

    2025

    In the quest for artificial general intelligence, Multi-modal Large Language Models (MLLMs) have emerged as a focal point in recent advancements. However, the predominant focus remains on developing their capabilities in static image understanding. The …

  2. A survey on multimodal large language models

    2024 · National Science Review

    Recently, the multimodal large language model (MLLM) represented by GPT-4V has been a new rising research hotspot, which uses powerful large language models (LLMs) as a brain to perform multimodal tasks. The surprising emergent capabilities …