GitHub

View BradyFU's full-sized avatar

👋

Chaoyou Fu BradyFU

👋

Organizations

@VITA-MLLM @MME-Benchmarks

Block or report BradyFU

Pinned Loading

  1. ✨✨Latest Advances on Multimodal Large Language Models

    18k 1.1k

  2. ✨✨[CVPR 2025] Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

    789 30

  3. ✨✨[NeurIPS 2025] VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

    Python 2.5k 182

  4. Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding

    Python 364 4

  5. ✨✨[NeurIPS 2025] VITA-Audio: Fast Interleaved Cross-Modal Token Generation for Efficient Large Speech-Language Model

    Python 684 62

  6. ✨✨Woodpecker: Hallucination Correction for Multimodal Large Language Models

    Python 649 28

Read the original on github.com ↗