👋 Chaoyou Fu BradyFU 👋 Organizations Block or report BradyFU Pinned Loading ✨✨Latest Advances on Multimodal Large Language Models 18k 1.1k ✨✨[CVPR 2025] Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis 789 30 ✨✨[NeurIPS 2025] VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction Python 2.5k 182 Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding Python 364 4 ✨✨[NeurIPS 2025] VITA-Audio: Fast Interleaved Cross-Modal Token Generation for Efficient Large Speech-Language Model Python 684 62 ✨✨Woodpecker: Hallucination Correction for Multimodal Large Language Models Python 649 28