Skip to content
@MME-Benchmarks

MME Benchmarks

Multimodal LLM Benchmarks of MME series

Pinned Loading

  1. Video-MME Video-MME Public

    ✨✨[CVPR 2025] Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

    791 30

  2. Video-MME-v2 Video-MME-v2 Public

    Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding

    Python 372 4

  3. MME-RealWorld MME-RealWorld Public

    ✨✨ [ICLR 2025] MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?

    Python 162 10

  4. MME-CoT MME-CoT Public

    MME-CoT: Benchmarking Chain-of-Thought in LMMs for Reasoning Quality, Robustness, and Efficiency

    Python 135 7

Repositories

Showing 5 of 5 repositories
  • Video-MME-v2 Public

    Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding

    MME-Benchmarks/Video-MME-v2's past year of commit activity
    Python 372 4 5 0 Updated Aug 5, 2026
  • Video-MME Public

    ✨✨[CVPR 2025] Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

    MME-Benchmarks/Video-MME's past year of commit activity
    791 30 11 0 Updated Dec 8, 2025
  • MME-RealWorld Public

    ✨✨ [ICLR 2025] MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?

    MME-Benchmarks/MME-RealWorld's past year of commit activity
    Python 162 10 5 0 Updated Oct 21, 2025
  • MME-CoT Public

    MME-CoT: Benchmarking Chain-of-Thought in LMMs for Reasoning Quality, Robustness, and Efficiency

    MME-Benchmarks/MME-CoT's past year of commit activity
    Python 135 7 3 0 Updated Aug 5, 2025
  • MME-Unify Public

    ✨✨ [ICLR 2026] MME-Unify: A Comprehensive Benchmark for Unified Multimodal Understanding and Generation Models

    MME-Benchmarks/MME-Unify's past year of commit activity
    Python 43 4 0 0 Updated Apr 10, 2025