Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

MMMU

non-profit
https://mmmu-benchmark.github.io/
MMMU-Benchmark
Activity Feed Request to join this org

AI & ML interests

Multimodal Model Evaluation

Recent Activity

zhangysk  authored a paper 6 days ago
HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness?
zhangysk  authored a paper 6 days ago
Aspire: Can Models Self-Evolve from Vague Goals?
zhangysk  authored a paper 6 days ago
S3Gym: Can LLMs Turn Self-Testing and Self-Judging into Self-Improvement?
View all activity

Xiang Yue's profile picture Yuansheng Ni's profile picture Kai Zhang's profile picture Yu Su's profile picture yuxuan sun's profile picture Ge Zhang's profile picture Huan Sun's profile picture Dongfu Jiang's profile picture Renliang Sun's profile picture Boyuan Zheng's profile picture Wenhu Chen's profile picture Yibo Liu's profile picture Ruibin Yuan's profile picture Weiming Ren's profile picture Ming Yin's profile picture TY.Zheng's profile picture Graham Neubig's profile picture

MMMU 's datasets 2

MMMU/MMMU

Viewer • Updated Jul 10 • 11.6k • 76.9k • 333

MMMU/MMMU_Pro

Benchmark • Updated Jul 10 • 5.19k • 22.3k • 66
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs