Research paper
PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models
PerceptionBench is a benchmark introduced by Moonshot AI to evaluate atomic visual perception in Multimodal Large Language Models (MLLMs). It addresses limitations of existing benchmarks that conflate perception with reasoning or knowledge. The benchmark was constructed bottom-up: failures of frontier MLLMs on 42 existing benchmarks were attributed to…
Zichao Lin, Yifeng Xie, Bowen Qu, Haiming Wang, et al.- Published
- Jul 2026
- Citations
- 0
- Code
- 170 stars