PerceptionBench: Evaluating Atomic Visual Perception in MLLMs

Kimi Team released PerceptionBench, a 3,000-question benchmark isolating 10 atomic visual perception capabilities, on which no frontier MLLM reaches 60% accuracy.

Multimodal model developers can now measure and diagnose which atomic perception capabilities their models lack, since models with nearly identical overall scores diverge sharply in what they actually perceive.

Sources

Read this as text

Back to the AI news