Moonshot’s PerceptionBench finds no tested vision model above 60%
Moonshot AI has released PerceptionBench, an open benchmark for testing what multimodal models actually see. Across 16 frontier models, the company says none reached 60% accuracy, with perception-related hallucination the weakest capability
Open discussion →