PerceptionBench

PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models

LLM OrchestrationOpen Source AI & Machine LearningOpen Source

About

PerceptionBench is a benchmark designed to assess the atomic visual perception abilities of multimodal large language models (MLLMs). It addresses limitations in existing benchmarks by isolating perceptual errors from reasoning or domain knowledge failures. The benchmark includes 3,000 verified questions that target ten atomic perceptual capabilities, providing a structured approach to evaluate and diagnose the visual perception boundaries of MLLMs.

Open Source Health

Not enough history
Stars
206
Forks
11
License
Apache-2.0
Last commit
2 months ago
Python

Related Categories

Using PerceptionBench?

Track its cost next to the rest of your stack and get a reminder before it renews.

Add to my stack

Vendor

Moonshot AI

Open-source projects on GitHub: Kimi-Dev, MoonEP and Checkpoint Engine

View vendor profile

More by Moonshot AI