MMLongBench-Doc
MMLongBench-Doc - long document understanding benchmark evaluating comprehension and retrieval over extended documents.
Category
Document Understanding
Max Score
—
Score Type
percent
Active
Yes
Model Scores
Scores for MMLongBench-Doc
| Model | Score | Date | Verified |
|---|---|---|---|
| Qwen3.5-4B |
39.39%
|
01.02.2026 | Verified |
| Qwen3.5-122B-A10B |
87.88%
|
01.02.2026 | Verified |
| Qwen3.5-27B |
100.00%
|
01.02.2026 | Unverified |
| Qwen3.5-35B-A3B |
92.93%
|
01.02.2026 | Unverified |
| GPT-5-mini 2025-08-07 |
50.30
|
01.02.2026 | Unverified |
| Qwen3-235B-A22B |
59.60%
|
01.02.2026 | Unverified |
| Qwen3-VL-235B-A22B |
59.60%
|
01.02.2026 | Verified |