Perception on MMStar (test)
72.3AccuracyVaLR-M
Evaluation Results
| Method | Links | |
|---|---|---|
| VaLR-MModel Category=Latent Reasoning models, Encoder=Multiple (DINOv3, SigLIPv2, pi^3)2026.02 | 72.3 | |
| VaLR-SModel Category=Latent Reasoning models, Encoder=Single (DINOv3)2026.02 | 70.8 | |
| CoVTModel Category=Latent Reasoning models2026.02 | 69.2 | |
| Qwen2.5-VL-7B + vanilla SFTModel Category=Base model, Training=vanilla SFT2026.02 | 67.5 | |
| Qwen2.5-VL-7BModel Category=Base model, Parameters=7B2026.02 | 67.1 | |
| GPT-4oModel Category=API models2026.02 | 65.2 | |
| LVTModel Category=Latent Reasoning models2026.02 | 64.4 | |
| Ocean-R1-7BModel Category=Reasoning models, Parameters=7B2026.02 | 62.6 | |
| Claude-3-SonnetModel Category=API models2026.02 | 58.8 | |
| R1-OneVision-7BModel Category=Reasoning models, Parameters=7B2026.02 | 55.6 | |
| MonetModel Category=Latent Reasoning models2026.02 | 53.3 |