Perception on V* (test)
86.9AccuracyVaLR-M
Evaluation Results
| Method | Links | |
|---|---|---|
| VaLR-MModel Category=Latent Reasoning models, Encoder=Multiple (DINOv3, SigLIPv2, pi^3)2026.02 | 86.9 | |
| VaLR-SModel Category=Latent Reasoning models, Encoder=Single (DINOv3)2026.02 | 86.4 | |
| MonetModel Category=Latent Reasoning models2026.02 | 83.3 | |
| LVTModel Category=Latent Reasoning models2026.02 | 81.7 | |
| Ocean-R1-7BModel Category=Reasoning models, Parameters=7B2026.02 | 78 | |
| Qwen2.5-VL-7B + vanilla SFTModel Category=Base model, Training=vanilla SFT2026.02 | 78 | |
| CoVTModel Category=Latent Reasoning models2026.02 | 78 | |
| Qwen2.5-VL-7BModel Category=Base model, Parameters=7B2026.02 | 76.4 | |
| R1-OneVision-7BModel Category=Reasoning models, Parameters=7B2026.02 | 59.2 | |
| GPT-4oModel Category=API models2026.02 | 42.9 | |
| Claude-3-SonnetModel Category=API models2026.02 | 15.2 |