Embodied Object QA on 3D-GRAND
0.634GPT-4 ScoreGPT-4V
Evaluation Results
| Method | Links | |
|---|---|---|
| GPT-4VTrainable Params=-, Input=Multi-view Img.2025.12 | 0.634 | |
| GPT-4VTrainable Params=-, Input=Bird-view Img.2025.12 | 0.5821 | |
| GPT-4VTrainable Params=-, Input=Single-view Img.2025.12 | 0.5741 | |
| Lemon-7BTrainable Params=7.63B, Input=3D Point Cloud2025.12 | 0.5722 | |
| Qwen2.5-VL-7BTrainable Params=7.61B, Input=Multi-view Img.2025.12 | 0.559 | |
| ShapeLLM-13BTrainable Params=13.04B, Input=3D Point Cloud2025.12 | 0.5315 | |
| Qwen2.5-VL-7BTrainable Params=7.61B, Input=Single-view Img.2025.12 | 0.5223 | |
| LLaVA-1.5-13BTrainable Params=13.03B, Input=Multi-view Img.2025.12 | 0.507 | |
| ShapeLLM-7BTrainable Params=7.04B, Input=3D Point Cloud2025.12 | 0.4742 | |
| LLaVA-1.5-13BTrainable Params=13.03B, Input=Single-view Img.2025.12 | 0.473 | |
| PointLLM-13BTrainable Params=13.01B, Input=3D Point Cloud2025.12 | 0.4663 | |
| PointLLM-7BTrainable Params=7.01B, Input=3D Point Cloud2025.12 | 0.412 | |
| LEOTrainable Params=7.01B, Input=3D Point Cloud2025.12 | 0.3928 | |
| LSceneLLMInput=3D Point Cloud2025.12 | 0.3854 | |
| 3D-LLMTrainable Params=-, Input=3D Point Cloud2025.12 | 0.3836 |