Medical Visual Question Answering on OBScan Closed
0.9939AccuracyPhi3.5V-Med
Evaluation Results
| Method | Links | |
|---|---|---|
| Phi3.5V-MedEvaluation Protocol=Fine-Tuning, Decoding Strategy=OPERA2025.12 | 0.9939 | |
| Phi3.5V-MedEvaluation Protocol=Fine-Tuning, Decoding Strategy=Greedy2025.12 | 0.9817 | |
| Phi3.5V-MedEvaluation Protocol=Fine-Tuning, Decoding Strategy=DoLA2025.12 | 0.9817 | |
| Phi3.5V-MedEvaluation Protocol=Fine-Tuning, Decoding Strategy=ARCD2025.12 | 0.9817 | |
| Phi3.5V-MedEvaluation Protocol=Fine-Tuning, Decoding Strategy=VCD2025.12 | 0.9756 | |
| Qwen-VL-7BEvaluation Protocol=General Domain, Decoding Strategy=Greedy2025.12 | 0.7195 | |
| GPT-4oEvaluation Protocol=General Domain, Decoding Strategy=Greedy2025.12 | 0.7073 | |
| Phi3.5V-MedEvaluation Protocol=Zero-Shot, Decoding Strategy=ARCD2025.12 | 0.6829 | |
| HuatuoV-34BEvaluation Protocol=Zero-Shot, Decoding Strategy=Greedy2025.12 | 0.6768 | |
| Phi3.5V-MedEvaluation Protocol=Zero-Shot, Decoding Strategy=Greedy2025.12 | 0.6585 | |
| Phi3.5V-MedEvaluation Protocol=Zero-Shot, Decoding Strategy=VCD2025.12 | 0.6402 | |
| Phi3.5V-MedEvaluation Protocol=Zero-Shot, Decoding Strategy=OPERA2025.12 | 0.6402 | |
| LLaVA-1.6-7BEvaluation Protocol=General Domain, Decoding Strategy=Greedy2025.12 | 0.628 | |
| Phi3.5V-MedEvaluation Protocol=Zero-Shot, Decoding Strategy=DoLA2025.12 | 0.628 | |
| LLaVA-Med-7BEvaluation Protocol=Zero-Shot, Decoding Strategy=Greedy2025.12 | 0.6037 | |
| Phi3.5V-4.2BEvaluation Protocol=General Domain, Decoding Strategy=Greedy2025.12 | 0.561 |