Intrinsic Reasoning on e-CARE
0.905Spearman CorrelationAlways Tell Me The Odds
Evaluation Results
| Method | Links | |
|---|---|---|
| Always Tell Me The OddsBackbone=Qwen2.5-14B-Instruct, Data Augmentation=+Syn, Training Strategy=+R2025.05 | 0.905 | |
| GPT-4oEvaluation Protocol=0-Shot2025.05 | 0.898 | |
| Always Tell Me The OddsBackbone=Qwen2.5-14B-Instruct2025.05 | 0.888 | |
| Always Tell Me The OddsBackbone=Qwen2.5-14B-Instruct, Data Augmentation=+Syn2025.05 | 0.884 | |
| DeepSeek-R1-Distill-Qwen-32BEvaluation Protocol=0-Shot2025.05 | 0.871 | |
| Always Tell Me The OddsBackbone=Qwen2.5-7B-Instruct2025.05 | 0.87 | |
| Llama-3-InstructEvaluation Protocol=Probe, Size=14B2025.05 | 0.856 | |
| Always Tell Me The OddsBackbone=Qwen2.5-8B-Instruct2025.05 | 0.855 | |
| RoBERTa-LType=Encoder2025.05 | 0.738 |