Intrinsic Reasoning on UNLI
0.813Spearman CorrelationAlways Tell Me The Odds
Evaluation Results
| Method | Links | |
|---|---|---|
| Always Tell Me The OddsBackbone=Qwen2.5-14B-Instruct2025.05 | 0.813 | |
| Always Tell Me The OddsBackbone=Qwen2.5-14B-Instruct, Data Augmentation=+Syn2025.05 | 0.812 | |
| Always Tell Me The OddsBackbone=Qwen2.5-14B-Instruct, Data Augmentation=+Syn, Training Strategy=+R2025.05 | 0.812 | |
| Always Tell Me The OddsBackbone=Qwen2.5-8B-Instruct2025.05 | 0.804 | |
| Always Tell Me The OddsBackbone=Qwen2.5-7B-Instruct2025.05 | 0.802 | |
| RoBERTa-LType=Encoder2025.05 | 0.707 | |
| GPT-4oEvaluation Protocol=0-Shot2025.05 | 0.699 | |
| Llama-3-InstructEvaluation Protocol=Probe, Size=14B2025.05 | 0.681 | |
| DeepSeek-R1-Distill-Qwen-32BEvaluation Protocol=0-Shot2025.05 | 0.629 |