Thematic fit estimation on Fer-Loc
0.69Spearman Correlation (ρ)Step-by-Step Prompting
Evaluation Results
| Method | Links | |
|---|---|---|
| Step-by-Step PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.3.1, Input Form=Lemma Tuples, Output Form=Numeric2024.10 | 0.69 | |
| Step-by-Step PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.3.2, Input Form=Lemma Tuples, Output Form=Categorical2024.10 | 0.68 | |
| Simple PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.1.1, Input Form=Lemma Tuples, Output Form=Numeric2024.10 | 0.63 | |
| Simple PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.1.2, Input Form=Lemma Tuples, Output Form=Categorical2024.10 | 0.61 | |
| GPT4.1 (Exp.3.1)Reasoning=Step-by-Step Prompting, Input=Lemma Tuples, Output=Numeric2024.10 | 0.59 | |
| Step-by-Step PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.4.2, Input Form=Generated Sentences, Output Form=Categorical2024.10 | 0.58 | |
| Step-by-Step PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.4.1, Input Form=Generated Sentences, Output Form=Numeric2024.10 | 0.57 | |
| GPT4.1 (Exp.3.2)Reasoning=Step-by-Step Prompting, Input=Lemma Tuples, Output=Categorical2024.10 | 0.56 | |
| GPT4.1 (Exp.4.1)Reasoning=Step-by-Step Prompting, Input=Generated Sentences, Output=Numeric2024.10 | 0.56 | |
| GPT4.1 (Exp.4.2)Reasoning=Step-by-Step Prompting, Input=Generated Sentences, Output=Categorical2024.10 | 0.53 | |
| GPT4.1 (Exp.1.1)Reasoning=Standard Prompting, Input=Lemma Tuples, Output=Numeric2024.10 | 0.51 | |
| GPT4.1 (Exp.1.2)Reasoning=Standard Prompting, Input=Lemma Tuples, Output=Categorical2024.10 | 0.5 | |
| Simple PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.2.1, Input Form=Generated Sentences, Output Form=Numeric2024.10 | 0.48 | |
| ResRoFA-MT2024.10 | 0.46 | |
| ResRoFA-MTMethod ID=B12024.10 | 0.46 | |
| ResRoFA-MTLabel=B12024.10 | 0.46 | |
| Simple PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.2.2, Input Form=Generated Sentences, Output Form=Categorical2024.10 | 0.46 | |
| NN RF2024.10 | 0.44 | |
| NN RFMethod ID=B02024.10 | 0.44 | |
| NN RFLabel=B02024.10 | 0.44 | |
| Exp.3.1Backbone=Qwen2024.10 | 0.41 | |
| Exp.3.2Backbone=Qwen2024.10 | 0.41 | |
| Exp.1.2Backbone=Qwen2024.10 | 0.38 | |
| Exp.2.2Backbone=Qwen2024.10 | 0.37 | |
| Glove-sharedtuning=tuned, data=20% subset2024.10 | 0.35 | |
| Glove-shared tuned with 20%MMethod ID=B3c2024.10 | 0.35 | |
| Glove-sharedLabel=B3c, Tuning Status=tuned, Training Subset=20%M2024.10 | 0.35 | |
| GPT4.1 (Exp.2.1)Reasoning=Standard Prompting, Input=Generated Sentences, Output=Numeric2024.10 | 0.32 | |
| Exp.2.1Backbone=Qwen2024.10 | 0.31 | |
| Exp.4.1Backbone=Qwen2024.10 | 0.31 | |
| Glove-sharedtuning=not-tuned2024.10 | 0.3 | |
| Glove-shared not-tunedMethod ID=B3a2024.10 | 0.3 | |
| Glove-sharedLabel=B3a, Tuning Status=not-tuned2024.10 | 0.3 | |
| GSD20152024.10 | 0.29 | |
| Glove-sharedtuning=tuned2024.10 | 0.29 | |
| GSD2015Method ID=BG2024.10 | 0.29 | |
| Glove-shared tunedMethod ID=B3b2024.10 | 0.29 | |
| Exp.1.1Backbone=Qwen2024.10 | 0.29 | |
| GSD2015Label=BG2024.10 | 0.29 | |
| Glove-sharedLabel=B3b, Tuning Status=tuned2024.10 | 0.29 | |
| GPT4.1 (Exp.2.2)Reasoning=Standard Prompting, Input=Generated Sentences, Output=Categorical2024.10 | 0.26 | |
| Exp.4.2Backbone=Qwen2024.10 | 0.22 |