Thematic fit estimation on Pado
0.59Spearman Correlation (ρ)Glove-shared
Evaluation Results
| Method | Links | |
|---|---|---|
| Glove-sharedtuning=tuned, data=20% subset2024.10 | 0.59 | |
| Glove-shared tuned with 20%MMethod ID=B3c2024.10 | 0.59 | |
| Glove-sharedLabel=B3c, Tuning Status=tuned, Training Subset=20%M2024.10 | 0.59 | |
| Step-by-Step PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.4.2, Input Form=Generated Sentences, Output Form=Categorical2024.10 | 0.59 | |
| Step-by-Step PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.4.1, Input Form=Generated Sentences, Output Form=Numeric2024.10 | 0.58 | |
| GSD20152024.10 | 0.53 | |
| ResRoFA-MT2024.10 | 0.53 | |
| Glove-sharedtuning=tuned2024.10 | 0.53 | |
| GSD2015Method ID=BG2024.10 | 0.53 | |
| ResRoFA-MTMethod ID=B12024.10 | 0.53 | |
| Glove-shared tunedMethod ID=B3b2024.10 | 0.53 | |
| GSD2015Label=BG2024.10 | 0.53 | |
| ResRoFA-MTLabel=B12024.10 | 0.53 | |
| Glove-sharedLabel=B3b, Tuning Status=tuned2024.10 | 0.53 | |
| NN RF2024.10 | 0.52 | |
| NN RFMethod ID=B02024.10 | 0.52 | |
| NN RFLabel=B02024.10 | 0.52 | |
| Glove-sharedtuning=not-tuned2024.10 | 0.49 | |
| Glove-shared not-tunedMethod ID=B3a2024.10 | 0.49 | |
| Glove-sharedLabel=B3a, Tuning Status=not-tuned2024.10 | 0.49 | |
| Step-by-Step PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.3.1, Input Form=Lemma Tuples, Output Form=Numeric2024.10 | 0.48 | |
| Step-by-Step PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.3.2, Input Form=Lemma Tuples, Output Form=Categorical2024.10 | 0.46 | |
| GPT4.1 (Exp.3.1)Reasoning=Step-by-Step Prompting, Input=Lemma Tuples, Output=Numeric2024.10 | 0.45 | |
| Simple PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.2.2, Input Form=Generated Sentences, Output Form=Categorical2024.10 | 0.45 | |
| 20% subset v22024.10 | 0.43 | |
| 20% subset v2Method ID=B22024.10 | 0.43 | |
| 20% subset v2Label=B22024.10 | 0.43 | |
| GPT4.1 (Exp.3.2)Reasoning=Step-by-Step Prompting, Input=Lemma Tuples, Output=Categorical2024.10 | 0.42 | |
| GPT4.1 (Exp.4.1)Reasoning=Step-by-Step Prompting, Input=Generated Sentences, Output=Numeric2024.10 | 0.42 | |
| GPT4.1 (Exp.4.2)Reasoning=Step-by-Step Prompting, Input=Generated Sentences, Output=Categorical2024.10 | 0.42 | |
| Simple PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.2.1, Input Form=Generated Sentences, Output Form=Numeric2024.10 | 0.42 | |
| GPT4.1 (Exp.1.1)Reasoning=Standard Prompting, Input=Lemma Tuples, Output=Numeric2024.10 | 0.4 | |
| GPT4.1 (Exp.1.2)Reasoning=Standard Prompting, Input=Lemma Tuples, Output=Categorical2024.10 | 0.4 | |
| Simple PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.1.2, Input Form=Lemma Tuples, Output Form=Categorical2024.10 | 0.4 | |
| Exp.4.2Backbone=Qwen2024.10 | 0.37 | |
| Exp.3.2Backbone=Qwen2024.10 | 0.36 | |
| Simple PromptingBackbone=GPT4-Turbo, Experiment ID=Exp.1.1, Input Form=Lemma Tuples, Output Form=Numeric2024.10 | 0.35 | |
| GPT4.1 (Exp.2.1)Reasoning=Standard Prompting, Input=Generated Sentences, Output=Numeric2024.10 | 0.33 | |
| GPT4.1 (Exp.2.2)Reasoning=Standard Prompting, Input=Generated Sentences, Output=Categorical2024.10 | 0.33 | |
| Exp.4.1Backbone=Qwen2024.10 | 0.33 | |
| Exp.3.1Backbone=Qwen2024.10 | 0.32 | |
| Exp.2.1Backbone=Qwen2024.10 | 0.31 | |
| Exp.1.1Backbone=Qwen2024.10 | 0.28 | |
| Exp.2.2Backbone=Qwen2024.10 | 0.22 | |
| Exp.1.2Backbone=Qwen2024.10 | 0.19 |