Profile Adherence on Curated Reddit dialogue dataset (test)
54.9Macro F1 MeanMoLA (Ours)
Evaluation Results
| Method | Links | ||||
|---|---|---|---|---|---|
| MoLA (Ours)Architecture=Feature-aware Mixture of Experts (MoE)2026.01 | 54.9 | 0.125 | 43 | 74 | |
| SFTBase Model=Llama-3-8B-Instruct, Training=Standard next-token prediction2026.01 | 51.5 | 0.16 | 36 | 76 | |
| Contrastive LearningBase Model=Llama-3-8B-Instruct, Training=SFT with disentanglement loss2026.01 | 48.4 | 0.178 | 34 | 85 | |
| GPT-4.1-miniMode=Zero-shot2026.01 | 30.1 | 0.131 | 16 | 58 | |
| Qwen-2.5-14B-InstructMode=Zero-shot2026.01 | 28.4 | 0.095 | 15 | 47 | |
| Llama-3-8B-InstructMode=Zero-shot2026.01 | 25.9 | 0.148 | 11 | 58 |