LLM Preference Alignment Evaluation on Stack-Exchange
66Preference Score (spec vs ctl)Spec Learning
Evaluation Results
| Method | Links | |||
|---|---|---|---|---|
| Spec LearningProposer (P)=Gemma 4 31B, Judge Model=GLM-5.12026.06 | 66 | 83 | 71 |
| Method | Links | |||
|---|---|---|---|---|
| Spec LearningProposer (P)=Gemma 4 31B, Judge Model=GLM-5.12026.06 | 66 | 83 | 71 |