Scientific-source retrieval on CheckThat! Task 1 (English) 2026 (dev)
0.717MRR@5With LLM judge
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| With LLM judgeStage=Final selective disagreement pipeline, LLM Judge Model=gpt-5.52026.05 | 0.717 | 0.0464 | |
| Plain rerankerStage=Reranking baseline2026.05 | 0.6706 | — |