Scientific-source retrieval on CheckThat! Task 1 (French) 2026 (dev)
77.72MRR@5With LLM judge
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| With LLM judgeStage=Final selective disagreement pipeline, LLM Judge Model=gpt-5.52026.05 | 77.72 | 0.0488 | |
| Plain rerankerStage=Reranking baseline2026.05 | 72.84 | — |