Long-context Evaluation on RULER 4k (Average 13 tasks)
0.842ScoreVanilla
Evaluation Results
| Method | Links | |||
|---|---|---|---|---|
| VanillaAttention Mechanism=Full Attention, Model Size=8.1B, Training Context=32k2026.05 | 0.842 | — | — | |
| Self-Pruned KVAttention Mechanism=Self-Pruned KV, tau=0.5, Model Size=8.1B, Training Context=32k2026.05 | 0.84 | -0.2 | 0.188 |