Open-ended Text Generation on Wikinews (test)
0.95Diversity (DIV)Typical Sampling
Evaluation Results
| Method | Links | |||
|---|---|---|---|---|
| Typical SamplingModel=GPT2-XL, Decoding Strategy=typical=0.952022.10 | 0.95 | 0.91 | 0.56 | |
| Typical SamplingModel=OPT-13B, Decoding Strategy=typical=0.952022.10 | 0.94 | 0.9 | 0.59 | |
| Contrastive DecodingModel=OPT-13B, Decoding Strategy=CD2022.10 | 0.94 | 0.94 | 0.69 | |
| Nucleus SamplingModel=GPT2-XL, Decoding Strategy=p=0.952022.10 | 0.94 | 0.9 | 0.6 | |
| Contrastive SearchModel=GPT2-XL, Decoding Strategy=CS2022.10 | 0.93 | 0.82 | 0.62 | |
| Nucleus SamplingModel=OPT-13B, Decoding Strategy=p=0.952022.10 | 0.92 | 0.92 | 0.62 | |
| Contrastive SearchModel=OPT-13B, Decoding Strategy=CS2022.10 | 0.92 | 0.87 | 0.59 | |
| Top-k SamplingModel=GPT2-XL, Decoding Strategy=k=502022.10 | 0.92 | 0.88 | 0.64 | |
| Contrastive DecodingModel=GPT2-XL, Decoding Strategy=CD2022.10 | 0.92 | 0.94 | 0.69 | |
| Top-k SamplingModel=OPT-13B, Decoding Strategy=k=502022.10 | 0.91 | 0.92 | 0.64 | |
| Greedy DecodingModel=OPT-13B, Decoding Strategy=max prob2022.10 | 0.08 | 0.3 | 0.65 | |
| Greedy DecodingModel=GPT2-XL, Decoding Strategy=max prob2022.10 | 0.04 | 0.14 | 0.65 |