Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
LLM (Chat & Instruction) AI research – Page 2 · SOTA2 Research
Back to Research
Research domain
LLM (Chat & Instruction)
Search tasks
Search
53 benchmarks · 24 papers
Dialogue Response Generation
49 benchmarks · 43 papers
Response Generation
49 benchmarks · 22 papers
LLM Inference
48 benchmarks · 34 papers
Conversational Question Answering
47 benchmarks · 27 papers
Instruction-based Image Editing
47 benchmarks · 12 papers
Model Routing
47 benchmarks · 85 papers
General Knowledge
42 benchmarks · 26 papers
LLM Alignment
41 benchmarks · 33 papers
Factuality Evaluation
41 benchmarks · 35 papers
Creative Writing
38 benchmarks · 61 papers
Long-context language modeling
38 benchmarks · 33 papers
Complex Reasoning
38 benchmarks · 44 papers
Long-context evaluation
38 benchmarks · 28 papers
Long-form Question Answering
37 benchmarks · 20 papers
Dialogue Evaluation
37 benchmarks · 26 papers
Jailbreak
37 benchmarks · 51 papers
Zero-shot Reasoning
37 benchmarks · 46 papers
Math Word Problem Solving
34 benchmarks · 14 papers
Conversational Recommendation
34 benchmarks · 30 papers
Human Preference Evaluation
33 benchmarks · 30 papers
Human Evaluation
33 benchmarks · 16 papers
Preference Prediction
33 benchmarks · 21 papers
STEM Reasoning
32 benchmarks · 2 papers
Abductive Explanation Generation
32 benchmarks · 5 papers
LLM Jailbreaking
32 benchmarks · 13 papers
Prompt Optimization
32 benchmarks · 24 papers
Alignment
31 benchmarks · 25 papers
Large Language Model Evaluation
30 benchmarks · 2 papers
Detection of LLM generated text
30 benchmarks · 14 papers
Personalization
30 benchmarks · 18 papers
Instruction Tuning
30 benchmarks · 31 papers
Multi-turn dialogue
Page 2 of 154
Previous
Next