Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
LLM (Chat & Instruction) AI research – Page 6 · SOTA2 Research
Back to Research
Research domain
LLM (Chat & Instruction)
Search tasks
Search
12 benchmarks · 1 papers
Few-shot example selection
12 benchmarks · 5 papers
Empathetic Dialogue
12 benchmarks · 4 papers
Memory
12 benchmarks · 2 papers
Step-wise Verification
12 benchmarks · 3 papers
LLM Training
12 benchmarks · 5 papers
Scenario Generation
12 benchmarks · 11 papers
Emotional Support Conversation
12 benchmarks · 1 papers
Safety and Informativeness Evaluation
12 benchmarks · 22 papers
Hallucination assessment
12 benchmarks · 4 papers
Next step prediction
12 benchmarks · 13 papers
Few-shot Learning
12 benchmarks · 1 papers
Pairwise Report Evaluation
12 benchmarks · 4 papers
Generative Task
12 benchmarks · 4 papers
Time to First Token (TTFT)
12 benchmarks · 1 papers
Self-Refinement
12 benchmarks · 6 papers
Creative Generation
12 benchmarks · 1 papers
Instruction Hierarchy Robustness
12 benchmarks · 5 papers
Medical Knowledge Evaluation
12 benchmarks · 8 papers
Empathetic Response Generation
12 benchmarks · 1 papers
Mobile Device Operation
12 benchmarks · 5 papers
Agent evaluation
12 benchmarks · 6 papers
LLM-as-a-judge evaluation
11 benchmarks · 5 papers
Visual Instruction Tuning
11 benchmarks · 9 papers
Needle-in-a-Haystack
11 benchmarks · 7 papers
Task Completion
11 benchmarks · 6 papers
Decoding
11 benchmarks · 2 papers
MLLM Unlearning
11 benchmarks · 8 papers
Preference Modeling
11 benchmarks · 1 papers
Unseen prompt generalization
11 benchmarks · 5 papers
Open-ended instruction following
11 benchmarks · 4 papers
Survey Generation
11 benchmarks · 4 papers
Context Compression
Page 6 of 154
Previous
Next