Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
LLM (Chat & Instruction) AI research – Page 7 · SOTA2 Research
Back to Research
Research domain
LLM (Chat & Instruction)
Search tasks
Search
11 benchmarks · 9 papers
Multimodal Understanding and Reasoning
11 benchmarks · 5 papers
Supervised Fine-tuning
11 benchmarks · 8 papers
Memory-augmented Question Answering
11 benchmarks · 4 papers
Survey Generation
11 benchmarks · 1 papers
Unseen prompt generalization
11 benchmarks · 8 papers
Preference Modeling
11 benchmarks · 1 papers
Instruction-driven 3D layout generation
11 benchmarks · 6 papers
Decoding
10 benchmarks · 1 papers
Autonomous LLM Fine-tuning
10 benchmarks · 5 papers
Safety & Helpfulness Evaluation
10 benchmarks · 2 papers
Diversity
10 benchmarks · 1 papers
MCQ Classification
10 benchmarks · 1 papers
Harmful Request Compliance
10 benchmarks · 18 papers
Graduate-Level Reasoning
10 benchmarks · 4 papers
Knowledge Conflict Resolution
10 benchmarks · 2 papers
Multimodal Preference Evaluation
10 benchmarks · 5 papers
Personalized Response Generation
10 benchmarks · 43 papers
Multi-turn Dialogue Evaluation
10 benchmarks · 4 papers
Steering
10 benchmarks · 11 papers
General Performance
10 benchmarks · 6 papers
Rationale Generation
10 benchmarks · 1 papers
Long-form Answer Generation
10 benchmarks · 8 papers
Knowledge & Reasoning
10 benchmarks · 12 papers
Travel Planning
10 benchmarks · 8 papers
Arithmetic
10 benchmarks · 4 papers
Prompt Recovery
10 benchmarks · 7 papers
Reasoning Quality Evaluation
10 benchmarks · 4 papers
Bargaining
10 benchmarks · 2 papers
Controllable Language Generation
10 benchmarks · 3 papers
Skill Routing
10 benchmarks · 1 papers
Instruction-following clustering
10 benchmarks · 7 papers
Clinical Reasoning
Page 7 of 154
Previous
Next