Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
LLM (Chat & Instruction) AI research – Page 37 · SOTA2 Research
Back to Research
Research domain
LLM (Chat & Instruction)
Search tasks
Search
3 benchmarks · 1 papers
multi-branch story generation
3 benchmarks · 1 papers
LLM Generation Efficiency
3 benchmarks · 1 papers
Reasoning failure prediction
3 benchmarks · 1 papers
Fine-tuning for knowledge acquisition and abstention preservation
3 benchmarks · 1 papers
Qualitative User Preference Evaluation
3 benchmarks · 3 papers
Psychological Counseling Dialogue Evaluation
3 benchmarks · 1 papers
Multi-agent Cognitive Orchestration
3 benchmarks · 1 papers
Attributed Text Generation
3 benchmarks · 2 papers
Story Generation Evaluation
3 benchmarks · 2 papers
Pairwise LLM Judging
3 benchmarks · 1 papers
Preference Labeling
3 benchmarks · 3 papers
Instruction-based Video Editing
3 benchmarks · 1 papers
Medical Instruction Following
3 benchmarks · 1 papers
Sycophancy Assessment
3 benchmarks · 3 papers
Instruction Following and Safety Evaluation
3 benchmarks · 5 papers
Multistep Soft Reasoning
3 benchmarks · 1 papers
Bayesian Assessment of Sycophancy
3 benchmarks · 2 papers
Competitive Mathematical Reasoning
3 benchmarks · 3 papers
Meeting Planning
3 benchmarks · 3 papers
Multi-Choice
3 benchmarks · 2 papers
Open-domain Conversation
3 benchmarks · 1 papers
Linear Concept Accessibility and Steering
3 benchmarks · 1 papers
Follow-up Questioning Consistency
3 benchmarks · 2 papers
Attention Operator Throughput
3 benchmarks · 1 papers
Policy Question Answering
3 benchmarks · 3 papers
Prefill
3 benchmarks · 1 papers
Content Generation (Micro Alignment)
3 benchmarks · 1 papers
Instruction Execution
3 benchmarks · 3 papers
High-Level Reasoning
3 benchmarks · 1 papers
Behavior Prediction (Micro Alignment)
3 benchmarks · 1 papers
Tool-use Inference
3 benchmarks · 1 papers
Goal-relevance Evaluation
Page 37 of 154
Previous
Next