Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
LLM (Chat & Instruction) AI research – Page 69 · SOTA2 Research
Back to Research
Research domain
LLM (Chat & Instruction)
Search tasks
Search
1 benchmarks · 1 papers
General Multi-modal Assistant Task
1 benchmarks · 1 papers
Clinical Intent Alignment
1 benchmarks · 1 papers
Creative Plot Generation
1 benchmarks · 1 papers
Long-horizon Dialogue Reasoning
1 benchmarks · 1 papers
Personalized memory reasoning
1 benchmarks · 1 papers
Multi-turn Chat Conversation
1 benchmarks · 1 papers
Logical Coherence
1 benchmarks · 1 papers
Multi-turn response generation
1 benchmarks · 1 papers
Analytical and personal anecdote writing
1 benchmarks · 1 papers
Spoken Task-Oriented Dialogue
1 benchmarks · 2 papers
Truthfulness and Informativeness
1 benchmarks · 1 papers
Value Alignment Evaluation
1 benchmarks · 1 papers
Agentic Presentation Generation
1 benchmarks · 1 papers
Value Consistency Evaluation
1 benchmarks · 1 papers
Author Response Generation
1 benchmarks · 2 papers
Language Model Unlearning
1 benchmarks · 1 papers
Pluralistic Alignment
1 benchmarks · 1 papers
Arabic Function Calling
1 benchmarks · 1 papers
Overton Alignment
1 benchmarks · 1 papers
Adversarial Prompt Detection
1 benchmarks · 1 papers
Autonomous research pipeline execution
1 benchmarks · 1 papers
Scattered Sequence Copying
1 benchmarks · 1 papers
Consistency Assessment of Generated Reference Points
1 benchmarks · 1 papers
Steerable Alignment
1 benchmarks · 1 papers
Memory-based Question Answering
1 benchmarks · 1 papers
Conversational Language Modeling
1 benchmarks · 1 papers
Steerable Value Alignment
1 benchmarks · 1 papers
Hallucination Interpretation
1 benchmarks · 1 papers
Expert Evaluation
1 benchmarks · 1 papers
Step-level Correctness Discrimination
1 benchmarks · 1 papers
Verbal Feedback Evaluation
1 benchmarks · 1 papers
Human Evaluation of Dialogue
Page 69 of 154
Previous
Next