Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
LLM (Chat & Instruction) AI research – Page 50 · SOTA2 Research
Back to Research
Research domain
LLM (Chat & Instruction)
Search tasks
Search
2 benchmarks · 1 papers
Sales interaction performance
2 benchmarks · 1 papers
Agreement with outcome human labels
2 benchmarks · 1 papers
Pun Explanation
2 benchmarks · 1 papers
MLLM Personalization
2 benchmarks · 2 papers
Knowledge / General QA
2 benchmarks · 1 papers
Citation Coverage Evaluation
2 benchmarks · 2 papers
Video generation reasoning
2 benchmarks · 1 papers
Agentic medical interaction
2 benchmarks · 2 papers
General Reasoning Average
2 benchmarks · 2 papers
Proof writing
2 benchmarks · 2 papers
Multi-Session Dialogue Generation
2 benchmarks · 1 papers
Generalization to Unseen Preferences
2 benchmarks · 1 papers
Reward Model Controllability
2 benchmarks · 1 papers
Best-of-N Selection
2 benchmarks · 2 papers
Math Question Answering
2 benchmarks · 1 papers
Cross-Lingual Knowledge Alignment
2 benchmarks · 1 papers
Preference Discrimination
2 benchmarks · 1 papers
Expert Preference Pairwise
2 benchmarks · 1 papers
Risk evaluation in caregiving responses
2 benchmarks · 1 papers
Chit-chat conversation evaluation correlation
2 benchmarks · 1 papers
Ordinal Preference Alignment
2 benchmarks · 1 papers
Reference-free Conversation Evaluation
2 benchmarks · 1 papers
Human Pairwise Comparison
2 benchmarks · 1 papers
Reward-wise QA fairness and alignment
2 benchmarks · 1 papers
Multi-hop Retrieval-Augmented Generation
2 benchmarks · 1 papers
Friend Recommendation
2 benchmarks · 1 papers
Price Negotiation
2 benchmarks · 1 papers
Scientific Verification
2 benchmarks · 1 papers
LLM steering evaluation
2 benchmarks · 2 papers
Language Understanding and Question Answering
2 benchmarks · 1 papers
Evaluating Context Influence and Input Regurgitation
2 benchmarks · 1 papers
Dialogue Annotation
Page 50 of 154
Previous
Next