Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
LLM (Chat & Instruction) AI research – Page 49 · SOTA2 Research
Back to Research
Research domain
LLM (Chat & Instruction)
Search tasks
Search
2 benchmarks · 1 papers
Agreement with process human labels
2 benchmarks · 1 papers
Construct Validity Verification
2 benchmarks · 1 papers
Knowledge-grounded Generation
2 benchmarks · 1 papers
Pun Explanation
2 benchmarks · 2 papers
Text-to-Image Preference Alignment
2 benchmarks · 1 papers
Personality Recovery
2 benchmarks · 1 papers
Knowledge Transfer
2 benchmarks · 1 papers
Scientific Verification
2 benchmarks · 1 papers
Concept Forgetting
2 benchmarks · 1 papers
Document Continuation
2 benchmarks · 8 papers
Dialogue Emotion Detection
2 benchmarks · 1 papers
Dialogue Strategy Alignment
2 benchmarks · 2 papers
General Capability Preservation
2 benchmarks · 2 papers
Long-context Multi-modal Understanding
2 benchmarks · 2 papers
Graduate-level Q&A
2 benchmarks · 1 papers
Locality
2 benchmarks · 1 papers
Motivational Interviewing Response Generation
2 benchmarks · 1 papers
Human forward simulatability
2 benchmarks · 1 papers
Out-of-scope refusal
2 benchmarks · 2 papers
General Reasoning Average
2 benchmarks · 2 papers
Proof writing
2 benchmarks · 1 papers
Reward Model Controllability
2 benchmarks · 1 papers
Progress Reasoning
2 benchmarks · 2 papers
Safety-Utility Trade-off
2 benchmarks · 1 papers
Ordinal Preference Alignment
2 benchmarks · 1 papers
Conversational Assistant
2 benchmarks · 1 papers
Agentic medical interaction
2 benchmarks · 1 papers
Generalization to Unseen Preferences
2 benchmarks · 1 papers
Reward-wise QA fairness and alignment
2 benchmarks · 1 papers
Bilingual Mathematical Reasoning
2 benchmarks · 1 papers
Reference-free Conversation Evaluation
2 benchmarks · 1 papers
Chit-chat conversation evaluation correlation
Page 49 of 154
Previous
Next