Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
LLM (Chat & Instruction) AI research – Page 127 · SOTA2 Research
Back to Research
Research domain
LLM (Chat & Instruction)
Search tasks
Search
1 benchmarks · 1 papers
Soft constrained reward optimization
1 benchmarks · 1 papers
Logic Puzzle
1 benchmarks · 1 papers
LaMP-5 Personalization
1 benchmarks · 1 papers
MK
1 benchmarks · 1 papers
Language Understanding (Law)
1 benchmarks · 1 papers
Math Find
1 benchmarks · 1 papers
Multi-turn conversational and instruction-following
1 benchmarks · 1 papers
Generative Operational Planning
1 benchmarks · 1 papers
Interactional pluralism evaluation
1 benchmarks · 1 papers
Knowledge-based Agent Reasoning
1 benchmarks · 2 papers
General speculative decoding performance
1 benchmarks · 1 papers
Pairwise Human Evaluation
1 benchmarks · 1 papers
Personal Assistant Agent Performance
1 benchmarks · 1 papers
Voice Chatting
1 benchmarks · 1 papers
Judge Performance
1 benchmarks · 1 papers
Text-based embodied AI
1 benchmarks · 1 papers
Cross-shot consistency
1 benchmarks · 1 papers
Overrefusal Detection
1 benchmarks · 1 papers
Compound LLM Collaboration
1 benchmarks · 1 papers
Intra-shot prompt-following alignment
1 benchmarks · 1 papers
Compound AI System Alignment
1 benchmarks · 1 papers
Persona-grounded Dialogue
1 benchmarks · 1 papers
Compound LLM Collaboration System
1 benchmarks · 1 papers
AI Interruption Detection
1 benchmarks · 1 papers
HR Scenario Completion and Governance
1 benchmarks · 1 papers
Reasoning and Math Problem Solving
1 benchmarks · 1 papers
Citation
1 benchmarks · 1 papers
Zero-shot Task Accuracy
1 benchmarks · 1 papers
LLM Behavior
1 benchmarks · 1 papers
Large Language Model Pre-training
1 benchmarks · 1 papers
Kernel-level Attention Speed and Memory Analysis
1 benchmarks · 1 papers
RL Training
Page 127 of 154
Previous
Next