Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
LLM (Chat & Instruction) AI research – Page 115 · SOTA2 Research
Back to Research
Research domain
LLM (Chat & Instruction)
Search tasks
Search
1 benchmarks · 1 papers
Plan-Grounded Answer Generation
1 benchmarks · 1 papers
Secret loyalty evaluation
1 benchmarks · 1 papers
Personalization Utility
1 benchmarks · 1 papers
Forward Continual Learning
1 benchmarks · 1 papers
Text Generation Throughput
1 benchmarks · 1 papers
Single-turn data-to-logic tasks
1 benchmarks · 1 papers
Instruction following verification
1 benchmarks · 5 papers
Harmful Prompt Refusal
1 benchmarks · 1 papers
Step-level error discrimination
1 benchmarks · 1 papers
Aggregate Multi-task Evaluation
1 benchmarks · 1 papers
Tool-Need Prediction
1 benchmarks · 1 papers
Tool-Risk Prediction
1 benchmarks · 1 papers
Tool-call Monitoring
1 benchmarks · 1 papers
Tool-use runtime evaluation
1 benchmarks · 1 papers
Long-context understanding and resolution
1 benchmarks · 1 papers
Downstream skill execution
1 benchmarks · 1 papers
Multimodal Cognition Evaluation
1 benchmarks · 1 papers
Health Reasoning
1 benchmarks · 1 papers
Multi-turn planning
1 benchmarks · 1 papers
Continual Knowledge Incorporation
1 benchmarks · 1 papers
Fact
1 benchmarks · 1 papers
Query routing and tool-calling accuracy evaluation
1 benchmarks · 1 papers
LLM-Assisted Scoring
1 benchmarks · 1 papers
Personalized LLM response generation
1 benchmarks · 1 papers
Downstream Generation Impact (LLaVA-1.5)
1 benchmarks · 1 papers
Reasoning Task
1 benchmarks · 1 papers
Downstream Generation Impact (LLaVA-1.6)
1 benchmarks · 1 papers
Downstream Generation Impact (InstructBLIP)
1 benchmarks · 1 papers
Downstream Generation Impact (Qwen2.5-VL)
1 benchmarks · 1 papers
Hybrid Retrieval-Augmented Generation
1 benchmarks · 1 papers
Downstream Generation Impact (GeoChat)
1 benchmarks · 1 papers
Research Assistant
Page 115 of 154
Previous
Next