Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Agents & Tool Use AI research – Page 6 · SOTA2 Research
Back to Research
Research domain
Agents & Tool Use
Search tasks
Search
5 benchmarks · 5 papers
Web Agent
5 benchmarks · 1 papers
Tool-Augmented Planning
5 benchmarks · 1 papers
End-to-end Tool-use
5 benchmarks · 1 papers
Agent Orchestration
5 benchmarks · 1 papers
Online Agent-Task Matching
5 benchmarks · 1 papers
Speech Function Calling
5 benchmarks · 1 papers
Agent Security (Indirect Prompt Injection)
5 benchmarks · 1 papers
Red-teaming tool-calling agents
5 benchmarks · 3 papers
Multimodal Agent Task
5 benchmarks · 4 papers
Agentic Tasks
5 benchmarks · 3 papers
Skill Selection
5 benchmarks · 1 papers
Tool-augmented Question Answering
5 benchmarks · 2 papers
Multi-agent Debate
5 benchmarks · 2 papers
Agent & Alignment
5 benchmarks · 1 papers
Multi-agent Selection (Pairwise Resolution)
5 benchmarks · 3 papers
Long-horizon Agent Performance
5 benchmarks · 1 papers
Multi-agent Observation Sharing
5 benchmarks · 1 papers
Autonomous Agent Task Execution under Injection
5 benchmarks · 5 papers
Autonomous Web Navigation
5 benchmarks · 3 papers
Agent Reasoning
5 benchmarks · 5 papers
Long-context memory management
5 benchmarks · 2 papers
Asynchronous planning
5 benchmarks · 1 papers
MTG Gameplay
5 benchmarks · 2 papers
Short-Term Planning
5 benchmarks · 2 papers
Multi-Agent Strategic Reasoning
5 benchmarks · 2 papers
Search-based planning
5 benchmarks · 6 papers
Tool Use Evaluation
4 benchmarks · 2 papers
Card Games
4 benchmarks · 1 papers
Function Invocation
4 benchmarks · 9 papers
Web Navigation Question Answering
4 benchmarks · 2 papers
Sequential Planning
4 benchmarks · 4 papers
GUI Agent Task Success
Page 6 of 49
Previous
Next