Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Agents & Tool Use AI research – Page 32 · SOTA2 Research
Back to Research
Research domain
Agents & Tool Use
Search tasks
Search
1 benchmarks · 1 papers
Interactive step-by-step task guidance
1 benchmarks · 1 papers
Agent Task Success
1 benchmarks · 1 papers
Multi-agent systems coordination
1 benchmarks · 1 papers
LLM agent alignment evaluation
1 benchmarks · 2 papers
Web security task completion
1 benchmarks · 1 papers
Reasoning and Decision-making
1 benchmarks · 1 papers
Web Search Agent
1 benchmarks · 1 papers
Reasoning and Planning
1 benchmarks · 1 papers
Tool-use and Complex Reasoning
1 benchmarks · 1 papers
TSG Execution
1 benchmarks · 1 papers
Web Navigation QA
1 benchmarks · 1 papers
Expert Assignment Recovery
1 benchmarks · 1 papers
Online Web Navigation
1 benchmarks · 1 papers
Agentic Reasoning in Quantum Chemistry
1 benchmarks · 1 papers
Autonomous Agent Problem Solving
1 benchmarks · 2 papers
Mobile GUI Agent Decision Making
1 benchmarks · 1 papers
Web Browsing and Navigation (Chinese)
1 benchmarks · 1 papers
Collaborative decision-making
1 benchmarks · 1 papers
Multi-agent System Safety and Welfare Evaluation
1 benchmarks · 1 papers
Tool Selection Quality
1 benchmarks · 1 papers
Multi-Agent Edge Computing Orchestration
1 benchmarks · 1 papers
Agent Memory Question Answering
1 benchmarks · 1 papers
Tool orchestration
1 benchmarks · 1 papers
API Calling
1 benchmarks · 1 papers
Workflow Planning
1 benchmarks · 1 papers
Atomic Task Execution
1 benchmarks · 1 papers
Atomic Task Success
1 benchmarks · 1 papers
Long-Horizon Tool Execution
1 benchmarks · 1 papers
Multi-agent interaction and social reasoning
1 benchmarks · 1 papers
Multi-agent research collaboration
1 benchmarks · 1 papers
Freight Negotiation
1 benchmarks · 1 papers
Guide target element selection
Page 32 of 49
Previous
Next