Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Agents & Tool Use AI research – Page 16 · SOTA2 Research
Back to Research
Research domain
Agents & Tool Use
Search tasks
Search
2 benchmarks · 1 papers
Single-turn Tool Calling
2 benchmarks · 1 papers
Tool-agent-user interaction
2 benchmarks · 2 papers
General AI Assistants Evaluation
2 benchmarks · 1 papers
Tool Sequence Recommendation
2 benchmarks · 1 papers
Personalized Agentic Social Support
2 benchmarks · 1 papers
Tool Use Accuracy
2 benchmarks · 1 papers
Tool Response Generation
2 benchmarks · 1 papers
Multi-agent Semantic Object Navigation
2 benchmarks · 2 papers
Tool-use Question Answering
2 benchmarks · 1 papers
Spreadsheet Understanding and Manipulation
2 benchmarks · 1 papers
Multi-Agent Task Scheduling
2 benchmarks · 2 papers
Role-playing agent evaluation
2 benchmarks · 1 papers
Price Negotiation
2 benchmarks · 1 papers
Slide Automation
2 benchmarks · 1 papers
Safety and Action Blocking
2 benchmarks · 1 papers
Multi-step reasoning and knowledge retrieval
2 benchmarks · 3 papers
GUI Task Completion
2 benchmarks · 1 papers
Web Navigation and Automation
2 benchmarks · 1 papers
Single-agent contract design
2 benchmarks · 1 papers
Steering Validation
2 benchmarks · 1 papers
Multi-agent contract design
2 benchmarks · 1 papers
Enterprise interface interaction
2 benchmarks · 1 papers
Enterprise interface task completion
2 benchmarks · 2 papers
Agent Routing
2 benchmarks · 1 papers
Knowledge-to-action question answering
2 benchmarks · 2 papers
General Agent Capability
2 benchmarks · 1 papers
Driving Hazard Explanation and Action Generation
2 benchmarks · 1 papers
device control
2 benchmarks · 1 papers
GUI Agent Automation
2 benchmarks · 1 papers
Agentic medical interaction
2 benchmarks · 5 papers
Mobile UI Control
2 benchmarks · 3 papers
Terminal-based task execution
Page 16 of 49
Previous
Next