Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Agents & Tool Use AI research – Page 38 · SOTA2 Research
Back to Research
Research domain
Agents & Tool Use
Search tasks
Search
1 benchmarks · 1 papers
Task 4-1 (S)
1 benchmarks · 1 papers
Agentic Text Synthesis
1 benchmarks · 1 papers
Tool Learning under Instruction with Missing Key Information
1 benchmarks · 1 papers
Multi-agent cooperation sustainability simulation
1 benchmarks · 1 papers
Tool Learning under Instruction with Multiple Requests
1 benchmarks · 1 papers
Template Following
1 benchmarks · 1 papers
Tool Learning under Instruction with Error
1 benchmarks · 1 papers
Task 6-1 (M)
1 benchmarks · 1 papers
UI Navigation / Task Completion
1 benchmarks · 1 papers
Environment-Intensive Task Generation
1 benchmarks · 1 papers
Agentic (multi-turn) evaluation
1 benchmarks · 1 papers
Tool Learning under Instructions Beyond Tool Capabilities
1 benchmarks · 1 papers
Behavioral Rule-Collision Resolution
1 benchmarks · 1 papers
Task 8-2 (S)
1 benchmarks · 1 papers
Agent Framework Feature Comparison
1 benchmarks · 1 papers
Knowledge Graph navigation
1 benchmarks · 1 papers
Task 9-2 (L)
1 benchmarks · 1 papers
Web Browsing/Site navigation
1 benchmarks · 1 papers
Mobile GUI Agents
1 benchmarks · 1 papers
Biological Tool Use
1 benchmarks · 1 papers
Desktop task completion
1 benchmarks · 1 papers
Web Browsing and Comparison
1 benchmarks · 1 papers
Web Browsing Automation
1 benchmarks · 1 papers
Proactive Autonomy
1 benchmarks · 3 papers
Agentic Workflow Success
1 benchmarks · 1 papers
Multitask Thai Language Evaluation
1 benchmarks · 1 papers
Meteorological diagnosis agentic workflow
1 benchmarks · 1 papers
Long-horizon item crafting
1 benchmarks · 1 papers
Bimanual tool-use (Scoop into bowl)
1 benchmarks · 1 papers
Single-turn Tool Use
1 benchmarks · 1 papers
Agent Adaptation
1 benchmarks · 1 papers
Due Diligence
Page 38 of 49
Previous
Next