Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Agents & Tool Use AI research – Page 20 · SOTA2 Research
Back to Research
Research domain
Agents & Tool Use
Search tasks
Search
2 benchmarks · 2 papers
Agent Memory
2 benchmarks · 2 papers
Multi-Agent Collaboration
2 benchmarks · 2 papers
GUI Agent Task Execution
2 benchmarks · 2 papers
Constraint Recall
2 benchmarks · 2 papers
Lie
2 benchmarks · 1 papers
Web browsing agent security and task completion
2 benchmarks · 1 papers
Mobile Agent Interaction
2 benchmarks · 1 papers
Agentic Affordance Reasoning
2 benchmarks · 1 papers
Open-set assistance
2 benchmarks · 3 papers
OS Task
2 benchmarks · 1 papers
Shopping Agent Task
2 benchmarks · 1 papers
Offline Constrained RLHF
2 benchmarks · 2 papers
DB Task
2 benchmarks · 3 papers
Real-World Agent
2 benchmarks · 4 papers
Code Agent
2 benchmarks · 1 papers
Agent Planning and API Calling
2 benchmarks · 2 papers
NL-to-STL Translation
2 benchmarks · 2 papers
General Downstream Evaluation
2 benchmarks · 1 papers
Tool Extraction
2 benchmarks · 1 papers
Long-Horizon Search Intelligence
2 benchmarks · 2 papers
Secure LLM Agent Task Completion
2 benchmarks · 1 papers
Agent success prediction
2 benchmarks · 2 papers
Agentic Function Calling
2 benchmarks · 8 papers
Interactive Tool-Use Agent Performance
1 benchmarks · 1 papers
Financial Agent
1 benchmarks · 1 papers
Mobile Agent Decision-making
1 benchmarks · 1 papers
Agentic Instruction Following
1 benchmarks · 1 papers
End-to-end task execution
1 benchmarks · 1 papers
Multi-step web search
1 benchmarks · 1 papers
Medical Visual Question Answering / Tool-use
1 benchmarks · 1 papers
General AI Assistant Task Completion
1 benchmarks · 1 papers
Efficiency Analysis of Search Agents
Page 20 of 49
Previous
Next