Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Agents & Tool Use AI research – Page 19 · SOTA2 Research
Back to Research
Research domain
Agents & Tool Use
Search tasks
Search
2 benchmarks · 5 papers
Mobile UI Control
2 benchmarks · 1 papers
Form-filling
2 benchmarks · 1 papers
Regions of Interest Discovery
2 benchmarks · 2 papers
Interactive tool-use
2 benchmarks · 1 papers
Enterprise interface task completion
2 benchmarks · 1 papers
Agentic medical interaction
2 benchmarks · 2 papers
Web Task
2 benchmarks · 1 papers
Agent & OpenClaw
2 benchmarks · 1 papers
Multi-agent contract design approximation
2 benchmarks · 2 papers
Multi-Agent Collaboration
2 benchmarks · 2 papers
GUI Agent Task Execution
2 benchmarks · 1 papers
Mobile Agent Interaction
2 benchmarks · 1 papers
Agentic Affordance Reasoning
2 benchmarks · 1 papers
Enterprise interface interaction
2 benchmarks · 1 papers
Agent success prediction
2 benchmarks · 1 papers
Agent Planning and API Calling
2 benchmarks · 1 papers
Offline Constrained RLHF
2 benchmarks · 2 papers
Framework Capability Comparison
2 benchmarks · 1 papers
Privacy-preserving Agent Interaction
2 benchmarks · 1 papers
Agricultural Agent Planning and Execution
2 benchmarks · 5 papers
LLM Agent Evaluation
2 benchmarks · 2 papers
Agent Memory
2 benchmarks · 2 papers
NL-to-STL Translation
2 benchmarks · 1 papers
Long-Horizon Search Intelligence
2 benchmarks · 2 papers
Constraint Recall
2 benchmarks · 2 papers
Lie
2 benchmarks · 1 papers
Web browsing agent security and task completion
2 benchmarks · 1 papers
Open-set assistance
2 benchmarks · 2 papers
Secure LLM Agent Task Completion
2 benchmarks · 1 papers
General AI Assistant Task Execution
2 benchmarks · 3 papers
OS Task
2 benchmarks · 2 papers
LLM Agent Security Evaluation
Page 19 of 49
Previous
Next