Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Agents & Tool Use AI research – Page 30 · SOTA2 Research
Back to Research
Research domain
Agents & Tool Use
Search tasks
Search
1 benchmarks · 1 papers
Interactive Science Reasoning
1 benchmarks · 1 papers
Agentic Person Search
1 benchmarks · 1 papers
Adversarial Attack on Go Agents
1 benchmarks · 1 papers
Agent Task-solving
1 benchmarks · 1 papers
Agentic Person Search (Spatial Reasoning)
1 benchmarks · 1 papers
Multi-step Scientific Tool-use
1 benchmarks · 1 papers
Agentic Person Search (Temporal Reasoning)
1 benchmarks · 1 papers
General Problem Solving
1 benchmarks · 1 papers
Task Planning and Allocation
1 benchmarks · 1 papers
GIS agentic spatial analysis
1 benchmarks · 1 papers
Robustness of Tool-use
1 benchmarks · 1 papers
Tool-Using Agent Safety
1 benchmarks · 1 papers
Tool Call Repair
1 benchmarks · 1 papers
Task Execution Simulation
1 benchmarks · 1 papers
Long-horizon GUI Interaction
1 benchmarks · 1 papers
GUI Operation
1 benchmarks · 1 papers
Agentic Planning
1 benchmarks · 1 papers
Multi-agent P2P energy trading optimization
1 benchmarks · 1 papers
General Assistant Reasoning
1 benchmarks · 1 papers
Multi-agent social interaction evaluation
1 benchmarks · 1 papers
Single Task Processing
1 benchmarks · 1 papers
Social Agent Evaluation
1 benchmarks · 1 papers
Contextual Planning
1 benchmarks · 2 papers
Tool usage simulation
1 benchmarks · 1 papers
GUI agent task planning
1 benchmarks · 1 papers
Proactive Seeking
1 benchmarks · 1 papers
LLM Agent Security
1 benchmarks · 1 papers
Result Feedback
1 benchmarks · 1 papers
Agent Core Capabilities Overall
1 benchmarks · 1 papers
Visual Agent Task
1 benchmarks · 1 papers
Multi-agent attack success evaluation
1 benchmarks · 2 papers
Web-based Agent QA
Page 30 of 49
Previous
Next