Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Code & Developer AI AI research – Page 37 · SOTA2 Research
Back to Research
Research domain
Code & Developer AI
Search tasks
Search
1 benchmarks · 1 papers
Program-aided math reasoning
1 benchmarks · 1 papers
Long-horizon Research Task Reproduction
1 benchmarks · 1 papers
Execution Accuracy
1 benchmarks · 1 papers
XML Compilation
1 benchmarks · 1 papers
Mobile kernel generation
1 benchmarks · 1 papers
Multi-turn Code Generation
1 benchmarks · 2 papers
Vulnerability Analysis and Reproduction
1 benchmarks · 1 papers
JAX
1 benchmarks · 1 papers
Mobile UI Automation
1 benchmarks · 1 papers
Mathematical Proof Construction
1 benchmarks · 1 papers
Disassembly
1 benchmarks · 1 papers
PPTX
1 benchmarks · 1 papers
NL to LTL translation
1 benchmarks · 1 papers
General Multitask Evaluation
1 benchmarks · 1 papers
Coding Collaboration
1 benchmarks · 1 papers
Multi-task Reasoning and Efficiency
1 benchmarks · 1 papers
Remote code execution (RCE) exploit generation
1 benchmarks · 1 papers
Unreferenced states & props detection
1 benchmarks · 1 papers
Generative Modeling Efficiency
1 benchmarks · 1 papers
Hardware-Aware Neural Architecture Search
1 benchmarks · 1 papers
Prop drilling detection
1 benchmarks · 1 papers
Layer Pruning
1 benchmarks · 1 papers
Test Writing
1 benchmarks · 1 papers
Malicious PowerShell Script Detection
1 benchmarks · 1 papers
API Call Prediction
1 benchmarks · 1 papers
Conditional auto-completion
1 benchmarks · 2 papers
Competitive Programming Agent Evaluation
1 benchmarks · 1 papers
Open-ended Computer Science Problem Solving
1 benchmarks · 1 papers
Capability Retention
1 benchmarks · 1 papers
Zero-day vulnerability detection
1 benchmarks · 1 papers
General-domain capability retention
1 benchmarks · 1 papers
Software Engineering Task Completion
Page 37 of 41
Previous
Next