Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| Astrid Orcesi, Romaric Audigier, Fritz Poka Toukam |
| 2022 |
| arxiv 2201.02396 |
| InstaFlow: One Step is Enough for High-Quality Diffusion-Based Text-to-Image Generation | Xingchao Liu, Xiwen Zhang, Jianzhu Ma | 2023 | arxiv 2309.06380 |
|---|
| ActionHub: A Large-scale Action Video Description Dataset for Zero-shot Action Recognition | Jiaming Zhou, Junwei Liang, Kun-Yu Lin | 2024 | arxiv 2401.11654 |
|---|
| DEYOv3: DETR with YOLO for Real-time Object Detection | Haodong Ouyang | 2023 | arxiv 2309.11851 |
|---|
| SyNet: An Ensemble Network for Object Detection in UAV Images | Berat Mert Albaba, Sedat Ozer | 2020 | arxiv 2012.12991 |
|---|
| Exploration of an End-to-End Automatic Number-plate Recognition neural network for Indian datasets | Sai Sirisha Nadiminti, Pranav Kant Gaur, Abhilash Bhardwaj | 2022 | arxiv 2207.06657 |
|---|
| Precise Single-stage Detector | Aisha Chandio, Gong Gui, Teerath Kumar | 2022 | arxiv 2210.04252 |
|---|
| Focal Modulation Networks | Jianwei Yang, Chunyuan Li, Xiyang Dai | 2022 | arxiv 2203.11926 |
|---|
| Auto-Context R-CNN | Bo Li, Tianfu Wu, Lun Zhang | 2018 | arxiv 1807.02842 |
|---|
| The Urban Vision Hackathon Dataset and Models: Towards Image Annotations and Accurate Vision Models for Indian Traffic | Akash Sharma, Chinmay Mhatre, Sankalp Gawali | 2025 | arxiv 2511.02563 |
|---|
| Are Neuro-Inspired Multi-Modal Vision-Language Models Resilient to Membership Inference Privacy Leakage? | David Amebley, Sayanton Dibbo | 2025 | arxiv 2511.20710 |
|---|