Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning | Taewhan Kim, Soeun Lee, Si-Woo Kim | 2024 | arxiv 2412.19289 |
|---|
| Recurrent Diffusion for Large-Scale Parameter Generation | Kai Wang, Dongwen Tang, Wangbo Zhao | 2025 | arxiv 2501.11587 |
|---|
| DeepSeek on a Trip: Inducing Targeted Visual Hallucinations via Representation Vulnerabilities | Chashi Mahiul Islam, Samuel Jacob Chacko, Preston Horne | 2025 | arxiv 2502.07905 |
|---|
| Multi-label out-of-distribution detection via evidential learning | Eduardo Aguilar, Bogdan Raducanu, Petia Radeva | 2025 | arxiv 2502.18224 |
|---|
| TIGeR: Unifying Text-to-Image Generation and Retrieval with Large Multimodal Models | Leigang Qu, Haochuan Li, Tan Wang | 2024 | arxiv 2406.05814 |
|---|
| Specifying What You Know or Not for Multi-Label Class-Incremental Learning | Aoting Zhang, Dongbao Yang, Chang Liu | 2025 | arxiv 2503.17017 |
|---|
| Beyond Language: Learning Commonsense from Images for Reasoning | Wanqing Cui, Yanyan Lan, Liang Pang | 2020 | arxiv 2010.05001 |
|---|
| FourierNet: Compact mask representation for instance segmentation using differentiable shape decoders | Hamd ul Moqeet Riaz, Nuri Benbarka, Andreas Zell | 2020 | arxiv 2002.02709 |
|---|
| Shooting Labels: 3D Semantic Labeling by Virtual Reality | Pierluigi Zama Ramirez, Claudio Paternesi, Luca De Luigi | 2019 | arxiv 1910.05021 |
|---|
| Learning to Generate Content-Aware Dynamic Detectors | Junyi Feng, Jiashen Hua, Baisheng Lai | 2020 | arxiv 2012.04265 |
|---|