Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| Huan-ang Gao, Mingju Gao, Jiaju Li |
| 2024 |
| arxiv 2403.09638 |
| Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation | Mengdan Zhu, Senhao Cheng, Guangji Bai | 2025 | arxiv 2505.21956 |
|---|
| DP-RDM: Adapting Diffusion Models to Private Domains Without Fine-Tuning | Jonathan Lebensold, Maziar Sanjabi, Pietro Astolfi | 2024 | arxiv 2403.14421 |
|---|
| CLIP as RNN: Segment Countless Visual Concepts without Training Endeavor | Shuyang Sun, Runjia Li, Philip Torr | 2023 | arxiv 2312.07661 |
|---|
| SDDGR: Stable Diffusion-based Deep Generative Replay for Class Incremental Object Detection | Junsu Kim, Hoseong Cho, Jihyeon Kim | 2024 | arxiv 2402.17323 |
|---|
| MPDIoU: A Loss for Efficient and Accurate Bounding Box Regression | Siliang Ma, Yong Xu | 2023 | arxiv 2307.07662 |
|---|
| Active Learning for Finely-Categorized Image-Text Retrieval by Selecting Hard Negative Unpaired Samples | Dae Ung Jo, Kyuewang Lee, JaeHo Chung | 2024 | arxiv 2405.16301 |
|---|
| LMPT: Prompt Tuning with Class-Specific Embedding Loss for Long-tailed Multi-Label Visual Recognition | Peng Xia, Di Xu, Ming Hu | 2023 | arxiv 2305.04536 |
|---|
| Visual Hallucinations of Multi-modal Large Language Models | Wen Huang, Hongbin Liu, Minxin Guo | 2024 | arxiv 2402.14683 |
|---|
| Introducing Routing Functions to Vision-Language Parameter-Efficient Fine-Tuning with Low-Rank Bottlenecks | Tingyu Qu, Tinne Tuytelaars, Marie-Francine Moens | 2024 | arxiv 2403.09377 |
|---|
| Context-Guided Spatial Feature Reconstruction for Efficient Semantic Segmentation | Zhenliang Ni, Xinghao Chen, Yingjie Zhai | 2024 | arxiv 2405.06228 |
|---|