Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| Pengxi Zeng, Alberto Presta, Jonah Reinis |
| 2024 |
| arxiv 2410.00582 |
| Dynamic Tuning Towards Parameter and Inference Efficiency for ViT Adaptation | Wangbo Zhao, Jiasheng Tang, Yizeng Han | 2024 | arxiv 2403.11808 |
|---|
| Is Your LiDAR Placement Optimized for 3D Scene Understanding? | Ye Li, Lingdong Kong, Hanjiang Hu | 2024 | arxiv 2403.17009 |
|---|
| VividMed: Vision Language Model with Versatile Visual Grounding for Medicine | Lingxiao Luo, Bingda Tang, Xuanzhong Chen | 2024 | arxiv 2410.12694 |
|---|
| iFormer: Integrating ConvNet and Transformer for Mobile Application | Chuanyang Zheng | 2025 | arxiv 2501.15369 |
|---|
| COSMOS: Cross-Modality Self-Distillation for Vision Language Pre-training | Sanghwan Kim, Rui Xiao, Mariana-Iuliana Georgescu | 2024 | arxiv 2412.01814 |
|---|
| Benchmarking Large Vision-Language Models via Directed Scene Graph for Comprehensive Image Captioning | Fan Lu, Wei Wu, Kecheng Zheng | 2024 | arxiv 2412.08614 |
|---|
| SN-LiDAR: Semantic Neural Fields for Novel Space-time View LiDAR Synthesis | Yi Chen, Tianchen Deng, Wentao Zhao | 2025 | arxiv 2504.08361 |
|---|
| Near, far: Patch-ordering enhances vision foundation models' scene understanding | Valentinos Pariza, Mohammadreza Salehi, Gertjan Burghouts | 2024 | arxiv 2408.11054 |
|---|
| Occlusion-Aware Self-Supervised Monocular Depth Estimation for Weak-Texture Endoscopic Images | Zebo Huang, Yinghui Wang | 2025 | arxiv 2504.17582 |
|---|
| Gompertz Linear Units: Leveraging Asymmetry for Enhanced Learning Dynamics | Indrashis Das, Mahmoud Safari, Steven Adriaensen | 2025 | arxiv 2502.03654 |
|---|