Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| Kuiliang Gao, Anzhu Yu, Xiong You |
| 2023 |
| arxiv 2305.09893 |
| Modulating Pretrained Diffusion Models for Multimodal Image Synthesis | Cusuh Ham, James Hays, Jingwan Lu | 2023 | arxiv 2302.12764 |
|---|
| MM-3DScene: 3D Scene Understanding by Customizing Masked Modeling with Informative-Preserved Reconstruction and Self-Distilled Consistency | Mingye Xu, Mutian Xu, Tong He | 2022 | arxiv 2212.09948 |
|---|
| Art Authentication with Vision Transformers | Ludovica Schaerf, Carina Popovici, Eric Postma | 2023 | arxiv 2307.03039 |
|---|
| Stroke Extraction of Chinese Character Based on Deep Structure Deformable Image Registration | Meng Li, Yahan Yu, Yi Yang | 2023 | arxiv 2307.04341 |
|---|
| High-Resolution Vision Transformers for Pixel-Level Identification of Structural Components and Damage | Kareem Eltouny, Seyedomid Sajedi, Xiao Liang | 2023 | arxiv 2308.03006 |
|---|
| Model-based inexact graph matching on top of CNNs for semantic scene understanding | Jérémy Chopin, Jean-Baptiste Fasquel, Harold Mouchère | 2023 | arxiv 2301.07468 |
|---|
| Improving Pixel-based MIM by Reducing Wasted Modeling Capability | Yuan Liu, Songyang Zhang, Jiacheng Chen | 2023 | arxiv 2308.00261 |
|---|
| Context Autoencoder for Self-Supervised Representation Learning | Xiaokang Chen, Mingyu Ding, Xiaodi Wang | 2022 | arxiv 2202.03026 |
|---|
| Future Video Prediction from a Single Frame for Video Anomaly Detection | Mohammad Baradaran, Robert Bergevin | 2023 | arxiv 2308.07783 |
|---|
| Food Image Classification and Segmentation with Attention-based Multiple Instance Learning | Valasia Vlachopoulou, Ioannis Sarafis, Alexandros Papadopoulos | 2023 | arxiv 2308.11452 |
|---|