Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model |
|---|
| Zheng Zhang, Yeyao Ma, Enming Zhang |
| 2024 |
| arxiv 2403.14598 |
| Estimating Human Poses Across Datasets: A Unified Skeleton and Multi-Teacher Distillation Approach | Muhammad Saif Ullah Khan, Dhavalkumar Limbachiya, Didier Stricker | 2024 | arxiv 2405.20084 |
|---|
| A Strong Baseline for Generalized Few-Shot Semantic Segmentation | Sina Hajimiri, Malik Boudiaf, Ismail Ben Ayed | 2022 | arxiv 2211.14126 |
|---|
| Detection Transformer with Stable Matching | Shilong Liu, Tianhe Ren, Jiayu Chen | 2023 | arxiv 2304.04742 |
|---|
| ScaleDet: A Scalable Multi-Dataset Object Detector | Yanbei Chen, Manchen Wang, Abhay Mittal | 2023 | arxiv 2306.04849 |
|---|
| DynaSeg: A Deep Dynamic Fusion Method for Unsupervised Image Segmentation Incorporating Feature Similarity and Spatial Continuity | Boujemaa Guermazi, Naimul Khan | 2024 | arxiv 2405.05477 |
|---|
| SIGN: A Statistically-Informed Gaze Network for Gaze Time Prediction | Jianping Ye, Michel Wedel | 2025 | arxiv 2501.17422 |
|---|
| LambdaNetworks: Modeling Long-Range Interactions Without Attention | Irwan Bello | 2021 | arxiv 2102.08602 |
|---|
| Data-Uncertainty Guided Multi-Phase Learning for Semi-Supervised Object Detection | Zhenyu Wang, Yali Li, Ye Guo | 2021 | arxiv 2103.16368 |
|---|
| Multimodal Contrastive Training for Visual Representation Learning | Xin Yuan, Zhe Lin, Jason Kuen | 2021 | arxiv 2104.12836 |
|---|