Loading the SOTA2 catalog…
DCFormer: Efficient 3D Vision-Language Modeling with Decomposed Convolutions · SOTA2 Research