Loading the SOTA2 catalog…
VIB-AVSR: Variational Information Bottleneck for Noise-Robust LLM-Based Audio-Visual Speech Recognition · SOTA2 Research