ResearchTasksText-and-Image to Audio-Visual GenerationFollowBenchmarksDataset NameSOTA MethodSortMost resultsRecently updatedMost papersApplyDataset NameSOTA methodMetricTrendResultsLast UpdatedAudio-Visual Generation Benchmark (test)LTX-24.15VA6May 12, 2026