Text-conditioned video prediction on BridgeData
246.3FVDSeer
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| SeerPre.-weight=txt-img, Text=Yes, Resolution=256 × 2562023.03 | 246.3 | 0.55 | |
| PVDMPre.-weight=No, Text=No, Resolution=256 × 2562023.03 | 490.4 | 122.4 | |
| VideoFusionPre.-weight=txt-video, Text=Yes, Resolution=256 × 2562023.03 | 501.2 | 1.45 | |
| Tune-A-VideoPre.-weight=txt-img, Text=Yes, Resolution=256 × 2562023.03 | 515.7 | 2.01 | |
| SimVPPre.-weight=No, Text=No, Resolution=64 × 642023.03 | 681.6 | 0.73 | |
| TATSPre.-weight=video, Text=No, Resolution=128 × 1282023.03 | 1,253 | 6,213 | |
| MCVDPre.-weight=No, Text=No, Resolution=256 × 2562023.03 | 1,427 | 2.5 | |
| MAGEPre.-weight=video, Text=Yes, Resolution=128 × 1282023.03 | 2,605 | 3.19 |