VidGen-1M A Large-Scale Dataset for Text-to-video Generation paper page: https://
huggingface.co/papers/2408.02
629
… The quality of video-text pairs fundamentally determines the upper bound of text-to-video models. Currently, the datasets used for training these models suffer from significant
VidGen-1M: Large-Scale Dataset for Text-to-Video Generation
By
–
