keyword
VidProM dataset
The VidProM dataset is a large-scale collection of real-world text prompts and artificial intelligence-generated videos designed to support research in text-to-video diffusion modeling. It comprises approximately 1.67 million unique text prompts submitted by real users, paired with around 6.69 million video clips generated across multiple text-to-video diffusion models. In addition to the paired prompts and videos, the dataset incorporates associated metadata such as text embeddings, safety classification scores, timestamps, and model identifiers. By providing a standardized repository of user input patterns and corresponding model syntheses, VidProM serves as a resource for advancing prompt engineering, improving video generation efficiency and quality, and exploring downstream safety applications such as deepfake detection and video copy detection.
1 item

