Video Generation Model Evaluation Dataset
2
9 commits
1 linked in READMEs
updated Jan 19, 2025
This dataset contains human annotations for videos generated by different video generation models. The annotations evaluate the quality of generated videos across multiple dimensions.
Each JSON file represents one evaluation dimension and follows this structure:
The dataset includes videos generated by 7 different models:
| Dimension | Description | Scale |
|---|---|---|
| Static Quality | ||
| Image Quality | Evaluates technical aspects including clarity and sharpness | 1-5 |
| Aesthetic Quality | Assesses visual appeal and artistic composition | 1-5 |
| Dynamic Quality | ||
| Temporal Consistency | Measures frame-to-frame coherence and smoothness | 1-5 |
| Motion Effects | Evaluates quality of movement and dynamics | 1-5 |
| Video-Text Alignment | ||
| Video-Text Consistency | Overall alignment with text prompt | 1-5 |
| Object-Class Consistency | Accuracy of object representation | 1-3 |
| Color Consistency | Matching of colors with text prompt | 1-3 |
| Action Consistency | Accuracy of depicted actions | 1-3 |
| Scene Consistency | Correctness of scene environment | 1-3 |
This dataset can be used for:
7 commits
2 commits
Video Generation Model Evaluation Dataset
2
9 commits
1 linked in READMEs
updated Jan 19, 2025
This dataset contains human annotations for videos generated by different video generation models. The annotations evaluate the quality of generated videos across multiple dimensions.
Each JSON file represents one evaluation dimension and follows this structure:
The dataset includes videos generated by 7 different models:
| Dimension | Description | Scale |
|---|---|---|
| Static Quality | ||
| Image Quality | Evaluates technical aspects including clarity and sharpness | 1-5 |
| Aesthetic Quality | Assesses visual appeal and artistic composition | 1-5 |
| Dynamic Quality | ||
| Temporal Consistency | Measures frame-to-frame coherence and smoothness | 1-5 |
| Motion Effects | Evaluates quality of movement and dynamics | 1-5 |
| Video-Text Alignment | ||
| Video-Text Consistency | Overall alignment with text prompt | 1-5 |
| Object-Class Consistency | Accuracy of object representation | 1-3 |
| Color Consistency | Matching of colors with text prompt | 1-3 |
| Action Consistency | Accuracy of depicted actions | 1-3 |
| Scene Consistency | Correctness of scene environment | 1-3 |
This dataset can be used for:
7 commits
2 commits