---
title: VCGBench-Diverse
canonical_url: "https://www.modelscope.cn/datasets/MBZUAI/VCGBench-Diverse"
md_url: "https://www.modelscope.cn/datasets/MBZUAI/VCGBench-Diverse.md"
repository: MBZUAI/VCGBench-Diverse
last_updated: 2026-02-10
license: mit
storage_size: "29 GB"
downloads: 1081
stars: 0
---

# VCGBench-Diverse

> VCGBench-Diverse - MBZUAI 在 ModelScope 开源的数据集。👁️ VCGBench-Diverse Benchmarks

MBZUAI/VCGBench-Diverse 是 ModelScope 魔搭社区上的数据集，存储大小 29 GB，采用 mit 许可。

- **Repository**: MBZUAI/VCGBench-Diverse
- **License**: mit
- **Storage size**: 29 GB
- **Downloads**: 1081
- **Stars**: 0
- **Last updated**: 2026-02-10

Source: https://www.modelscope.cn/datasets/MBZUAI/VCGBench-Diverse

---

# 👁️ VCGBench-Diverse Benchmarks

---
## 📝 Description
Recognizing the limited diversity in existing video conversation benchmarks, we introduce VCGBench-Diverse to comprehensively evaluate the generalization ability of video LMMs. While VCG-Bench provides an extensive evaluation protocol, it is limited to videos from the ActivityNet200 dataset. Our benchmark comprises a total of 877 videos, 18 broad video categories and 4,354 QA pairs, ensuring a robust evaluation framework.


<p align="center">
  <img src="vcgbench_diverse.png" alt="Contributions">
</p>


## Dataset Contents
1. `vcgbench_diverse_qa.json` - Contains VCGBench-Diverse question-answer pairs.
2. `videos.tar.gz` - Contains the videos corresponding to `vcgbench_diverse_qa.json`.
3. `human_annotated_video_descriptions` - Contains original human-annotated dense descriptions of the videos.
4. `gpt_evaluation_scripts` - Contains the GPT-3.5-Turbo evaluation scripts to evaluate a model's predictions.
5. `sample_predictions` - Contains the VideoGPT+ predictions on the VCGBench-Diverse. Compatible with `gpt_evaluation_scripts`.

In order to evaluate your model on `VCGBench-Diverse`, use question-answer pairs in `vcgbench_diverse_qa.json` to generate your model's predictions in format same as
`sample_predictions` and then use `gpt_evaluation_scripts` for the evalution.


## 💻 Download
To get started, follow these steps:
   ```
   git lfs install
   git clone https://huggingface.co/MBZUAI/VCGBench-Diverse
   ```

## 📚 Additional Resources
- **Paper:** [ArXiv](https://arxiv.org/abs/2406.09418).
- **GitHub Repository:** For training and updates: [GitHub](https://github.com/mbzuai-oryx/VideoGPT-plus).
- **HuggingFace Collection:** For downloading the pretrained checkpoints, VCGBench-Diverse Benchmarks and Training data, visit [HuggingFace Collection - VideoGPT+](https://huggingface.co/collections/MBZUAI/videogpt-665c8643221dda4987a67d8d).

## 📜 Citations and Acknowledgments

```bibtex
  @article{Maaz2024VideoGPT+,
      title={VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding},
      author={Maaz, Muhammad and Rasheed, Hanoona and Khan, Salman and Khan, Fahad Shahbaz},
      journal={arxiv},
      year={2024},
      url={https://arxiv.org/abs/2406.09418}
  }
