---
title: DocVQA
canonical_url: "https://www.modelscope.cn/datasets/lmms-lab/DocVQA"
md_url: "https://www.modelscope.cn/datasets/lmms-lab/DocVQA.md"
repository: lmms-lab/DocVQA
last_updated: 2024-11-21
license: "Apache License 2.0"
storage_size: "11 GB"
downloads: 12114
stars: 2
---

# DocVQA

> DocVQA - lmms-lab 在 ModelScope 开源的数据集。Large-scale Multi-modality Models Evaluation Suite

lmms-lab/DocVQA 是 ModelScope 魔搭社区上的数据集，存储大小 11 GB，采用 Apache License 2.0 许可。

- **Repository**: lmms-lab/DocVQA
- **License**: Apache License 2.0
- **Storage size**: 11 GB
- **Downloads**: 12114
- **Stars**: 2
- **Last updated**: 2024-11-21

Source: https://www.modelscope.cn/datasets/lmms-lab/DocVQA

---

<p align="center" width="100%">
<img src="https://i.postimg.cc/g0QRgMVv/WX20240228-113337-2x.png"  width="100%" height="80%">
</p>

# Large-scale Multi-modality Models Evaluation Suite

> Accelerating the development of large-scale multi-modality models (LMMs) with `lmms-eval`

🏠 [Homepage](https://lmms-lab.github.io/) | 📚 [Documentation](docs/README.md) | 🤗 [Huggingface Datasets](https://huggingface.co/lmms-lab)

# This Dataset

This is a formatted version of [DocVQA](https://arxiv.org/abs/2007.00398). It is used in our `lmms-eval` pipeline to allow for one-click evaluations of large multi-modality models.

```
@article{mathew2020docvqa,
  title={DocVQA: A Dataset for VQA on Document Images. CoRR abs/2007.00398 (2020)},
  author={Mathew, Minesh and Karatzas, Dimosthenis and Manmatha, R and Jawahar, CV},
  journal={arXiv preprint arXiv:2007.00398},
  year={2020}
}
```
