---
title: VoiceAssistant-430K-vocalnet
canonical_url: "https://www.modelscope.cn/datasets/VocalNet/VoiceAssistant-430K-vocalnet"
md_url: "https://www.modelscope.cn/datasets/VocalNet/VoiceAssistant-430K-vocalnet.md"
repository: VocalNet/VoiceAssistant-430K-vocalnet
last_updated: 2025-04-23
license: "Apache License 2.0"
storage_size: "86 GB"
downloads: 5919
stars: 0
---

# VoiceAssistant-430K-vocalnet

> VoiceAssistant-430K-vocalnet - VocalNet 在 ModelScope 开源的数据集。VoiceAssistant-430K-vocalnet

VocalNet/VoiceAssistant-430K-vocalnet 是 ModelScope 魔搭社区上的数据集，存储大小 86 GB，采用 Apache License 2.0 许可。

- **Repository**: VocalNet/VoiceAssistant-430K-vocalnet
- **License**: Apache License 2.0
- **Storage size**: 86 GB
- **Downloads**: 5919
- **Stars**: 0
- **Last updated**: 2025-04-23

Source: https://www.modelscope.cn/datasets/VocalNet/VoiceAssistant-430K-vocalnet

---

# VoiceAssistant-430K-vocalnet

This dataset supports the reproduction of [VocalNet](https://github.com/SJTU-OmniAgent/VocalNet).

## Data Construction

1. **Data Source**: We used the [VoiceAssistant-400K](https://huggingface.co/datasets/gpt-omni/VoiceAssistant-400K) from Mini-Omni, which contains about 470K instances.
2. **Data Filtering**: We removed samples with excessively long data. The resulted corpus contains 430K instances.
3. **Response Speech**: We perform speech synthesis using [CosyVoice](https://github.com/FunAudioLLM/CosyVoice) to generate the response speech.
4. **Response Token**: We generate the speech token using [CosyVoice2](https://github.com/FunAudioLLM/CosyVoice).

## Acknowledgment

1. The original data is from [Mini-Omni](https://github.com/gpt-omni/mini-omni).
2. The generation of speech wave and token is from [CosyVoice/CosyVoice2](https://github.com/gpt-omni/mini-omni).
