---
title: medical-symptoms-english-audio
canonical_url: "https://www.modelscope.cn/datasets/Kratos-AI/medical-symptoms-english-audio"
md_url: "https://www.modelscope.cn/datasets/Kratos-AI/medical-symptoms-english-audio.md"
repository: Kratos-AI/medical-symptoms-english-audio
last_updated: 2025-08-14
license: cc-by-4.0
storage_size: "74 MB"
downloads: 250
stars: 0
---

# medical-symptoms-english-audio

> medical-symptoms-english-audio - Kratos-AI 在 ModelScope 开源的数据集。Medical Symptoms English Audio Dataset

Kratos-AI/medical-symptoms-english-audio 是 ModelScope 魔搭社区上的数据集，存储大小 74 MB，采用 cc-by-4.0 许可。

- **Repository**: Kratos-AI/medical-symptoms-english-audio
- **License**: cc-by-4.0
- **Storage size**: 74 MB
- **Downloads**: 250
- **Stars**: 0
- **Last updated**: 2025-08-14

Source: https://www.modelscope.cn/datasets/Kratos-AI/medical-symptoms-english-audio

---

# Medical Symptoms English Audio Dataset

*This dataset contains intentionally low-quality (“B-grade”) data. It has been curated to include noisy, imperfect, or otherwise suboptimal samples for the purpose of testing model robustness and performance under degraded input conditions

**Text spoken by all participants:**
"Doctor, I'm constantly tired, like a heavy fog I can't shake. Sharp headaches hit, worse at night, and sleep is tough. I get dizzy, and my stomach feels uneasy after meals. I'm really worried it’s serious. Please help me figure out what's wrong."

The dataset supports training and evaluation of models in:

- Automatic Speech Recognition (ASR)
- Emotional tone classification
- Voice synthesis and generation
- Emotion-aware conversational agents

---

## Intended Uses

### ✅ Direct Use

- Training and benchmarking ASR models with Indian-accented English
- Emotion detection and classification from voice
- Research in affective computing and empathetic AI

### ❌ Out-of-Scope Use

- Real-time or production-grade systems
- Commercial use without proper CC BY 4.0 attribution
- Clinical or diagnostic use cases

---

## Considerations and Limitations

- ❗ The dataset is small (<1,000 samples) and not fully representative of India's linguistic and emotional diversity
- 💡 Emotions are subjective — classification results may vary by listener or model
- 🔄 Future versions will aim to expand multilingual support and speaker diversity

---

## License

**CC BY 4.0** — You can use, modify, and share the dataset with appropriate credit.

---

## Contact

  - For queries or collaborations related to datasets, contact at :
    - anoushka@kgen.io
    - abhishek.vadapalli@kgen.io

---
