---
title: CareBot_Medical_multi-llama3-8b-instruct
canonical_url: "https://www.modelscope.cn/models/BAAI/CareBot_Medical_multi-llama3-8b-instruct"
md_url: "https://www.modelscope.cn/models/BAAI/CareBot_Medical_multi-llama3-8b-instruct.md"
repository: BAAI/CareBot_Medical_multi-llama3-8b-instruct
last_updated: 2026-08-24
license: apache-2.0
pipeline_tag: text-generation
tasks:
  - text-generation
model_type:
  - llama
architectures:
  - LlamaForCausalLM
base_model:
  - XiaofengAlg/CareBot_Medical_multi-llama3-8b-base
base_model_relation: finetune
parameters: 8.0B
tensor_type:
  - BF16
library_name:
  - safetensors
  - pytorch
frameworks:
  - pytorch
downloads: 2577
stars: 22
tags:
  - "医疗对话模型"
  - "中英文多语种医疗对话模型"
  - chatmodel
  - "arxiv:2412.15236"
---

# CareBot_Medical_multi-llama3-8b-instruct

> CareBot_Medical_multi-llama3-8b-instruct - BAAI 在 ModelScope 开源的模型。This model is trained from XiaofengAlg/CareBotMedicalmulti-llama3-8b-base using BAAI/IndustryInstructionHealth-Medicine. To enhance the model's ability to follow medical instructions and…

BAAI/CareBot_Medical_multi-llama3-8b-instruct 是 ModelScope 魔搭社区上的 8.0B 参数text-generation模型，采用 apache-2.0 许可，基于 XiaofengAlg/CareBot_Medical_multi-llama3-8b-base 构建。

- **Repository**: BAAI/CareBot_Medical_multi-llama3-8b-instruct
- **License**: apache-2.0
- **Tasks**: text-generation
- **Parameters**: 8.0B
- **Base model**: XiaofengAlg/CareBot_Medical_multi-llama3-8b-base
- **Tags**: 医疗对话模型, 中英文多语种医疗对话模型, chatmodel, arxiv:2412.15236
- **Downloads**: 2577
- **Stars**: 22
- **Last updated**: 2026-08-24

Source: https://www.modelscope.cn/models/BAAI/CareBot_Medical_multi-llama3-8b-instruct

---

This model is trained from [XiaofengAlg/CareBot_Medical_multi-llama3-8b-base](https://huggingface.co/XiaofengAlg/CareBot_Medical_multi-llama3-8b-base) using BAAI/IndustryInstruction_Health-Medicine. To enhance the model's ability to follow medical instructions and better adapt to specific medical scenarios, we conduct supervised fine-tuning. This process involves using conversational-style data (comprising both queries and responses) to fine-tune the pretrained LLM. In the following sections, we explore the data construction and training methods.

## Data Construction

Our SFT dataset comprises a diverse array of question types, including multiple-choice questions from medical exams, single-turn disease diagnoses, and multi-turn health consultations. It integrates data from seven publicly available sources: Chinese Medical Dialogue Data\footnote{https://github.com/Toyhom/Chinese-medical-dialogue-data}, Huatuo26M , MedDialog , ChatMed Consult Dataset , ChatDoctor , CMB\footnote{https://github.com/FreedomIntelligence/CMB}, and MedQA . We preserve portions of authentic doctor-patient conversations and augment the dataset by rewriting the remaining content. For these rewrites, we use real-world medical scenarios as prompts and generate responses via GPT-4. We believe this ensures the diversity of the SFT dataset, which can help the CareBot better adapt to different types of medical problems and patient situations, thereby improving its performance in a variety of scenarios.

## evaluation

evaluation on benchmark is bellow.
![image/png](https://cdn-uploads.huggingface.co/production/uploads/642f6c64f945a8a5c9ee5b5d/kqvLfcFtkw6lHcHtCySLr.png)

![image/png](https://cdn-uploads.huggingface.co/production/uploads/642f6c64f945a8a5c9ee5b5d/UiokfV8qcYEyCWEa__820.png)


gsb result with other medical LLMS
![image/png](https://cdn-uploads.huggingface.co/production/uploads/642f6c64f945a8a5c9ee5b5d/rOnnIoY9MaXPTFD_R10r1.png)




## Citation

```bibtex
@inproceedings{zhao2025carebot,
  title={CareBot: A Pioneering Full-Process Open-Source Medical Language Model},
  author={Lulu Zhao and Weihao Zeng and Xiaofeng Shi and Hua Zhou},
  booktitle={Proceedings of the AAAI Conference on Artificial Intelligence},
  volume={39},
  number={24},
  pages={26039--26047},
  year={2025},
  doi={10.1609/aaai.v39i24.34799},
  url={https://doi.org/10.1609/aaai.v39i24.34799}
}
```

# Acknowledgements

This work is supported by the National Science and Technology Major Project (No. 2022ZD0116314).

本项目受新一代人工智能国家科技重大专项（No. 2022ZD0116314）支持。
