---
title: chinese-mixtral-instruct
canonical_url: "https://www.modelscope.cn/models/ChineseAlpacaGroup/chinese-mixtral-instruct"
md_url: "https://www.modelscope.cn/models/ChineseAlpacaGroup/chinese-mixtral-instruct.md"
repository: ChineseAlpacaGroup/chinese-mixtral-instruct
last_updated: 2024-04-24
license: apache-2.0
pipeline_tag: fill-mask
tasks:
  - fill-mask
model_type:
  - mixtral
architectures:
  - MixtralForCausalLM
parameters: 46.7B
tensor_type:
  - BF16
  - F16
library_name:
  - safetensors
  - pytorch
frameworks:
  - Pytorch
language:
  - zh
  - en
inference_backends:
  - "deploy_task text/emb"
  - "lmdeploy 0.9.1"
  - "lmdeploy_turbomind 0.9.1"
  - "sglang 0.5.2"
  - "vllm 0.9.2"
downloads: 165
stars: 0
tags:
  - moe
---

# chinese-mixtral-instruct

> chinese-mixtral-instruct - ChineseAlpacaGroup 在 ModelScope 开源的模型。Chinese-Mixtral-Instruct

ChineseAlpacaGroup/chinese-mixtral-instruct 是 ModelScope 魔搭社区上的 46.7B 参数fill-mask模型，采用 apache-2.0 许可，可用 deploy_task text/emb、lmdeploy 0.9.1、lmdeploy_turbomind 0.9.1 部署。

- **Repository**: ChineseAlpacaGroup/chinese-mixtral-instruct
- **License**: apache-2.0
- **Tasks**: fill-mask
- **Parameters**: 46.7B
- **Inference backends**: deploy_task text/emb, lmdeploy 0.9.1, lmdeploy_turbomind 0.9.1, sglang 0.5.2, vllm 0.9.2
- **Tags**: moe
- **Downloads**: 165
- **Stars**: 0
- **Last updated**: 2024-04-24

Source: https://www.modelscope.cn/models/ChineseAlpacaGroup/chinese-mixtral-instruct

---

# Chinese-Mixtral-Instruct
<p align="center">
    <a href="https://github.com/ymcui/Chinese-Mixtral"><img src="https://ymcui.com/images/chinese-mixtral-banner.png" width="600"/></a>
</p>

**Chinese Mixtral GitHub repository: https://github.com/ymcui/Chinese-Mixtral**

This repository contains **Chinese-Mixtral-Instruct**, which is further tuned with instruction data on [Chinese-Mixtral](https://huggingface.co/hfl/chinese-mixtral), where Chinese-Mixtral is build on top of [Mixtral-8x7B-v0.1](https://huggingface.co/mistralai/Mixtral-8x7B-v0.1).

**Note: this is an instruction (chat) model, which can be used for conversation, QA, etc.**

## Others

- For LoRA-only model, please see: https://huggingface.co/hfl/chinese-mixtral-instruct-lora

- For GGUF model (llama.cpp compatible), please see: https://huggingface.co/hfl/chinese-mixtral-instruct-gguf

- If you have questions/issues regarding this model, please submit an issue through https://github.com/ymcui/Chinese-Mixtral/.

## Citation

Please consider cite our paper if you use the resource of this repository.
Paper link: https://arxiv.org/abs/2403.01851
```
@article{chinese-mixtral,
      title={Rethinking LLM Language Adaptation: A Case Study on Chinese Mixtral}, 
      author={Cui, Yiming and Yao, Xin},
      journal={arXiv preprint arXiv:2403.01851},
      url={https://arxiv.org/abs/2403.01851},
      year={2024}
}
```
