qwen2.5

Qwen2.5 模型在阿里巴巴最新的超大规模数据集上进行了预训练，该数据集包含高达 18 万亿个词元。该模型支持高达 128K 个词元，并具有多语言支持。

工具 0.5b 1.5b 3b 7b 14b 32b 72b

1.9M 拉取更新 7 周前

133 个标签

7 周前更新

7 周前

b439066a7c50 · 339MB

Apache 许可证 2.0 版，200 年 1 月

11kB

自述文件

Qwen2.5 是最新的 Qwen 大型语言模型系列。对于 Qwen2.5，发布了一系列基础语言模型和指令微调模型，其大小范围从 0.5 到 720 亿个参数。与 Qwen2 相比，Qwen2.5 引入了以下改进

它拥有**更丰富的知识**，并在**编码**和**数学**方面有了很大的提升，这得益于这些领域中的专业专家模型。
它在**指令遵循**、**长文本生成**（超过 8K 个词元）、**理解结构化数据**（例如表格）和**生成结构化输出**（尤其是 JSON 格式）方面取得了重大进展。它也**对各种系统提示更具弹性**，改善了聊天机器人的角色扮演和条件设置。
它支持高达 128K 个词元的**长上下文**，并可以生成高达 8K 个词元。
它提供**多语言支持**，支持 29 种以上的语言，包括中文、英语、法语、西班牙语、葡萄牙语、德语、意大利语、俄语、日语、韩语、越南语、泰语、阿拉伯语等等。

请注意：除了 3B 和 72B 模型之外，所有模型都根据 Apache 2.0 许可证发布，而 3B 和 72B 模型则根据 Qwen 许可证发布。

参考文献

GitHub

博客文章

HuggingFace

Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, a range of base language models and instruction-tuned models are released, with sizes ranging from 0.5 to 72 billion parameters. Qwen2.5 introduces the following improvements over Qwen2:

- It possesses **significantly more knowledge** and has greatly enhanced capabilities in **coding** and **mathematics**, due to specialized expert models in these domains.
- It demonstrates significant advancements in **instruction following**, **long-text generation** (over 8K tokens), **understanding structured data** (e.g., tables), and **generating structured outputs**, especially in JSON format. It is also **more resilient to diverse system prompts**, improving role-play and condition-setting for chatbots.
- It supports **long contexts** of up to 128K tokens and can generate up to 8K tokens.
- It offers **multilingual support** for over 29 languages, including Chinese, English, French, Spanish, Portuguese, German, Italian, Russian, Japanese, Korean, Vietnamese, Thai, Arabic, and more.

Please note: all models except the 3B and 72B are released under the Apache 2.0 license, while the 3B and 72B models are under the Qwen license.

## References

[GitHub](https://github.com/QwenLM/Qwen2.5)

[Blog post](https://qwenlm.github.io/blog/qwen2.5/)

[HuggingFace](https://hugging-face.cn/collections/Qwen/qwen25-66e81a666513e518adb90d9e)

粘贴、拖放或点击上传图像（.png、.jpeg、.jpg、.svg、.gif）