本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。
これは何か
Qwen is a language-model series Alibaba began open-sourcing in August 2023. It spans parameters from 0.5B to 72B and covers both text-only and multimodal versions, with the Qwen-VL line accepting image input. Alibaba releases base models alongside instruction-tuned versions and specialised variants for coding and mathematics. Generations such as Qwen2.5 support a context window in the 128K range, with weights downloadable under a permissive licence. Its complete ladder of sizes makes it a widely used base for secondary development in the open-source community.
なぜ覚えておく価値があるか
With a full ladder from 0.5B to 72B plus both text and multimodal lines released openly, it lets researchers always find a size that fits their compute — one of the broadest catalogues in the open-weight ecosystem.
主な仕様
- Parameters
- From 0.5B to 72B (Qwen2.5 family)
- Context window
- 128K tokens (Qwen2.5)
- Modality
- Text, image in (Qwen-VL); text out
- Open weights
- Yes
対応する能力
関連する概念
同種の製品
Llama
2023オープンウェイト路線を主流にした汎用モデル群
DeepSeek-V3
2024総パラメータ 671B、1 トークンあたり 37B だけを活性化するオープンウェイト MoE
Mistral Large
2024欧州のオープンウェイト研究所による商用フラッグシップモデル
GLM
2023自己回帰空白穴埋め事前学習から始まった中国語の汎用モデル
Phi
2023小さな規模と厳選データで作られた高効率な小モデル
DeepSeek-R1
2025強化学習で推論の連鎖を訓練し、MIT ライセンスで重みを公開した推論モデル
Kimi
2023長い文脈の処理を得意とする中国語の対話アシスタント
ERNIE
2019知識増強の事前学習から始まった中国語モデル、その初期の代表的版
Hunyuan
2023開放ウェイト版を含むテンセントの汎用モデル群
MiniMax-M
2025ハイブリッド注意と100万トークン文脈をもつオープンウェイトの推論モデル
Yi
2023中英バイリンガルで超長文脈版もあるオープンウェイトモデル
Baichuan
2023中国語シーンに向けたオープンウェイトの汎用モデル
DBRX
2024DatabricksのオープンウェイトMoE言語モデル