이 페이지의 본문은 영어로 제공됩니다. 제목과 요약은 한국어로 번역되었습니다.
무엇인가
Qwen is a language-model series Alibaba began open-sourcing in August 2023. It spans parameters from 0.5B to 72B and covers both text-only and multimodal versions, with the Qwen-VL line accepting image input. Alibaba releases base models alongside instruction-tuned versions and specialised variants for coding and mathematics. Generations such as Qwen2.5 support a context window in the 128K range, with weights downloadable under a permissive licence. Its complete ladder of sizes makes it a widely used base for secondary development in the open-source community.
기억할 만한 이유
With a full ladder from 0.5B to 72B plus both text and multimodal lines released openly, it lets researchers always find a size that fits their compute — one of the broadest catalogues in the open-weight ecosystem.
주요 사양
- Parameters
- From 0.5B to 72B (Qwen2.5 family)
- Context window
- 128K tokens (Qwen2.5)
- Modality
- Text, image in (Qwen-VL); text out
- Open weights
- Yes
소속 능력
관련 개념
동종 제품
Llama
2023개방형 가중치 노선을 주류로 만든 범용 모델 계열
DeepSeek-V3
2024총 671B, 토큰당 37B만 활성화하는 오픈웨이트 MoE
Mistral Large
2024유럽 오픈웨이트 연구소의 플래그십 상용 모델
GLM
2023자기회귀 빈칸 채우기 사전학습에서 출발한 중국어 범용 모델
Phi
2023작은 규모와 선별 데이터로 만든 효율적 소형 모델
DeepSeek-R1
2025강화학습으로 추론 사슬을 훈련하고 MIT 라이선스로 가중치를 공개한 추론 모델
Kimi
2023긴 문맥 처리에 강한 중국어 대화 어시스턴트
ERNIE
2019지식 강화 사전학습에서 출발한 중국어 모델, 초기 대표 버전
Hunyuan
2023오픈웨이트 버전을 포함한 텐센트의 범용 모델 계열
MiniMax-M
2025하이브리드 어텐션과 백만 토큰 문맥을 갖춘 오픈웨이트 추론 모델
Yi
2023중영 이중언어, 초장문 문맥 버전을 제공하는 오픈웨이트 모델
Baichuan
2023중국어 환경을 겨냥한 오픈웨이트 범용 모델
DBRX
2024Databricks의 오픈 웨이트 MoE 언어 모델