이 페이지의 본문은 영어로 제공됩니다. 제목과 요약은 한국어로 번역되었습니다.
무엇인가
Llama is Meta’s large language model family, first released in February 2023 and distributed as open weights that can be downloaded and run in one’s own environment. From Llama 2 onward, Meta permitted commercial use and offered several sizes such as 7B, 13B and 70B; Llama 3.1 extended the largest version to 405B parameters. It is a pretrained base on which the community performs extensive instruction tuning and domain adaptation. Its open weights made it a default starting point for many open-source projects.
기억할 만한 이유
By releasing large models as open weights under a commercially usable licence, it made “download a base model and fine-tune it yourself” common practice and opened the open-weight camp’s direct competition with the closed frontier.
주요 사양
- Parameters
- 8B / 70B / 405B (Llama 3.1 family)
- Context window
- 128K tokens (Llama 3.1)
- Open weights
- Yes
- Released
- 2023-02
소속 능력
관련 개념
동종 제품
Mistral Large
2024유럽 오픈웨이트 연구소의 플래그십 상용 모델
Qwen
2023여러 규모와 멀티모달 버전을 아우르는 오픈웨이트 모델 계열
DeepSeek-V3
2024총 671B, 토큰당 37B만 활성화하는 오픈웨이트 MoE
GPT-4o
2024네이티브 멀티모달 범용 모델. 텍스트·이미지·오디오를 한 창구에서 다룬다
Claude
2023긴 문맥과 안전 정렬로 알려진 범용 대화 모델
Gemini
2023네이티브 멀티모달에 초장문 문맥을 다루는 범용 모델
Phi
2023작은 규모와 선별 데이터로 만든 효율적 소형 모델
Command R
2024검색 증강과 도구 호출을 위해 설계된 상용 모델
Jamba
2024상태공간 모델과 Transformer를 혼합한 오픈웨이트 모델
Yi
2023중영 이중언어, 초장문 문맥 버전을 제공하는 오픈웨이트 모델
DBRX
2024Databricks의 오픈 웨이트 MoE 언어 모델