Anthropic
Claude와 해석가능성 연구로 알려진 연구소
이 페이지의 본문은 영어로 제공됩니다. 제목과 요약은 한국어로 번역되었습니다.
무엇인가
Anthropic was founded in 2021 by a team largely from OpenAI, with AI safety and interpretability as its central themes. It trains the Claude family and was among the first to put Constitutional AI into its alignment pipeline — using a written set of principles so the model critiques and revises itself rather than relying only on human labels. It also publishes interpretability work that tries to identify the features and circuits inside a model.
왜 중요한가
It brought Constitutional AI into mainstream alignment practice, making "constrain the model with a written set of principles" a route alongside RLHF; and it moved interpretability research from the margins onto a frontier lab’s official agenda.
주요 이정표
- 2021
Founded by former OpenAI members
- 2023
Released the Claude model
- 2025
Released the Claude Code coding tool