本文へスキップ
AI図鑑

Anthropic

Claude と解釈可能性研究で知られる研究所

アメリカ合衆国 最先端ラボ クローズド

本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。

これは何か

Anthropic was founded in 2021 by a team largely from OpenAI, with AI safety and interpretability as its central themes. It trains the Claude family and was among the first to put Constitutional AI into its alignment pipeline — using a written set of principles so the model critiques and revises itself rather than relying only on human labels. It also publishes interpretability work that tries to identify the features and circuits inside a model.

なぜ重要なのか

It brought Constitutional AI into mainstream alignment practice, making "constrain the model with a written set of principles" a route alongside RLHF; and it moved interpretability research from the margins onto a frontier lab’s official agenda.

主なマイルストーン

  1. 2021

    Founded by former OpenAI members

  2. 2023

    Released the Claude model

  3. 2025

    Released the Claude Code coding tool

代表的な製品

2

主な能力領域