対話と指示への追従
やり取りの中で意図を汲み、指示どおりに動く
本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。
この能力とは何か
Takes a multi-turn message sequence with roles (system, user, assistant) and outputs the next assistant turn. It is not merely continuation: the model must hold a persona across turns, remember earlier constraints, and obey format and tone instructions. It differs from plain text generation by its explicit role structure and the demand to follow instructions.
技術的にどう実現するか
The base is the same autoregressive language model; the difference lies in chat templates and post-training. System prompts and history are concatenated into a fixed sequence of special tokens, then instruction tuning teaches the model to answer rather than continue a web page. Preference data then shapes it through RLHF or a direct-preference method to favour helpful, harmless and honest behaviour. Tools and memory are injected into the prompt by external systems.
代表的な製品
30ChatGPT
2022大規模言語モデルを誰もが使える対話画面にした
Claude
2023長い文脈と安全性の調整で知られる汎用対話モデル
Gemini
2024検索・オフィス・マルチモーダルモデルを一つの対話入口に集約
Kimi
2023長い文脈の処理を得意とする中国語の対話アシスタント
Doubao
2023バイトダンスの汎用対話モデルとアプリ
Character.AI
2022自分で作ったキャラクターと長く会話できるロールプレイ
Microsoft Copilot
2023対話型 AI を OS とオフィスソフトに組み込む
MiniMax-M
2025ハイブリッド注意と100万トークン文脈をもつオープンウェイトの推論モデル
o3
2025答える前に長い推論を重ね、推論時の計算で正答率を上げる
DeepSeek-R1
2025強化学習で推論の連鎖を訓練し、MIT ライセンスで重みを公開した推論モデル
DeepSeek-V3
2024総パラメータ 671B、1 トークンあたり 37B だけを活性化するオープンウェイト MoE
Apple Intelligence
2024端末とプライベートクラウドで分担する OS 級の AI 機能
GPT-4o
2024ネイティブにマルチモーダルな汎用モデル。テキスト・画像・音声をひとつの入口で扱う
Command R
2024検索拡張とツール呼び出しのために設計された商用モデル
Jamba
2024状態空間モデルと Transformer を混合したオープンウェイトモデル
DBRX
2024DatabricksのオープンウェイトMoE言語モデル
Mistral Large
2024欧州のオープンウェイト研究所による商用フラッグシップモデル
Gemini
2023ネイティブにマルチモーダルで、超長文脈を扱う汎用モデル
Grok
2023ソーシャルプラットフォームのデータと結びついた対話モデル
Yi
2023中英バイリンガルで超長文脈版もあるオープンウェイトモデル
Hunyuan
2023開放ウェイト版を含むテンセントの汎用モデル群
Qwen
2023多様な規模とマルチモーダル版を備えたオープンウェイトのモデル群
Phi
2023小さな規模と厳選データで作られた高効率な小モデル
Step
2023マルチモーダルとオンデバイスを狙う汎用モデル群
Baichuan
2023中国語シーンに向けたオープンウェイトの汎用モデル
GLM
2023自己回帰空白穴埋め事前学習から始まった中国語の汎用モデル
Llama
2023オープンウェイト路線を主流にした汎用モデル群
Pangu
2020華為の盤古シリーズ基盤モデル
ERNIE
2019知識増強の事前学習から始まった中国語モデル、その初期の代表的版
Siri
2011音声アシスタントを主流のスマホに広めた初期の存在
関連する組織
代表的な用途
- General assistants and support bots
- Coding and study tutoring
- Role-play and companion apps
- Internal knowledge-helpdesk entry points
どう評価するか
- Human-preference Elo
- Ranking by blind pairwise win rate
- Instruction-following rate
- Share of verifiable constraints (format, length, banned words) satisfied
- Safety violation rate
- Rate of violations under red-team prompts
限界と難しさ
- Constraints from early turns, especially format rules, are forgotten or softened in long chats
- It tends to agree with the user even when the premise is wrong — sycophancy
- Carefully crafted prompts can bypass safety policy; jailbreaks are hard to eliminate