本文へスキップ
AI図鑑

テキスト生成と続き書き

前の文脈を受けて一語ずつ書き進める

言語と知識初級 #01
入力テキストテキスト

本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。

この能力とは何か

Given a passage as context, the model writes what comes next. Both input and output are text and no particular transformation is targeted: it can continue a story, finish an explanation or draft an email. It differs from conversation in that no instruction format is required, and from summarisation or translation in that neither compression nor a language switch is the goal — both ends stay in the same modality.

技術的にどう実現するか

The dominant route is an autoregressive Transformer language model: text is split into tokens, the model predicts the probability of the next token position by position, and training maximises likelihood over the corpus. GPT, Llama and Qwen all follow it, differing in scale, data and post-training recipe. At inference, temperature, top-k and top-p shape diversity, while long outputs reuse cached keys and values to cut cost.

代表的な製品

27

GPT-4o

2024
OpenAI

ネイティブにマルチモーダルな汎用モデル。テキスト・画像・音声をひとつの入口で扱う

モデル クローズド
テキスト画像音声テキスト音声

Llama

2023
Meta AI (FAIR)

オープンウェイト路線を主流にした汎用モデル群

モデル オープンウェイト
テキストテキスト

Qwen

2023
Alibaba (Qwen)

多様な規模とマルチモーダル版を備えたオープンウェイトのモデル群

モデル オープンウェイト
テキスト画像テキスト

DeepSeek-V3

2024
DeepSeek

総パラメータ 671B、1 トークンあたり 37B だけを活性化するオープンウェイト MoE

モデル オープンウェイト
テキストテキスト

Mistral Large

2024
Mistral AI

欧州のオープンウェイト研究所による商用フラッグシップモデル

モデル クローズド
テキストテキスト

Gemini

2023
Google DeepMind

ネイティブにマルチモーダルで、超長文脈を扱う汎用モデル

モデル クローズド
テキスト画像音声動画テキスト

Claude

2023
Anthropic

長い文脈と安全性の調整で知られる汎用対話モデル

モデル クローズド
テキスト画像テキスト

MiniMax-M

2025
MiniMax

ハイブリッド注意と100万トークン文脈をもつオープンウェイトの推論モデル

モデル オープンウェイト
テキストテキスト

DeepSeek-R1

2025
DeepSeek

強化学習で推論の連鎖を訓練し、MIT ライセンスで重みを公開した推論モデル

モデル オープンウェイト
テキストテキスト

Apple Intelligence

2024
Apple

端末とプライベートクラウドで分担する OS 級の AI 機能

アプリ クローズド
テキスト画像音声テキスト画像

Command R

2024
Cohere

検索拡張とツール呼び出しのために設計された商用モデル

モデル オープンウェイト
テキストテキスト

Jamba

2024
AI21 Labs

状態空間モデルと Transformer を混合したオープンウェイトモデル

モデル オープンウェイト
テキストテキスト

DBRX

2024
Databricks

DatabricksのオープンウェイトMoE言語モデル

モデル オープンウェイト
テキストテキスト

Gemini

2024
Google DeepMind

検索・オフィス・マルチモーダルモデルを一つの対話入口に集約

アプリ フリーミアム
テキスト画像音声テキスト画像音声

Grok

2023
xAI

ソーシャルプラットフォームのデータと結びついた対話モデル

モデル クローズド
テキスト画像テキスト

Yi

2023
01.AI

中英バイリンガルで超長文脈版もあるオープンウェイトモデル

モデル オープンウェイト
テキストテキスト

Kimi

2023
Moonshot AI

長い文脈の処理を得意とする中国語の対話アシスタント

モデル クローズド
テキストテキスト

Hunyuan

2023
Tencent (Hunyuan)

開放ウェイト版を含むテンセントの汎用モデル群

モデル オープンウェイト
テキスト画像テキスト

Microsoft Copilot

2023
Microsoft

対話型 AI を OS とオフィスソフトに組み込む

アプリ フリーミアム
テキスト画像音声テキスト画像

Doubao

2023
ByteDance (Seed)

バイトダンスの汎用対話モデルとアプリ

モデル クローズド
テキスト画像テキスト

Phi

2023
Microsoft

小さな規模と厳選データで作られた高効率な小モデル

モデル オープンウェイト
テキストテキスト

Step

2023
StepFun

マルチモーダルとオンデバイスを狙う汎用モデル群

モデル クローズド
テキスト画像テキスト

Baichuan

2023
Baichuan AI

中国語シーンに向けたオープンウェイトの汎用モデル

モデル オープンウェイト
テキストテキスト

GLM

2023
Zhipu AI

自己回帰空白穴埋め事前学習から始まった中国語の汎用モデル

モデル クローズド
テキストテキスト

ChatGPT

2022
OpenAI

大規模言語モデルを誰もが使える対話画面にした

アプリ フリーミアム
テキスト画像音声テキスト画像音声

Character.AI

2022
Character.AI

自分で作ったキャラクターと長く会話できるロールプレイ

アプリ フリーミアム
テキスト音声テキスト音声

ERNIE

2019
Baidu

知識増強の事前学習から始まった中国語モデル、その初期の代表的版

モデル クローズド
テキストテキスト

関連する組織

代表的な用途

  • First drafts and rewriting
  • Drafting email, documents and copy
  • Code comments and API docs
  • Test data and synthetic corpora

どう評価するか

Perplexity
Predictive uncertainty on held-out text; lower is better
Human-preference Elo
Preference ranking from pairwise comparison
ROUGE/BLEU on constrained tasks
Reference overlap when an answer key exists

限界と難しさ

  • Fluency is not truth: the model states wrong facts with confidence — hallucination
  • Over long outputs, earlier setup drifts and the text contradicts itself
  • Poor sampling settings can trap the model in repetitive or degenerate loops

背景にある概念