Skip to content
AI Atlas

ERNIE

A Chinese model that began with knowledge-enhanced pretraining, an early landmark version

Baidu Model ClosedLegacy
inTextText

WHAT IT IS

ERNIE is Baidu’s large language model series; the earliest, ERNIE 1.0, was released in March 2019 and stands for Enhanced Representation through Knowledge Integration. Its key idea is knowledge-enhanced pretraining: entity- and phrase-level masking during pretraining forces the model to learn knowledge relations between words, not just surface co-occurrence of neighbouring characters. The ERNIE line later grew into a large family spanning retrieval, dialogue and multimodality, and underpins Baidu’s ERNIE Bot product. This entry records the ERNIE starting point as represented by the early 2019 versions.

Why it matters

It introduced entity- and phrase-level masking into pretraining so the model learns knowledge relations rather than mere character co-occurrence — an idea widely borrowed in early Chinese pretraining. These early versions have since been superseded by later ERNIE generations.

Key specs

Architecture
Knowledge-enhanced pretraining (knowledge masking)
Released
2019-03 (ERNIE 1.0)
Modality
Text
Open weights
No

Capabilities

Related concepts

Comparable products