Chuyển đến nội dung
Bản đồ AI

GLOSSARY

Bảng thuật ngữ

Mọi thuật ngữ chính nằm rải rác trong các mục từ, được tập hợp thành một mục lục tra cứu.

tổng cộng 193 thuật ngữ

3 1

3D Gaussian Splatting

Representing a scene with many 3D Gaussian ellipsoids for fast rendering

TừSinh đa phương thức

A 9

Accelerator

A high-throughput parallel unit such as a GPU or TPU

TừHạ tầng huấn luyện và suy luận
Action A

What the agent can do; either discrete or continuous

TừQuá trình quyết định Markov
Action value Q(s, a)

Expected discounted return after forcing the first action to be a

TừHàm giá trị và Q-learning
Activation function

A function that applies a nonlinear transform to the weighted sum

TừNơ-ron và perceptron
Actor / Critic

The policy network and the value network: one acts, one scores

TừGradient chính sách
Advantage A(s, a)

How much better an action is than the average at that state

TừGradient chính sách
Alignment

Making model behaviour match human intent and values

TừAn toàn, căn chỉnh và chèn lệnh (prompt injection)
Anomaly detection

Finding the few samples that deviate from the bulk distribution

TừHọc không giám sát
Automatic differentiation

Letting a framework compute exact gradients automatically, not by numerical approximation

TừLan truyền ngược

B 8

Batch size

How many samples estimate the gradient per step

TừGradient và hạ gradient
Bias

How far the model’s average prediction departs from the true regularity

TừĐánh đổi độ lệch – phương sai
Bias

A learnable offset applied to the threshold

TừNơ-ron và perceptron
Bit depth

How many bits encode each channel; 8 bits give 256 levels

TừBiểu diễn số của ảnh
Bottleneck

The low-dimensional layer holding the latent code, limiting its bandwidth

TừAutoencoder và VAE
Bounding box

A rectangle represented as (x, y, w, h) or corner points

TừPhát hiện đối tượng
BPE

Byte-Pair Encoding: bottom-up merging of frequent symbol pairs

TừToken hóa
BPTT

Backpropagation through time after unrolling

TừMạng nơ-ron hồi tiếp

C 20

Catastrophic forgetting

Rapid loss of old abilities while learning a new task

TừTiền huấn luyện và tinh chỉnh
Chain rule

The derivative of a composition is the product of the local derivatives

TừLan truyền ngược
Chain-of-thought (CoT)

Making the model write out intermediate reasoning steps

TừKỹ thuật prompt và căn chỉnh
Channel

A distinct measurement at the same location, such as R/G/B or alpha

TừBiểu diễn số của ảnh
Chunking

Splitting long documents into retrievable pieces

TừSinh văn bản tăng cường truy xuất
Classifier-free guidance

Extrapolating between conditional and unconditional predictions to control prompt fidelity

TừMô hình khuếch tán
Clustering

Grouping samples by similarity (k-means, hierarchical clustering)

TừHọc không giám sát
Colour space

A coordinate system for colour values, such as sRGB, HSV or Lab

TừBiểu diễn số của ảnh
Computation graph

A computation expressed as nodes and directed edges over which derivatives propagate

TừLan truyền ngược
Confusion matrix

A cross-tabulation of true versus predicted classes

TừĐánh giá mô hình và kiểm định chéo
Continuous batching

Re-forming the batch every step to keep the GPU busy

TừTối ưu suy luận và phục vụ
Contrastive learning

Learning representations by pulling positives together and pushing negatives apart

TừThị giác tự giám sát và học tương phản đa phương thức
Contrastive loss

A loss that pulls same-class embeddings together and pushes different-class ones apart

TừHàm mất mát
ControlNet

A bypass network guiding structure from a condition map

TừKhuếch tán trong không gian ẩn và điều khiển có điều kiện
Cosine similarity

The alignment of two vector directions, from −1 to 1

TừNhúng từ (Word Embeddings)
Cross-attention

Query from one sequence, Key/Value from another

TừCơ chế chú ý (Attention)
Cross-attention

The attention mechanism letting image features query text vectors

TừKhuếch tán trong không gian ẩn và điều khiển có điều kiện
Cross-entropy

The information needed to encode data from P using distribution Q

TừEntropy và lý thuyết thông tin
Cross-entropy

Negative log-probability of the correct class; the default classification loss

TừHàm mất mát
Cross-entropy loss

The standard objective for classification training

TừPhân loại ảnh

D 12

DDIM

Deterministic sampling achieving comparable quality in a few dozen steps

TừMô hình khuếch tán
DDPM

Discrete Markov diffusion, typically needing a thousand sampling steps

TừMô hình khuếch tán
Degradation problem

Deeper networks with higher training error, and not from overfitting

TừChuẩn hóa và kết nối dư
Density estimation

Estimating the probability distribution the data follows

TừHọc không giám sát
Dimension

The number of entries in a vector

TừVector và không gian vector
Dimensionality reduction

Compressing high-dimensional data to fewer dimensions while preserving structure (PCA, t-SNE, UMAP)

TừHọc không giám sát
Discount factor γ

Between 0 and 1; how much future rewards are valued

TừQuá trình quyết định Markov
Discriminator

The network judging real versus fake, serving as the loss

TừMạng đối kháng tạo sinh
Double descent

The modern counterexample where test error falls again past the interpolation point

TừĐánh đổi độ lệch – phương sai
DPO

Direct preference optimisation without an explicit reward model

TừKỹ thuật prompt và căn chỉnh
Dropout

Randomly silencing units during training to prevent co-adaptation

TừQuá khớp và điều chuẩn
Dying ReLU

A neuron stuck in the negative region with zero gradient, no longer updating

TừHàm kích hoạt

E 10

Early stopping

Halting training before validation loss turns upward

TừQuá khớp và điều chuẩn
Eigenvector / eigenvalue

A vector whose direction is unchanged by the map, and the factor by which it is scaled

TừPhép toán ma trận và biến đổi tuyến tính
ELBO

A lower bound on the log-likelihood: the reconstruction term minus the KL term; a VAE’s actual objective

TừAutoencoder và VAE
Embedding

The layer, or its output, that maps a discrete object into a continuous vector

TừVector và không gian vector
Embedding model

A model that encodes text into vectors

TừSinh văn bản tăng cường truy xuất
Empirical risk

The model’s average loss on the training samples

TừHọc có giám sát
Equivariance

When the input shifts, the output shifts accordingly rather than changing

TừMạng nơ-ron tích chập
Evidence

The total probability of the data across all hypotheses; it normalises the result

TừĐịnh lý Bayes
Experience replay

Store past transitions and sample randomly to break correlation

TừHọc tăng cường sâu
Explicit density

A model that writes down or approximates p(x), e.g. autoregressive or diffusion

TừTổng quan về mô hình sinh

F 3

F1

The harmonic mean of precision and recall

TừĐánh giá mô hình và kiểm định chéo
FID

Fréchet distance between generated and real distributions in Inception feature space; lower is better

TừTổng quan về mô hình sinh
Function calling

The model emitting structured arguments to invoke an external function

TừTác tử và sử dụng công cụ

G 5

Gating

Using 0–1 coefficients from Sigmoid to control how much information passes

TừMạng nơ-ron hồi tiếp
Generalisation

Performance on data the model has not seen

TừQuá khớp và điều chuẩn
Generator

The network mapping noise to samples

TừMạng đối kháng tạo sinh
Gradient flow

The magnitude and stability of gradients as they propagate layer by layer

TừLan truyền ngược
Guardrail

Checks and constraints bounding what an agent may do

TừTác tử và sử dụng công cụ

H 3

Hidden state

A continuously updated "summary so far" vector

TừMạng nơ-ron hồi tiếp
Hinge loss

Requires the correct class to win by a margin; the heart of the SVM

TừHàm mất mát
Hypothesis space

The set of all functions the model can represent

TừHọc có giám sát

I 9

Identity shortcut

The path in a residual connection that adds the input straight back to the output

TừChuẩn hóa và kết nối dư
Implicit density

A model that offers only a sampler, not a probability, e.g. a GAN

TừTổng quan về mô hình sinh
In-context learning

Solving a task from prompt examples without updating parameters

TừKỹ thuật prompt và căn chỉnh
InfoNCE

The standard contrastive loss; essentially a multi-class cross-entropy

TừThị giác tự giám sát và học tương phản đa phương thức
Inner product

Element-wise product summed over entries; the numerator of cosine similarity

TừVector và không gian vector
Input x

The feature vector fed to the model

TừHọc có giám sát
Internal covariate shift

The shifting distribution of inputs to later layers during training

TừChuẩn hóa và kết nối dư
IoU

The ratio of the intersection to the union of two boxes

TừPhát hiện đối tượng
Irreducible error

The unavoidable error floor caused by label noise

TừĐánh đổi độ lệch – phương sai

J 1

Jailbreak

Inducing a model past its safety training

TừAn toàn, căn chỉnh và chèn lệnh (prompt injection)

K 7

Kernel / filter

A set of learnable weights that slides over the input

TừMạng nơ-ron tích chập
Kernel / filter

The small weight matrix that is learned

TừPhép tích chập
KL divergence

Cross-entropy minus true entropy; non-negative and asymmetric

TừEntropy và lý thuyết thông tin
KL divergence

Measures how far the encoded distribution deviates from a standard normal; acts as a regulariser

TừAutoencoder và VAE
KL penalty

Penalises divergence from the reference policy to prevent degeneration

TừHọc tăng cường từ phản hồi của con người
Knowledge distillation

Training a small model on a large model’s soft outputs

TừNén mô hình
KV cache

Caching past tokens’ keys and values to avoid recomputation

TừTối ưu suy luận và phục vụ

L 12

Label y

The correct output for each sample; the source of supervision

TừHọc có giám sát
Latent space

The low-dimensional representation space produced by the autoencoder

TừKhuếch tán trong không gian ẩn và điều khiển có điều kiện
Learning rate η

How far each step moves

TừGradient và hạ gradient
Likelihood

The probability of the observed data under given parameters

TừXác suất và phân phối xác suất
Likelihood

The probability of observed data given that the hypothesis is true

TừĐịnh lý Bayes
Linearly separable

A hyperplane exists that separates the two classes perfectly

TừNơ-ron và perceptron
Log-derivative trick

Turns the gradient of an expectation into a weighted sum of log-probabilities

TừGradient chính sách
Long-range dependency

Influence between elements far apart in a sequence

TừMạng nơ-ron hồi tiếp
LoRA

Low-rank adapters training only a tiny number of new parameters

TừTiền huấn luyện và tinh chỉnh
LoRA

Low-rank adaptation increments for low-cost customisation

TừKhuếch tán trong không gian ẩn và điều khiển có điều kiện
Loss surface

The high-dimensional terrain of loss values over parameter space

TừGradient và hạ gradient
Low-rank factorisation

Approximating a large matrix by a product of two smaller ones

TừNén mô hình

M 9

mAP

Mean average precision across classes and IoU thresholds

TừPhát hiện đối tượng
Masked language modelling

Hide random words and recover them, a bidirectional objective

TừTiền huấn luyện và tinh chỉnh
MCTS

An algorithm that evaluates moves via sampled rollouts to guide search

TừHọc tăng cường sâu
Mean squared error (MSE)

The average squared difference between prediction and label; the default regression loss

TừHàm mất mát
mIoU

The mean of per-class IoU, the primary segmentation metric

TừPhân đoạn ngữ nghĩa
Mode collapse

When a generator covers only a few modes of the data distribution

TừTổng quan về mô hình sinh
Mode collapse

The generator covers few modes and loses diversity

TừMạng đối kháng tạo sinh
Multi-armed bandit

The simplest sequential model: unknown reward distributions, one pull per round

TừKhám phá và khai thác
Multi-head attention

Several attentions in parallel, each learning a different focus

TừCơ chế chú ý (Attention)

N 5

Negative sampling

Replacing full-vocabulary softmax with a few random negatives

TừNhúng từ (Word Embeddings)
NeRF

A neural network representing a scene’s radiance field for novel-view synthesis

TừSinh đa phương thức
NMS

Non-maximum suppression, removing duplicate boxes

TừPhát hiện đối tượng
Noise schedule

The timetable of noise added per step, described by βₜ or ᾱₜ

TừMô hình khuếch tán
Norm

A function measuring a vector’s "length"; L2 is the common choice

TừVector và không gian vector

O 3

Off-policy

The behaviour policy may differ from the policy being learned

TừHàm giá trị và Q-learning
One-hot

A sparse vector with a single 1; distinct words are fully orthogonal

TừNhúng từ (Word Embeddings)
Out-of-vocabulary (OOV)

A word absent from the vocabulary, spelled out from subwords

TừToken hóa

P 16

Padding

Adding zeros at the border to control output size

TừPhép tích chập
Perplexity

The exponential of the cross-entropy; the effective number of options the model hesitates among per step

TừEntropy và lý thuyết thông tin
Pipeline parallelism

Placing different layers on different devices and filling bubbles with micro-batches

TừHạ tầng huấn luyện và suy luận
Pixel

The smallest sampling unit of an image, carrying one or more channel values

TừBiểu diễn số của ảnh
Policy π

A mapping from states to actions, or to a distribution over actions

TừQuá trình quyết định Markov
Positional encoding

An explicit order signal, sinusoidal or RoPE

TừKiến trúc Transformer
Posterior

The updated degree of belief after incorporating the evidence

TừĐịnh lý Bayes
Pre-activation

A layout placing normalisation before the convolution, which trains more stably

TừChuẩn hóa và kết nối dư
Pre-LN

Placing layer norm before each sublayer for stability

TừKiến trúc Transformer
Precision & recall

Precision asks how many alerts are real; recall asks how many real cases were caught

TừĐánh giá mô hình và kiểm định chéo
Preference pair

Two candidate outputs for one input plus the human’s choice between them

TừHọc tăng cường từ phản hồi của con người
Prior

The degree of belief in a hypothesis before seeing data

TừĐịnh lý Bayes
Probability density

The "thickness" of probability for a continuous variable; its integral over an interval is the probability

TừXác suất và phân phối xác suất
Projection head

The MLP the contrastive loss is applied to, usually discarded after training

TừThị giác tự giám sát và học tương phản đa phương thức
Prompt injection

Smuggling malicious instructions as data for the model to follow

TừAn toàn, căn chỉnh và chèn lệnh (prompt injection)
Pruning

Removing low-impact weights or whole structures

TừNén mô hình

Q 2

Quantisation

Representing float weights and activations with low-bit integers

TừNén mô hình
Query / Key / Value

The three vector roles: what you seek, what is on offer, what is carried

TừCơ chế chú ý (Attention)

R 14

Random variable

A function mapping outcomes of a random experiment to numbers

TừXác suất và phân phối xác suất
Rank

The number of independent directions the map actually spans; at most rows or columns

TừPhép toán ma trận và biến đổi tuyến tính
ReAct

A prompting paradigm alternating reasoning and action

TừTác tử và sử dụng công cụ
Receptive field

The region of the original input that a given output covers

TừMạng nơ-ron tích chập
Receptive field

The input region that one output pixel depends on

TừPhép tích chập
Red teaming

Actively hunting for failure and misuse paths

TừAn toàn, căn chỉnh và chèn lệnh (prompt injection)
Regret

The gap between realised cumulative reward and always picking the best arm

TừKhám phá và khai thác
Reparameterisation

Writing sampling as a deterministic transform plus external noise so gradients flow

TừAutoencoder và VAE
Reranking

Rescoring candidate passages with a more accurate model

TừSinh văn bản tăng cường truy xuất
Residual connection

Adding the input past a sublayer to ease vanishing gradients in depth

TừKiến trúc Transformer
Reward hacking

Exploiting the proxy reward instead of genuinely completing the task

TừHọc tăng cường từ phản hồi của con người
Reward model

A model fitting human preferences and emitting a differentiable score

TừHọc tăng cường từ phản hồi của con người
ROC-AUC

Area under the ROC curve, measuring ranking ability across all thresholds

TừĐánh giá mô hình và kiểm định chéo
RoPE

Rotary Position Embedding: relative position with better extrapolation

TừKiến trúc Transformer

S 19

Saturation

A function whose derivative tends to 0 at the extremes, blocking gradients

TừHàm kích hoạt
Self-attention

Attention whose Q, K and V all come from one sequence

TừCơ chế chú ý (Attention)
Self-information

The information of a single event, −log p

TừEntropy và lý thuyết thông tin
Self-play

Generating training data by having an agent play against its past selves

TừHọc tăng cường sâu
Self-supervised

Labels manufactured from the data itself, no manual annotation

TừTiền huấn luyện và tinh chỉnh
Semantic / instance / panoptic

Class → class + instance → the two unified

TừPhân đoạn ngữ nghĩa
SentencePiece

A subword toolkit that runs directly on the character/byte stream

TừToken hóa
SFT

Supervised fine-tuning on instruction–response pairs

TừKỹ thuật prompt và căn chỉnh
SGD

Approximating the full gradient with a mini-batch

TừGradient và hạ gradient
Shannon entropy

The average information, or uncertainty, of a random variable

TừEntropy và lý thuyết thông tin
Singular value decomposition

Writing any matrix as the product "rotate · stretch · rotate"

TừPhép toán ma trận và biến đổi tuyến tính
Skip connection

Routing shallow high-resolution features into deep layers to preserve boundaries

TừPhân đoạn ngữ nghĩa
Skip-gram

A training objective that predicts surrounding words from the centre

TừNhúng từ (Word Embeddings)
Softmax

Turns a set of real scores into a probability distribution summing to 1

TừXác suất và phân phối xác suất
Spatiotemporal patch

A local unit spanning frames in video, used to model motion

TừSinh đa phương thức
Speculative decoding

A small model drafts and the large model verifies in parallel to speed up generation

TừTối ưu suy luận và phục vụ
State S

The variables describing the present situation; must satisfy the Markov property

TừQuá trình quyết định Markov
State value V(s)

Expected discounted return from s under policy π

TừHàm giá trị và Q-learning
Stride

How many pixels the window jumps each step

TừPhép tích chập

T 11

Target network

A slowly updated copy of the network providing stable bootstrap targets

TừHọc tăng cường sâu
TD error

The gap between the fresh target and the old estimate

TừHàm giá trị và Q-learning
Tensor parallelism

Splitting a single layer’s large matrices across devices

TừHạ tầng huấn luyện và suy luận
Thompson sampling

Sample from the posterior and pick the max, auto-directing exploration to uncertainty

TừKhám phá và khai thác
Time to first token (TTFT)

Time from sending a request to receiving the first token

TừTối ưu suy luận và phục vụ
Tool

One external capability an agent may invoke

TừTác tử và sử dụng công cụ
Top-1 / top-5 error

Whether the top prediction / top five include the true label

TừPhân loại ảnh
Transfer learning

Pre-train on a large dataset, then fine-tune on a small task

TừPhân loại ảnh
Transpose

Flip a matrix across its diagonal so rows become columns

TừPhép toán ma trận và biến đổi tuyến tính
Transposed convolution

An upsampling operation common in segmentation decoders

TừPhân đoạn ngữ nghĩa
Trust region / KL constraint

Bounds how far the new policy may drift from the old

TừGradient chính sách

V 6

Vanishing gradient

Gradients shrinking exponentially as they are multiplied across layers

TừHàm kích hoạt
Variance

How sensitive the model is to perturbations of the training set

TừĐánh đổi độ lệch – phương sai
Vector database

A store providing nearest-neighbour search over high-dimensional vectors

TừSinh văn bản tăng cường truy xuất
ViT

An architecture that applies a Transformer to image patches

TừPhân loại ảnh
Vocabulary

The fixed set of all tokens and their indices

TừToken hóa
Vocoder

The component that turns acoustic features back into a waveform

TừSinh đa phương thức

W 4

Wasserstein distance

An earth-mover distance between distributions, better behaved for training than JS divergence

TừMạng đối kháng tạo sinh
Weight

How strongly an input influences the output; may be positive or negative

TừNơ-ron và perceptron
Weight decay (L2)

Adding a squared-weight penalty to the loss to suppress large weights

TừQuá khớp và điều chuẩn
Weight sharing

Reusing one set of weights across all spatial positions

TừMạng nơ-ron tích chập

Z 3

ZeRO

Sharding optimiser states, gradients and parameters to cut per-device memory

TừHạ tầng huấn luyện và suy luận
Zero-centred

Outputs symmetric about 0, which aids optimisation

TừHàm kích hoạt
Zero-shot classification

Classifying directly with text prompts, without fine-tuning

TừThị giác tự giám sát và học tương phản đa phương thức

Ε 1

ε-greedy

Explore at random with probability ε, exploit the current best otherwise

TừKhám phá và khai thác