Chuyển đến nội dung
Bản đồ AI

Học tăng cường

Học chuỗi quyết định từ thử và sai cùng phần thưởng trễ

OVERVIEW

Tất cả mục từ
6
Cơ bản
2
Trung cấp
3
Chuyên gia
1

Reinforcement learning addresses problems that have no answer key, only consequences: whether a step was good may only be revealed much later as reward. It introduces the vocabulary of agent, environment, state, action and reward, and approaches optimal behaviour through two families of ideas — value functions and policy gradients.