o3
Suy luận dài trước khi trả lời, đổi tính toán lúc suy luận lấy độ chính xác ổn định hơn
Toàn văn của trang này được trình bày bằng tiếng Anh; tiêu đề và phần dẫn đã được bản địa hóa.
NÓ LÀ GÌ
o3 is a reasoning model OpenAI released in April 2025, part of the o series. Unlike models that answer directly, it performs an internal chain of reasoning before producing its output, then derives the answer from that reasoning. The series is trained with reinforcement learning so the model learns to spend more reasoning steps when they buy a more reliable answer. o3 accepts text and image input and serves as the main version for mathematics, coding and scientific reasoning tasks.
Vì sao đáng ghi nhớ
It turned test-time compute into a product: by reasoning at length before answering, o3 spends more compute at inference to buy steadier correctness, and made this the default shape of mainstream reasoning models.
Thông số chính
- Type
- Reasoning model
- Context window
- 200K tokens
- Modality
- Text, image in; text out
- Released
- 2025-04
Năng lực liên quan
Khái niệm liên quan
Học tăng cường từ phản hồi của con người
Khi «câu trả lời tốt» không viết được thành công thức, hãy để con người đóng vai hàm thưởng
Kỹ thuật prompt và căn chỉnh
Làm cho mô hình hữu ích, trung thực và vô hại khó hơn việc chỉ phóng to nó
Sản phẩm cùng loại
DeepSeek-R1
2025Mô hình suy luận huấn luyện bằng RL trên chuỗi suy nghĩ, trọng số mở theo giấy phép MIT
Gemini
2023Mô hình đa phương thức gốc, xử lý ngữ cảnh siêu dài
Grok
2023Mô hình hội thoại gắn với dữ liệu mạng xã hội
Claude
2023Mô hình hội thoại phổ thông nổi bật với ngữ cảnh dài và căn chỉnh an toàn
MiniMax-M
2025Mô hình suy luận trọng số mở với attention lai và ngữ cảnh một triệu token