मुख्य सामग्री पर जाएँ

MiniMax-M

हाइब्रिड अटेंशन और दस लाख टोकन संदर्भ वाला ओपन-वेट तर्क मॉडल

MiniMax मॉडल खुले वेट
इनपुटटेक्स्टटेक्स्ट

यह पृष्ठ अंग्रेज़ी में प्रस्तुत है; शीर्षक और सारांश का स्थानीयकरण किया गया है।

यह क्या है

MiniMax-M1 is an open-weight reasoning model MiniMax released in June 2025. MiniMax was founded in Shanghai in 2021 by Yan Junjie and others. M1 uses a hybrid attention design that combines standard full-attention layers with linear-attention layers to cut the compute cost of long-sequence reasoning, and its stated context window reaches the million-token range. With 456B total parameters activating about 46B per token, it is a large-scale MoE reasoning model whose weights are released under an open licence. The company also runs generation product lines such as Hailuo.

यह क्यों महत्वपूर्ण है

It mixes linear and full attention and pushes context to a million tokens to cut the cost of long-sequence reasoning; among the larger open-weight reasoning models, it is a concrete landing of the open-weight camp on the reasoning track.

मुख्य विशिष्टताएँ

Parameters
456B (MoE, ~46B active)
Context window
1M tokens
Architecture
Hybrid linear and full attention
Open weights
Yes

संबंधित क्षमताएँ

संबंधित अवधारणाएँ

समान उत्पाद

DeepSeek-R1

2025
DeepSeek

तर्क-श्रृंखला पर RL से प्रशिक्षित तर्क मॉडल, भार MIT लाइसेंस के तहत खुले

मॉडल खुले वेट
टेक्स्टटेक्स्ट

Qwen

2023
Alibaba (Qwen)

अनेक आकारों और बहुविध संस्करणों वाला ओपन-वेट परिवार

मॉडल खुले वेट
टेक्स्टइमेजटेक्स्ट

o3

2025
OpenAI

उत्तर देने से पहले लंबी तर्क-श्रृंखला, अनुमान के समय गणना से अधिक सटीकता

मॉडल बंद स्रोत
टेक्स्टइमेजटेक्स्ट

GLM

2023
Zhipu AI

ऑटोरेग्रेसिव ब्लैंक-भरने वाले प्रीट्रेनिंग से शुरू हुआ चीनी सामान्य मॉडल

मॉडल बंद स्रोत
टेक्स्टटेक्स्ट

Doubao

2023
ByteDance (Seed)

बाइटडांस का सामान्य संवाद मॉडल और ऐप

मॉडल बंद स्रोत
टेक्स्टइमेजटेक्स्ट

Step

2023
StepFun

बहुविधता और ऑन-डिवाइस उपयोग पर लक्षित सामान्य मॉडल परिवार

मॉडल बंद स्रोत
टेक्स्टइमेजटेक्स्ट