मुख्य सामग्री पर जाएँ

Jamba

स्टेट-स्पेस मॉडल और Transformer को मिलाने वाला ओपन-वेट मॉडल

AI21 Labs मॉडल खुले वेट
इनपुटटेक्स्टटेक्स्ट

यह पृष्ठ अंग्रेज़ी में प्रस्तुत है; शीर्षक और सारांश का स्थानीयकरण किया गया है।

यह क्या है

Jamba is an open-weight model AI21 Labs released in March 2024, interleaving Mamba state-space layers with Transformer attention layers inside one mixture-of-experts architecture. AI21 Labs was founded in Tel Aviv in 2017 by Amnon Shashua and others. This hybrid design holds long context while cutting memory and compute relative to a pure-attention model. Jamba ships as a 52B-parameter MoE that activates about 12B per token, with a context window in the 256K range.

यह क्यों महत्वपूर्ण है

It replaced most attention layers with state-space layers, testing whether a non-pure-Transformer architecture can save compute at long context — a concrete instance of hybrid architectures entering open-weight models.

मुख्य विशिष्टताएँ

Parameters
52B (MoE, ~12B active)
Context window
256K tokens
Architecture
Mamba state-space layers interleaved with Transformer layers
Open weights
Yes

संबंधित क्षमताएँ

संबंधित अवधारणाएँ

समान उत्पाद