Figure 02
दृष्टि-भाषा-क्रिया मॉडल से चलने वाला दूसरी पीढ़ी का मानव-सदृश रोबोट
यह पृष्ठ अंग्रेज़ी में प्रस्तुत है; शीर्षक और सारांश का स्थानीयकरण किया गया है।
यह क्या है
Figure 02 is the second-generation humanoid robot Figure AI released in August 2024, designed for settings such as industry. It pairs a vision-language-action model with robot hardware so the robot can understand spoken instructions and carry out grasping and carrying tasks. Figure worked with OpenAI and demonstrated talking to the robot by voice and giving it tasks.
यह क्यों महत्वपूर्ण है
By connecting a vision-language-action model and a conversational speech model to a physical robot, it offers a direct demonstration of models entering the physical world.
मुख्य विशिष्टताएँ
- Form
- Bipedal humanoid robot, battery-powered
- Perception
- Camera vision
- Control
- Vision-language-action model
संबंधित अवधारणाएँ
गहन सुदृढ़ीकरण अधिगम
तंत्रिका जाल को सीधे पिक्सेल से निर्णय लेने दें, पुराने विचारों से स्थिर करें
बहु-मॉडल जनरेशन
एक ही मॉडल बोलना, चित्र बनाना, हिलना, और यहाँ तक कि 3D संसार का नमूना बनाना सीखता है
वस्तु संसूचन
“चित्र में क्या है” से “क्या, कहाँ और कितने” तक
सुरक्षा, संरेखण और प्रॉम्प्ट इंजेक्शन
मॉडल वह अनुकूलित करता है जो हमने loss में लिखा, वह नहीं जो हम वास्तव में चाहते हैं — यही अंतराल संरेखण समस्या का पूरा सार है