يُعرض النص الكامل باللغة الإنجليزية؛ وقد تمت ترجمة العنوان والملخص.
ما هو
Figure 02 is the second-generation humanoid robot Figure AI released in August 2024, designed for settings such as industry. It pairs a vision-language-action model with robot hardware so the robot can understand spoken instructions and carry out grasping and carrying tasks. Figure worked with OpenAI and demonstrated talking to the robot by voice and giving it tasks.
لماذا يستحق التذكّر
By connecting a vision-language-action model and a conversational speech model to a physical robot, it offers a direct demonstration of models entering the physical world.
المواصفات الأساسية
- Form
- Bipedal humanoid robot, battery-powered
- Perception
- Camera vision
- Control
- Vision-language-action model
المفاهيم ذات الصلة
التعلّم المعزّز العميق
دع شبكة عصبية تقرّر من البكسلات مباشرة، مع تثبيتها بأفكار قديمة
التوليد متعدد الوسائط
نموذج واحد يتعلّم الكلام والرسم والحركة، بل ونمذجة العالم ثلاثي الأبعاد
كشف الأجسام
من «ما الذي في الصورة» إلى «ماذا وأين وكم العدد»
السلامة والمواءمة وحقن الأوامر
لا يحسّن النموذج ما نريده فعلًا، بل مؤشرًا بديلًا كتبناه في دالة الخسارة — والفجوة بينهما هي مشكلة المواءمة كلها