What problem does it solve?
This Skill provides an efficient Vision-Language-Action policy for robotic manipulation tasks, enabling robots to perform tasks like picking up objects, placing them, and manipulating them in a controlled manner.
Core Features & Use Cases
- Vision-Language-Action: Combines vision, language, and action to enable robots to understand and execute tasks described in natural language.
- SmolVLA Model: Utilizes the SmolVLA model, which is finetuned on the LIBERO task suite for efficient performance.
- Embodiment Support: Supports various robot embodiments, including Franka Panda, SO-100 follower arm, and any 6-7 DOF manipulator.
- Use Case: Imagine you have a robot equipped with a camera and a gripper. Use this Skill to instruct the robot to pick up a red cube from a table and place it in a drawer using natural language commands.
Quick Start
Use the rskill-smolvla-libero skill to perform a task on the Franka Panda robot.