How a 450M-parameter Vision-Language-Action model learns to coordinate two robot arms for fabric folding — from just 50 human demonstrations.
Teaching Robots to Fold Clothes: SmolVLA for…
How a 450M-parameter Vision-Language-Action model learns to coordinate two robot arms for fabric folding — from just 50 human demonstrations.