Adaptive tactile fusion and joint visual-tactile prediction enable VLA models to handle contact-rich manipulation where vision alone fails, achieving 71% success on dexterous tasks.
DeCAL is a vision-language-action model for robot hands that combines visual and tactile (touch) sensing to perform complex manipulation tasks. It uses specialized AI experts working together to understand scenes, imagine future states, and generate actions—all while explicitly modeling contact dynamics that cause visual occlusions during dexterous manipulation.