Nearby in the stack

FM-VLA: Force-based Memory for Vision-Language-Action Models in Contact-Rich Manipulation · arXivDesk