Nearby in the stack

OmniVTLA: Vision-Tactile-Language-Action Models with Semantic-Aligned Tactile Sensing · arXivDesk