HF Daily Papers
· Papers
N_0-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile Tokens
We present N_0-VTLA, a vision-tactile-language-action (VTLA) foundation model capable of (1) fine-grained contact-rich manipulation with tactile perception and tactile-feedback control, and (2) offline policy improvement from stored deployment data. Building on current vision-based backbones, we pro