Researchers introduce T-Rex, a framework that unifies vision, language, and tactile sensing so robots can respond to physical contact in real time rather than relying on vision alone
Summary
Researchers introduced T-Rex, a framework that integrates vision, language, and tactile sensing, enabling robots to respond to physical contact in real time rather than relying solely on vision.
Similar Articles
@rohanpaul_ai: A lot of embodied AI still feels like AI modules bolted onto a robot. TARS is taking a different architectural bet with…
TARS launches AWE 3.5, an embodied-native foundation model that integrates action, perception, geometry, and touch into one model for general-purpose physical AI, with claims of 2x task execution efficiency over PI0.5.
N_0-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile Tokens
Introduces N_0-VTLA, a vision-tactile-language-action foundation model for contact-rich manipulation, featuring large-scale tactile pretraining and advantage-conditioned offline policy improvement, with strong results on real-robot and simulation benchmarks.
OmniTacTune: Policy-Agnostic Real-World RL for Tactile Residual Adaptation of Visual Policies
OmniTacTune introduces a two-stage reinforcement learning pipeline for adapting tactile feedback to pretrained visual robot policies, achieving 85-100% success on contact-rich manipulation tasks within 40-80 minutes.
'Touch dreaming' helps humanoid robots handle five tricky tasks with 90.9% higher success
Researchers from CMU and Bosch Center for AI introduced the Humanoid Transformer with Touch Dreaming (HTD) model, which uses tactile signal prediction to improve humanoid robot manipulation, achieving a 90.9% higher average success rate over the ACT baseline across five real-world tasks.
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination
RxBrain is an embodied cognition foundation model that jointly reasons with language and visual imagination to represent embodied plans, using a unified multimodal Mixture-of-Transformers architecture. It achieves promising real-robot performance without large-scale action data.