Back to Research papers
Research paper index

The Potential of Haptic Foundation Models

Jianquan Wang, Haiwei Dong, Abdulmotaleb El Saddik

arXiv:2608.28664Published August 23, 20260 citations
  • cs.RO
  • cs.CV
  • cs.MM
  • trajectory
  • action
  • robot
  • foundation model
  • embodied

Abstract

Despite the success of foundation models in language and vision, their expansion into embodied AI is bottlenecked by a lack of generalized touch sensing. This limitation is especially relevant to consumer electronics, where smartphones, wearables, VR controllers, home robots, and health monitoring devices require safe and adaptive physical interaction. Constrained by hardware heterogeneity and the necessity of active physical data collection, current haptic models remain rigidly task-specific. To overcome these limitations, this article explores the transformative potential and developmental trajectory of Haptic Foundation Models (HFMs). We detail the paradigm shift required to transition from passive Large Language Models and Vision Language Models into active HFMs across four core dimensions: action coupling, physical dynamical representation space, continuous time-series data granularity, and action-conditioned future state prediction. Furthermore, we synthesize existing large-scale tactile datasets and benchmark UniTouch, AnyTouch, T3, and Sparsh on TacBench for force estimation, slip detection, and relative pose estimation.

Read the original paper

This page indexes public paper metadata. The manuscript remains with its original publisher and authors.