メインナビゲーションにスキップ 検索にスキップ メインコンテンツにスキップ

Visuo-Tactile Zero-Shot Object Recognition with Vision-Language Model

  • Shiori Ueda
  • , Atsushi Hashimoto
  • , Masashi Hamaya
  • , Kazutoshi Tanaka
  • , Hideo Saito

研究成果: Conference contribution

抄録

Tactile perception is vital, especially when distinguishing visually similar objects. We propose an approach to incorporate tactile data into a Vision-Language Model (VLM) for visuo-tactile zero-shot object recognition. Our approach leverages the zero-shot capability of VLMs to infer tactile properties from the names of tactilely similar objects. The proposed method translates tactile data into a textual description solely by annotating object names for each tactile sequence during training, making it adaptable to various contexts with low training costs. The proposed method was evaluated on the FoodReplica and Cube datasets, demonstrating its effectiveness in recognizing objects that are difficult to distinguish by vision alone.

本文言語English
ホスト出版物のタイトル2024 IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS 2024
出版社Institute of Electrical and Electronics Engineers Inc.
ページ7243-7250
ページ数8
ISBN(電子版)9798350377705
DOI
出版ステータスPublished - 2024
イベント2024 IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS 2024 - Abu Dhabi, United Arab Emirates
継続期間: 2024 10月 142024 10月 18

出版物シリーズ

名前IEEE International Conference on Intelligent Robots and Systems
ISSN(印刷版)2153-0858
ISSN(電子版)2153-0866

Conference

Conference2024 IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS 2024
国/地域United Arab Emirates
CityAbu Dhabi
Period24/10/1424/10/18

ASJC Scopus subject areas

  • 制御およびシステム工学
  • ソフトウェア
  • コンピュータ ビジョンおよびパターン認識
  • コンピュータ サイエンスの応用

フィンガープリント

「Visuo-Tactile Zero-Shot Object Recognition with Vision-Language Model」の研究トピックを掘り下げます。これらがまとまってユニークなフィンガープリントを構成します。

引用スタイル