Fugu-MT 論文翻訳(概要): HMR-1: Hierarchical Massage Robot with Vision-Language-Model for Embodied Healthcare

論文の概要: HMR-1: Hierarchical Massage Robot with Vision-Language-Model for Embodied Healthcare

arxiv url: http://arxiv.org/abs/2603.08817v1
Date: Mon, 09 Mar 2026 18:17:33 GMT
ステータス: 翻訳完了
システム内更新日: 2026-03-23 08:17:42.121661
Title: HMR-1: Hierarchical Massage Robot with Vision-Language-Model for Embodied Healthcare
Title（参考訳）: HMR-1:身体医療のための視覚言語モデルを用いた階層型マッサージロボット
Authors: Rongtao Xu, Mingming Yu, Xiaofeng Han, Yu Zhang, Kaiyi Hu, Zhe Feng, Zenghuang Fu, Changwei Wang, Weiliang Meng, Xiaopeng Zhang,
Abstract要約: 身体知性は医療、特に理学療法やリハビリテーションにおいて変革の機会を開いている。我々は、12,190の画像と174,177のQAペアを含むマルチモーダルデータセットを構築し、様々な照明条件と背景をカバーした。本稿では,ハイレベルなアキューポイント接地モジュールと低レベルな制御モジュールを備えた階層型エンボディマッサージフレームワークを提案する。
参考スコア（独自算出の注目度）: 28.230151467353647
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Abstract: The rapid advancement of Embodied Intelligence has opened transformative opportunities in healthcare, particularly in physical therapy and rehabilitation. However, critical challenges remain in developing robust embodied healthcare solutions, such as the lack of standardized evaluation benchmarks and the scarcity of open-source multimodal acupoint massage datasets. To address these gaps, we construct MedMassage-12K - a multimodal dataset containing 12,190 images with 174,177 QA pairs, covering diverse lighting conditions and backgrounds. Furthermore, we propose a hierarchical embodied massage framework, which includes a high-level acupoint grounding module and a low-level control module. The high-level acupoint grounding module uses multimodal large language models to understand human language and identify acupoint locations, while the low-level control module provides the planned trajectory. Based on this, we evaluate existing MLLMs and establish a benchmark for embodied massage tasks. Additionally, we fine-tune the Qwen-VL model, demonstrating the framework's effectiveness. Physical experiments further confirm the practical applicability of the framework.Our dataset and code are publicly available at https://github.com/Xiaofeng-Han-Res/HMR-1.
Abstract（参考訳）: エボディード・インテリジェンス(Embodied Intelligence)の急速な進歩は、医療、特に理学療法やリハビリテーションにおいて変革の機会を開いた。しかしながら、標準化された評価ベンチマークの欠如や、オープンソースのマルチモーダル・アキューポイント・マッサージデータセットの不足など、堅牢な医療ソリューションの開発において重要な課題が残っている。これらのギャップに対処するため、MedMassage-12Kという、12,190の画像と174,177のQAペアを含むマルチモーダルデータセットを構築し、様々な照明条件と背景をカバーした。さらに,高レベルアキューポイント接地モジュールと低レベル制御モジュールを含む階層型エンボディマッサージフレームワークを提案する。高レベルアキューポイントグラウンドモジュールは、多モードの大規模言語モデルを使用して、人間の言語を理解し、アキューポイントの位置を特定する。そこで本研究では,既存のMLLMを評価し,マッサージタスクを具体化するためのベンチマークを構築した。さらに、Qwen-VLモデルを微調整し、フレームワークの有効性を示す。我々のデータセットとコードはhttps://github.com/Xiaofeng-Han-Res/HMR-1.comで公開されている。

論文の概要: HMR-1: Hierarchical Massage Robot with Vision-Language-Model for Embodied Healthcare

関連論文リスト