論文の概要: ArtLang: Structured Language-to-Kinematics Grounding for Articulated 3D Actuation
- arxiv url: http://arxiv.org/abs/2608.15419v1
- Date: Sat, 15 Aug 2026 21:39:47 GMT
- ステータス: 翻訳完了
- システム内更新日: 2026-08-18 19:59:03.294269
- Title: ArtLang: Structured Language-to-Kinematics Grounding for Articulated 3D Actuation
- Title(参考訳): ArtLang:Articulated 3Dアクティベーションのための構造化言語とキネマティクスグラウンド
- Authors: Sylvia Yuan, Dan Wang, Ravi Ramamoorthi, Xinrui Cui,
- Abstract要約: 永続的再構成資産のオープン語彙制御のためのフレームワークであるArtLangを提案する。
ArtLangは、セマンティックな調音グラフとして資産を表し、その表面を言語の特徴とグラフ制約された動きで拡張する。
- 参考スコア(独自算出の注目度): 22.43942785482678
- License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
- Abstract: Articulated-object reconstructions recover explicit geometry and kinematics, but their parts often remain semantically anonymous and must be controlled through part indices and numerical joint parameters. We present ArtLang, a framework for open-vocabulary language control of persistent reconstructed articulated assets. ArtLang represents an asset as a semantic-kinematic articulation graph and augments its surface with language features and graph-constrained motion. Open-vocabulary proposals are bound to reconstructed parts while allowing uncertain parts to remain unnamed. A typed parser converts a command into a directive graph containing referring expressions, actions, magnitudes, reference frames, and relations. We then solve a global graph-to-graph grounding problem that jointly reasons about semantic, spatial, relational, and kinematic compatibility, with support for null assignments and abstention under ambiguity. Accepted directives are converted into continuous joint targets within the observed motion range and executed through forward kinematics. Experiments on synthetic reconstructions, mesh-based assets, and real captures demonstrate reliable language grounding and continuous articulated control across repeated parts, spatial references, relational commands, and ambiguous instructions.
- Abstract(参考訳): 人工物再構成は、明示的な幾何学と運動学を復元するが、それらの部分はしばしば意味的に匿名であり、部分的な指標と数値的な関節パラメータによって制御されなければならない。
永続的再構成された資産のオープン語彙制御のためのフレームワークであるArtLangを提案する。
ArtLangは、セマンティックな調音グラフとして資産を表し、その表面を言語の特徴とグラフ制約された動きで拡張する。
オープン語彙の提案は、不確実な部品が未命名のまま残ることを許しながら、再構成された部品に拘束される。
型付きパーサは、コマンドを参照式、アクション、サイズ、参照フレーム、リレーションを含むディレクティブグラフに変換する。
次に,意味的,空間的,リレーショナル,キネマティック整合性の両立を両立させるグローバルグラフとグラフの接地問題を,あいまいさの下でのヌル代入と棄権をサポートすることで解決する。
受容指令は、観測された運動範囲内で連続的な関節目標に変換され、前方運動学を通して実行される。
合成再構成、メッシュベースの資産、実際のキャプチャの実験は、繰り返し部分、空間参照、リレーショナルコマンド、あいまいな指示に対して、信頼性の高い言語基底と連続的な調音制御を示す。
関連論文リスト
- Articulated Object Reconstruction from Rest-State Observation [33.09154459685742]
インタラクティブなデジタル双生児を作るには、3D幾何学と、物体の関節の仕方を管理する運動構造の両方を復元する必要がある。
一つの閉じた構成から明瞭な物体を再構成する静止状態定式化を導入する。
論文 参考訳(メタデータ) (2026-07-30T06:42:03Z) - Beyond Isolated Objects: Relationship-aware Open Vocabulary Scene Understanding via 3D Scene Graph Analysis [59.527737790821305]
本稿では,3次元シーングラフを用いてオープンな3次元理解を促進する関係認識フレームワークを提案する。
本手法は,物体関係の推測に視覚言語推論を活用することで,多視点観測からリレーショナルシーングラフを構築する。
ScanNetV2、ScanNet200、ScanNet$++$、Replicaの実験は、強力なパフォーマンスと一般化能力を示している。
論文 参考訳(メタデータ) (2026-07-06T17:29:49Z) - UnfoldArt: Zero-Shot Recovery of Full Articulated 3D Objects from Text or Image [58.40343808003224]
本稿では,テキストや画像からの3Dオブジェクトの再構成について,議論を巻き起こしたアプローチを提案する。
高レベルのエージェントは、視覚言語やビデオモデルからの知識を用いて、オブジェクトの意味論と動きを推論する。
低レベルエージェントは調音パラメータと相互作用点を推定する。
同じビデオは、合意された調律に条件付けされ、各部分をその動きを通して駆動し、隠された内部と幾何学を露呈する。
論文 参考訳(メタデータ) (2026-06-29T17:44:53Z) - ReLaGS: Relational Language Gaussian Splatting [20.136674901612334]
本稿では,階層型言語で区切られたガウシアンシーンと,シーン固有の訓練を伴わない3Dセマンティックシーングラフを構築する新しいフレームワークを提案する。
この階層の上に、視覚言語由来のアノテーションとグラフニューラルネットワークに基づくリレーショナル推論を備えたオープンな3Dシーングラフを構築します。
本手法は,階層的セマンティクスとオブジェクト間の相互関係を共同でモデル化することにより,効率的でスケーラブルな3次元推論を可能にする。
論文 参考訳(メタデータ) (2026-03-18T11:18:23Z) - Beyond Global Alignment: Fine-Grained Motion-Language Retrieval via Pyramidal Shapley-Taylor Learning [56.6025512458557]
動き言語検索は、自然言語と人間の動きの間の意味的ギャップを埋めることを目的としている。
既存のアプローチは主に、全動作シーケンスとグローバルテキスト表現の整合性に重点を置いている。
本研究では,微粒な動き言語検索のためのPST学習フレームワークを提案する。
論文 参考訳(メタデータ) (2026-01-29T16:00:12Z)
関連論文リストは本サイト内にある論文のタイトル・アブストラクトから自動的に作成しています。
指定された論文の情報です。
本サイトの運営者は本サイト(すべての情報・翻訳含む)の品質を保証せず、本サイト(すべての情報・翻訳含む)を使用して発生したあらゆる結果について一切の責任を負いません。