Fugu-MT 論文翻訳(概要): Exploring High-Order Self-Similarity for Video Understanding

論文の概要: Exploring High-Order Self-Similarity for Video Understanding

arxiv url: http://arxiv.org/abs/2604.20760v1
Date: Wed, 22 Apr 2026 16:48:43 GMT
ステータス: 翻訳完了
システム内更新日: 2026-04-23 15:36:11.237866
Title: Exploring High-Order Self-Similarity for Video Understanding
Title（参考訳）: 映像理解のための高次自己相似性探索
Authors: Manjin Kim, Heeseung Kwon, Karteek Alahari, Minsu Cho,
Abstract要約: MOSS(Multi-Order Self-Similarity)は、マルチオーダーSTSS機能の学習と統合を目的とした軽量ニューラルネットワークモジュールである。多様なビデオタスクに適用することで、限界計算コストとメモリ使用量のみを消費しながら、モーションモデリング能力を向上させることができる。
参考スコア（独自算出の注目度）: 55.52840327834189
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Abstract: Space-time self-similarity (STSS), which captures visual correspondences across frames, provides an effective way to represent temporal dynamics for video understanding. In this work, we explore higher-order STSS and demonstrate how STSSs at different orders reveal distinct aspects of these dynamics. We then introduce the Multi-Order Self-Similarity (MOSS) module, a lightweight neural module designed to learn and integrate multi-order STSS features. It can be applied to diverse video tasks to enhance motion modeling capabilities while consuming only marginal computational cost and memory usage. Extensive experiments on video action recognition, motion-centric video VQA, and real-world robotic tasks consistently demonstrate substantial improvements, validating the broad applicability of MOSS as a general temporal modeling module. The source code and checkpoints will be publicly available.
Abstract（参考訳）: 空間時間自己相似性(STSS)は、フレーム間の視覚的対応を捉え、ビデオ理解のための時間的ダイナミクスを表現する効果的な方法を提供する。本研究では、高次STSSを探索し、異なる順序でのSTSSがどのようにこれらのダイナミクスの異なる側面を明らかにするかを実証する。次に,Multi-Order Self-Similarity (MOSS)モジュールを紹介した。多様なビデオタスクに適用することで、限界計算コストとメモリ使用量のみを消費しながら、モーションモデリング能力を向上させることができる。ビデオアクション認識、モーション中心ビデオVQA、実世界のロボットタスクに関する広範囲にわたる実験は、一般的な時間モデリングモジュールとしてのMOSSの適用性を検証し、一貫して顕著な改善を証明している。ソースコードとチェックポイントが公開されている。

論文の概要: Exploring High-Order Self-Similarity for Video Understanding

関連論文リスト