論文の概要: AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan
- arxiv url: http://arxiv.org/abs/2604.08184v1
- Date: Thu, 09 Apr 2026 12:38:19 GMT
- ステータス: 翻訳完了
- システム内更新日: 2026-04-10 18:34:05.919697
- Title: AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan
- Title(参考訳): AT-ADD:オール型オーディオディープフェイク検出課題評価計画
- Authors: Yuankun Xie, Haonan Cheng, Jiayi Zhou, Xiaoxuan Guo, Tao Wang, Jian Liu, Weiqiang Wang, Ruibo Fu, Xiaopeng Wang, Hengyan Huang, Xiaoying Huang, Long Ye, Guangtao Zhai,
- Abstract要約: ACMマルチメディア2026におけるオールタイプオーディオディープフェイク検出(AT-ADD)グランドチャレンジを提案する。
AT-ADDは、堅牢で一般化可能なオーディオ法医学技術の開発を加速することを目的としている。
- 参考スコア(独自算出の注目度): 64.09595490689874
- License: http://creativecommons.org/licenses/by/4.0/
- Abstract: The rapid advancement of Audio Large Language Models (ALLMs) has enabled cost-effective, high-fidelity generation and manipulation of both speech and non-speech audio, including sound effects, singing voices, and music. While these capabilities foster creativity and content production, they also introduce significant security and trust challenges, as realistic audio deepfakes can now be generated and disseminated at scale. Existing audio deepfake detection (ADD) countermeasures (CMs) and benchmarks, however, remain largely speech-centric, often relying on speech-specific artifacts and exhibiting limited robustness to real-world distortions, as well as restricted generalization to heterogeneous audio types and emerging spoofing techniques. To address these gaps, we propose the All-Type Audio Deepfake Detection (AT-ADD) Grand Challenge for ACM Multimedia 2026, designed to bridge controlled academic evaluation with practical multimedia forensics. AT-ADD comprises two tracks: (1) Robust Speech Deepfake Detection, which evaluates detectors under real-world scenarios and against unseen, state-of-the-art speech generation methods; and (2) All-Type Audio Deepfake Detection, which extends detection beyond speech to diverse, unknown audio types and promotes type-agnostic generalization across speech, sound, singing, and music. By providing standardized datasets, rigorous evaluation protocols, and reproducible baselines, AT-ADD aims to accelerate the development of robust and generalizable audio forensic technologies, supporting secure communication, reliable media verification, and responsible governance in an era of pervasive synthetic audio.
- Abstract(参考訳): オーディオ大言語モデル(ALLM)の急速な進歩により、音声効果、歌声、音楽を含む音声と非音声の両方をコスト効率が高く、高忠実に生成し、操作することが可能になった。
これらの機能はクリエイティビティとコンテンツ生産を促進する一方で、現実的なオーディオディープフェイクを大規模に生成し、普及させることができるため、セキュリティと信頼の面での重大な課題も導入している。
しかし、既存の音声ディープフェイク検出(ADD)とベンチマークは、主に音声中心であり、しばしば音声固有のアーティファクトに依存し、実世界の歪みに制限された堅牢性を示す。
これらのギャップに対処するために,ACMマルチメディア2026におけるオールタイプオーディオディープフェイク検出(AT-ADD)グランドチャレンジを提案する。
AT-ADD は,(1) 実環境のシナリオ下で検出者を評価するロバスト音声深度検出,(2) 音声以外の様々な未知の音声タイプへの検出を拡張し,音声,音,歌,音楽のタイプ非依存的な一般化を促進する全型音声深度検出,の2つのトラックから構成される。
標準化されたデータセット、厳密な評価プロトコル、再現可能なベースラインを提供することにより、AT-ADDは、堅牢で一般化可能なオーディオ法医学技術の開発を加速し、セキュアな通信、信頼できるメディア検証、広汎な合成オーディオの時代における責任あるガバナンスをサポートすることを目的としている。
関連論文リスト
- Detect All-Type Deepfake Audio: Wavelet Prompt Tuning for Enhanced Auditory Perception [19.10177637063233]
既存の対策 (CM) は単一型オーディオディープフェイク検出 (ADD) では良好に機能するが, クロスタイプのシナリオでは性能が低下する。
我々は、音声、音声、歌声、音楽のクロスタイプディープフェイク検出を取り入れ、現在のCMを評価するためのオールタイプADDベンチマークを包括的に確立した最初の人物である。
異なる音声タイプの聴覚知覚を考慮し,タイプ不変の聴覚深度情報をキャプチャするためのウェーブレット・プロンプト・チューニング(WPT)-SSL法を提案する。
論文 参考訳(メタデータ) (2025-04-09T10:18:45Z) - Where are we in audio deepfake detection? A systematic analysis over generative and detection models [59.09338266364506]
SONARはAI-Audio Detection FrameworkとBenchmarkの合成である。
最先端のAI合成聴覚コンテンツを識別するための総合的な評価を提供する。
従来のモデルベース検出システムと基礎モデルベース検出システムの両方で、AIオーディオ検出を均一にベンチマークする最初のフレームワークである。
論文 参考訳(メタデータ) (2024-10-06T01:03:42Z) - Partially Fake Audio Detection by Self-attention-based Fake Span
Discovery [89.21979663248007]
本稿では,部分的に偽の音声を検出する自己認識機構を備えた質問応答(フェイクスパン発見)戦略を導入することで,新たな枠組みを提案する。
ADD 2022の部分的に偽の音声検出トラックで第2位にランクインした。
論文 参考訳(メタデータ) (2022-02-14T13:20:55Z)
関連論文リストは本サイト内にある論文のタイトル・アブストラクトから自動的に作成しています。
指定された論文の情報です。
本サイトの運営者は本サイト(すべての情報・翻訳含む)の品質を保証せず、本サイト(すべての情報・翻訳含む)を使用して発生したあらゆる結果について一切の責任を負いません。