論文の概要: Language Models as Measurement Apparatus for Culture
- arxiv url: http://arxiv.org/abs/2607.02459v1
- Date: Thu, 02 Jul 2026 17:25:55 GMT
- ステータス: 翻訳完了
- システム内更新日: 2026-07-03 19:45:08.942102
- Title: Language Models as Measurement Apparatus for Culture
- Title(参考訳): 文化計測装置としての言語モデル
- Authors: Kent K. Chang,
- Abstract要約: 言語モデルは、文化現象の定量化にますます利用されている。
なぜそのような測定が文化的に異なるのか?
本論文は, 測定対象の文化的現実の構成に, モデル, データ, アノテーション, 評価などの装置が関与していることを論じる。
- 参考スコア(独自算出の注目度): 3.3842793760651566
- License: http://creativecommons.org/licenses/by-nc-sa/4.0/
- Abstract: Language models are increasingly used to quantify cultural phenomena, but what makes such measurement distinctively cultural? This paper argues that NLP work on culture is a material-discursive practice: the apparatus -- model, data, annotation, evaluation -- participates in constituting the cultural reality it measures, rather than passively recording it. Drawing on Karen Barad's concept of the agential cut -- the contingent boundary between phenomenon and instrument -- I show that the apparatus's substantive design choices draw such boundaries, and that the boundary is entangled from the start because language models have already internalized much of the cultural material they measure. I illustrate this through three case studies on television and film dialogue (measuring structure, interaction, and deviation) and three examinations of the apparatus itself (erasure of cultural markers, attunement to historical material, and agency in an agentic workflow). This big picture analysis proposes a research program that is theory-driven, empirically rigorous, and culturally contingent, treating each agential cut as a conscious commitment, at once methodological and ethical.
- Abstract(参考訳): 言語モデルは文化現象の定量化にますます使われていますが、そのような測定が文化的に顕著な理由は何でしょうか?
本論では,NLPの文化への取り組みは,受動的に記録するのではなく,その測定する文化的現実の構成に,モデル,データ,アノテーション,評価といった装置が関与する,物質的展開的な実践である,と論じる。
カレン・バラド(Karen Barad)のエージェントカットの概念(現象と楽器の境界)に基づき、私は装置の実質的な設計選択がそのような境界を描き、言語モデルが既に測定した文化的材料の多くを内部化しているため、境界は最初から絡み合っていることを示す。
本稿では,テレビと映画の対話(構造,相互作用,逸脱の測定)と装置自体の3つの検査(文化マーカーの測定,史料の添付,エージェントワークフローにおけるエージェンシー)について説明する。
この大局的な分析は、理論駆動で、実証的に厳密で、文化的に随伴する研究プログラムを提案し、各エージェントカットを、一度に方法論的かつ倫理的に、意識的なコミットメントとして扱う。
関連論文リスト
- Hire Your Anthropologist! Rethinking Culture Benchmarks Through an Anthropological Lens [9.000522371422628]
ベンチマークのフレームカルチャーを分類する4つのフレームワークを紹介します。
20の文化指標を質的に検討し,6つの方法論的問題を同定した。
我々の目標は、静的リコールタスクを超える文化ベンチマークの開発をガイドすることです。
論文 参考訳(メタデータ) (2025-10-07T13:42:44Z) - CultureScope: A Dimensional Lens for Probing Cultural Understanding in LLMs [57.653830744706305]
CultureScopeは、大規模な言語モデルにおける文化的理解を評価するための、これまでで最も包括的な評価フレームワークである。
文化的な氷山理論に触発されて、文化知識分類のための新しい次元スキーマを設計する。
実験結果から,文化的理解を効果的に評価できることが示唆された。
論文 参考訳(メタデータ) (2025-09-19T17:47:48Z) - Culture is Everywhere: A Call for Intentionally Cultural Evaluation [36.20861746863831]
文献的文化的評価について論じる: 評価のあらゆる側面に埋め込まれた文化的仮定を体系的に検証するアプローチ。
我々は、現在のベンチマークプラクティスを超えて、意味と今後の方向性について議論する。
論文 参考訳(メタデータ) (2025-09-01T09:39:21Z) - From Word to World: Evaluate and Mitigate Culture Bias in LLMs via Word Association Test [50.51344198689069]
我々は,人中心語関連テスト(WAT)を拡張し,異文化間認知による大規模言語モデルのアライメントを評価する。
文化選好に対処するために,モデルの内部表現空間に直接,文化固有の意味的関連性を直接埋め込む革新的なアプローチであるCultureSteerを提案する。
論文 参考訳(メタデータ) (2025-05-24T07:05:10Z) - Extrinsic Evaluation of Cultural Competence in Large Language Models [53.626808086522985]
本稿では,2つのテキスト生成タスクにおける文化能力の評価に焦点をあてる。
我々は,文化,特に国籍の明示的なキューが,そのプロンプトに乱入している場合のモデル出力を評価する。
異なる国におけるアウトプットのテキスト類似性とこれらの国の文化的価値との間には弱い相関関係がある。
論文 参考訳(メタデータ) (2024-06-17T14:03:27Z)
関連論文リストは本サイト内にある論文のタイトル・アブストラクトから自動的に作成しています。
指定された論文の情報です。
本サイトの運営者は本サイト(すべての情報・翻訳含む)の品質を保証せず、本サイト(すべての情報・翻訳含む)を使用して発生したあらゆる結果について一切の責任を負いません。