論文の概要: Constructing Parallel Multidimensional Chromatic Lexicons for Corpus-Assisted Analysis of Russian and English Texts
- arxiv url: http://arxiv.org/abs/2608.01752v1
- Date: Mon, 03 Aug 2026 06:20:23 GMT
- ステータス: 翻訳完了
- システム内更新日: 2026-08-04 15:07:25.349144
- Title: Constructing Parallel Multidimensional Chromatic Lexicons for Corpus-Assisted Analysis of Russian and English Texts
- Title(参考訳): コーパス支援によるロシア語と英語のテキスト分析のための並列多次元クロマティック辞書の構築
- Authors: Larisa Nikitina,
- Abstract要約: ロシア語(224エントリ)と英語(141エントリ)の2つの多次元色レキシコンの開発について述べる。
レキシコンは、色調、彩度、温度に応じてエントリを分類する。
アンドレイ・ベーリーとエミリー・ディキンソンによる詩集を純粋にサンプリングした試験的な研究が行われた。
- 参考スコア(独自算出の注目度): 0.0
- License: http://creativecommons.org/licenses/by/4.0/
- Abstract: This article addresses the relative scarcity of research tools for the corpus-assisted linguistic analysis of colour terms in literary texts. It describes the development of two multidimensional chromatic lexicons: one for Russian (224 entries) and one for English (141 entries). Lexicon construction involved sourcing colour vocabulary from specialised resources and research literature, comparing the two language inventories, manually checking translated candidates, and addressing language-specific morphological features. In addition to identifying colour terms and visual descriptors, the lexicons classify entries according to hue, saturation, and temperature. To demonstrate their practical application, a pilot study was conducted on purposively sampled corpora of poetry by Andrei Bely (20,373 tokens) and Emily Dickinson (28,479 tokens). All retrieved matches were checked in context and classified as Confirmed_chromatic, Ambiguous_visual, or Excluded. The analysis was implemented in two main stages: a strict analysis including confirmed chromatic lexis only, followed by a sensitivity analysis incorporating both confirmed and ambiguous chromatic lexis to determine whether coding decisions about borderline cases affected the main findings. The quantitative results indicated marked differences in the use of colour terms, visual descriptors, hue, saturation, and temperature. Specifically, the analysis revealed that confirmed chromatic terms occurred 3.4 times more frequently in the sampled Bely corpus than in the Dickinson corpus. These findings demonstrate the analytical value of a multidimensional approach, with the main contribution of this study being a transparent and reusable procedure for constructing and applying multilingual chromatic lexicons.
- Abstract(参考訳): 本稿では,文文における色彩用語のコーパス支援言語分析のための研究ツールの相対的不足について論じる。
ロシア語(224エントリ)と英語(141エントリ)の2つの多次元色レキシコンの開発について記述している。
辞書の構築には、特殊リソースと研究文献からカラー語彙を抽出し、2つの言語在庫を比較し、翻訳された候補を手動でチェックし、言語固有の形態的特徴に対処することが含まれていた。
色用語と視覚ディスクリプタの識別に加えて、レキシコンは色、彩度、温度に応じてエントリを分類する。
その実践的応用を実証するために、アンドレイ・ベーリー(20,373トークン)とエミリー・ディキンソン(28,479トークン)による詩の純粋にサンプリングされたコーパスについて試験的な研究が行われた。
検索したマッチはすべてコンテキストでチェックされ、Confirmed_chromatic, Ambiguous_visual, Excludedに分類された。
この分析は2つの主要な段階において実施された: 確定色レキシスのみを含む厳密な分析、および、確定色レキシスと曖昧色レキシスを併用した感度分析により、境界線のケースに関するコーディング決定が主な発見に影響を与えるか否かを判定した。
その結果,色調,視覚ディスクリプタ,色調,彩度,温度に有意な差が認められた。
具体的には、確認された色調項はディキンソン・コーパスの3.4倍の頻度で発生したことが明らかとなった。
これらの結果は多次元アプローチの解析的価値を示し,本研究の主な貢献は多言語カラーレキシコンの構築と適用のための透明かつ再利用可能な手順である。
関連論文リスト
- Floating or Suggesting Ideas? A Large-Scale Contrastive Analysis of Metaphorical and Literal Verb-Object Constructions [53.690096725532726]
本研究では,2Mコーパス文中の297の英語動詞オブジェクト対(例:float idea vs. suggest idea)を分析した。
5つのNLPツールを用いて,感情的,語彙的,統語的,言論的な特徴を捉えた認知的・言語的特徴2,293点を抽出した。
クロスペアの結果は, 語彙頻度, 凝集度, 構造規則性が高く, 比喩的文脈は感情負荷, イメージ性, 語彙多様性, 構造的特異性を示す。
論文 参考訳(メタデータ) (2026-04-09T14:08:57Z) - Understanding Stigmatizing Language Lexicons: A Comparative Analysis in Clinical Contexts [45.61748951587092]
言語を安定させると、医療的不平等が生じる。
医療においてどの単語、用語、フレーズがスティグマタイズ言語を構成するかを定義する普遍的または標準化された語彙は存在しない。
論文 参考訳(メタデータ) (2025-09-09T07:41:20Z) - Underutilization of Syntactic Processing by Chinese Learners of English in Comprehending English Sentences, Evidenced from Adapted Garden-Path Ambiguity Experiment [0.0]
本研究は, 統語処理の非活用を, 統語的観点から強調する。
この研究は、部分的および完全という2種類のパーシングアンダーユーティライゼーションを識別する。
構文処理を文理解に完全に統合する新しい構文解析法の開発の基礎を築いた。
論文 参考訳(メタデータ) (2024-12-21T01:32:10Z) - Análise de ambiguidade linguística em modelos de linguagem de grande escala (LLMs) [0.35069196259739965]
言語的曖昧さは、自然言語処理(NLP)システムにとって重要な課題である。
近年のChatGPTやGeminiのような教育モデルの成功に触発されて,これらのモデルにおける言語的あいまいさを分析し,議論することを目的とした。
論文 参考訳(メタデータ) (2024-04-25T14:45:07Z) - Perceptual Structure in the Absence of Grounding for LLMs: The Impact of
Abstractedness and Subjectivity in Color Language [2.6094835036012864]
定義色空間と言語モデルで定義される特徴空間との間にはかなりの整合性があることが示される。
その結果,色空間のアライメントはモノレキセミックで実用的な色記述を保ちつつも,実際の言語的利用の要素を示す例の存在感は著しく低下することがわかった。
論文 参考訳(メタデータ) (2023-11-22T02:12:36Z) - A bilingual approach to specialised adjectives through word embeddings
in the karstology domain [3.92181732547846]
単語埋め込みを用いた特定の意味関係を表現する形容詞の抽出実験を行う。
実験の結果は徹底的に分析され、形式的または意味的な類似性を示す形容詞のグループに分類される。
論文 参考訳(メタデータ) (2022-03-31T08:27:15Z) - A Latent-Variable Model for Intrinsic Probing [93.62808331764072]
固有プローブ構築のための新しい潜在変数定式化を提案する。
我々は、事前訓練された表現が言語間交互に絡み合ったモルフォシンタクスの概念を発達させる経験的証拠を見出した。
論文 参考訳(メタデータ) (2022-01-20T15:01:12Z) - Clinical Named Entity Recognition using Contextualized Token
Representations [49.036805795072645]
本稿では,各単語の意味的意味をより正確に把握するために,文脈型単語埋め込み手法を提案する。
言語モデル(C-ELMo)とC-Flair(C-Flair)の2つの深い文脈型言語モデル(C-ELMo)を事前訓練する。
明示的な実験により、静的単語埋め込みとドメインジェネリック言語モデルの両方と比較して、我々のモデルは劇的に改善されている。
論文 参考訳(メタデータ) (2021-06-23T18:12:58Z) - AM2iCo: Evaluating Word Meaning in Context across Low-ResourceLanguages
with Adversarial Examples [51.048234591165155]
本稿では, AM2iCo, Adversarial and Multilingual Meaning in Contextを提案する。
言語間文脈における単語の意味の同一性を理解するために、最先端(SotA)表現モデルを忠実に評価することを目的としている。
その結果、現在のSotAプリトレーニングエンコーダは人間のパフォーマンスにかなり遅れていることが明らかとなった。
論文 参考訳(メタデータ) (2021-04-17T20:23:45Z)
関連論文リストは本サイト内にある論文のタイトル・アブストラクトから自動的に作成しています。
指定された論文の情報です。
本サイトの運営者は本サイト(すべての情報・翻訳含む)の品質を保証せず、本サイト(すべての情報・翻訳含む)を使用して発生したあらゆる結果について一切の責任を負いません。