Fugu-MT 論文翻訳(概要): AgenticShop: Benchmarking Agentic Product Curation for Personalized Web Shopping

論文の概要: AgenticShop: Benchmarking Agentic Product Curation for Personalized Web Shopping

arxiv url: http://arxiv.org/abs/2602.12315v1
Date: Thu, 12 Feb 2026 17:25:45 GMT
ステータス: 翻訳完了
システム内更新日: 2026-02-16 23:37:53.711774
Title: AgenticShop: Benchmarking Agentic Product Curation for Personalized Web Shopping
Title（参考訳）: AgenticShop: パーソナライズされたWebショッピングのためのエージェント製品キュレーションのベンチマーク
Authors: Sunghwan Kim, Ryang Heo, Yongsik Seo, Jinyoung Yeo, Dongha Lee,
Abstract要約: 我々は、オープンウェブ環境におけるパーソナライズされた製品キュレーションにおけるエージェントシステム評価のための最初のベンチマークであるAgenticShopを紹介する。提案手法は,現実的なショッピングシナリオ,多様なユーザプロファイル,検証可能なチェックリストによるパーソナライズ評価フレームワークを特徴とする。
参考スコア（独自算出の注目度）: 20.52047960513448
License: http://creativecommons.org/licenses/by/4.0/
Abstract: The proliferation of e-commerce has made web shopping platforms key gateways for customers navigating the vast digital marketplace. Yet this rapid expansion has led to a noisy and fragmented information environment, increasing cognitive burden as shoppers explore and purchase products online. With promising potential to alleviate this challenge, agentic systems have garnered growing attention for automating user-side tasks in web shopping. Despite significant advancements, existing benchmarks fail to comprehensively evaluate how well agentic systems can curate products in open-web settings. Specifically, they have limited coverage of shopping scenarios, focusing only on simplified single-platform lookups rather than exploratory search. Moreover, they overlook personalization in evaluation, leaving unclear whether agents can adapt to diverse user preferences in realistic shopping contexts. To address this gap, we present AgenticShop, the first benchmark for evaluating agentic systems on personalized product curation in open-web environment. Crucially, our approach features realistic shopping scenarios, diverse user profiles, and a verifiable, checklist-driven personalization evaluation framework. Through extensive experiments, we demonstrate that current agentic systems remain largely insufficient, emphasizing the need for user-side systems that effectively curate tailored products across the modern web.
Abstract（参考訳）: eコマースの普及により、Webショッピングプラットフォームは、巨大なデジタルマーケットプレースをナビゲートする顧客にとって重要なゲートウェイとなっている。しかし、この急速な拡大により、ノイズと断片化された情報環境が生まれ、買い物客が商品をオンラインで探したり購入したりすることで認知的負担が増大した。この課題を緩和する有望な可能性を秘めたエージェントシステムは,Webショッピングにおけるユーザ側タスクの自動化に注目が集まっている。大幅な進歩にもかかわらず、既存のベンチマークでは、エージェントシステムがいかにオープンなWeb設定で製品をキュレートできるかを包括的に評価することができない。具体的には,探索探索ではなく,単一プラットフォーム検索の簡易化にのみ焦点を絞った,ショッピングシナリオのカバー範囲が限られている。さらに、評価におけるパーソナライズを見落とし、エージェントがリアルなショッピングコンテキストにおいて多様なユーザー嗜好に適応できるかどうかも不明である。このギャップに対処するために、オープンウェブ環境におけるパーソナライズされた製品キュレーションにおけるエージェントシステム評価のための最初のベンチマークであるAgenticShopを提案する。当社のアプローチは,現実的なショッピングシナリオ,多様なユーザプロファイル,検証可能なチェックリストによるパーソナライズ評価フレームワークを備えている。大規模な実験を通じて、現在のエージェントシステムは依然としてほとんど不十分であり、現代のウェブ全体にわたって効果的にカスタマイズされた製品をキュレートするユーザ側システムの必要性を強調した。

論文の概要: AgenticShop: Benchmarking Agentic Product Curation for Personalized Web Shopping

関連論文リスト