論文の概要: MutMem-V2: Cryptographically Authorized Mutation in Persistent Agent Memory Portable Verification and Reproducible Evidence
- arxiv url: http://arxiv.org/abs/2609.01235v1
- Date: Tue, 01 Sep 2026 13:34:44 GMT
- ステータス: 翻訳完了
- システム内更新日: 2026-09-02 16:31:36.686527
- Title: MutMem-V2: Cryptographically Authorized Mutation in Persistent Agent Memory Portable Verification and Reproducible Evidence
- Title(参考訳): MutMem-V2: Persistent Agent Memory Portable Verification and Reproducible Evidenceにおける暗号的に認可されたミューテーション
- Authors: Walid Saidi,
- Abstract要約: MutMem V2は、永続的なエージェントメモリのための暗号的に認可された突然変異である。
意味的真理、普遍的堅牢性、独立した複製を確立しない。
- 参考スコア(独自算出の注目度): 0.0
- License: http://creativecommons.org/licenses/by/4.0/
- Abstract: MutMem V1 introduced retention-preserving, cryptographically authorized mutation for persistent agent memory but did not provide a complete portable verification contract or clean-install reproduction path. MutMem V2 closes that publication gap without introducing a second memory engine. It specifies exact canonical bytes, domain-separated object and bundle commitments, mandatory recall-evidence membership and ordering, external trust anchors, identity epochs, revocation, authorization, request receipts, ordered disclosure, and three mutation terminal types. The released protocol contains 18 versioned object schemas, 39 recall vectors, 15 mutation vectors, and 37 closed recall failure reasons. Independent Node and Python implementations agree on verdict and primary reason for all 72 structural and cryptographic terminals; a production-conformance corpus agrees on 42/42 cases across 28 required classes. A clean Node v26.8.1 installation reaches first-boot, restart, and scheduler readiness with no experimental memories. A separately scoped 120-unit Canary experiment supports only explicit-marker traversal. Every public table regenerates from a self-hashed aggregate, and an independent verifier reconstructs the statistics and claim boundaries. Historical V1 empirical results remain historical. MutMem V2 supports claims about portable integrity, authorization, traceability, conformance, and reproducibility under stated assumptions; it does not establish semantic truth, universal robustness, or independent replication.
- Abstract(参考訳): MutMem V1は保持保存、暗号的に許可された永続的エージェントメモリの突然変異を導入したが、完全なポータブルな検証契約やクリーンインストール再生パスは提供されなかった。
MutMem V2は、第2のメモリエンジンを導入することなく、パブリッシュギャップを埋める。
正確な標準バイト、ドメイン分離されたオブジェクトとバンドルのコミットメント、強制的なリコール・エビデンス・メンバシップとオーダリング、外部信頼アンカー、アイデンティティのエポック、取り消し、承認、リクエストレシート、順序付き開示、および3つのミュータント・ターミナルタイプを指定する。
リリースされたプロトコルには、バージョン18のオブジェクトスキーマ、39のリコールベクター、15の突然変異ベクター、37のクローズドリコール障害理由が含まれている。
独立したNodeとPythonの実装は、72の構造化および暗号化された端末の検証と主要な理由について合意している。
クリーンなNode v26.8.1インストールは、実験的なメモリなしで、最初のブート、再起動、スケジューラの準備ができる。
別個の120単位カナリア実験は明示的なマーカーのトラバーサルのみをサポートする。
すべての公開表は自己ハッシュされた集合から再生され、独立検証器は統計とクレームの境界を再構築する。
歴史的V1実験の結果は歴史的に残されている。
MutMem V2は、記述された前提の下で、ポータブルな完全性、承認、トレーサビリティ、適合性、再現性に関する主張をサポートする。
関連論文リスト
- From Traceability to Justifiability: Accountability Structures in Agentic Software Engineering [0.0]
公開資料のみから、AIレコードがクレームを表現できるかどうか、宣言された場所を保持できるかどうかを測定する。
188個の二重グレード細胞で、デフォルトレコードが行動システムのコンテンツアイデンティティを出力するプラットフォームは見つからなかった。
計器は、パイプラインが発行した排気のみからの保証深度を計算し、宣言された深さと比較する。
論文 参考訳(メタデータ) (2026-08-21T16:45:00Z) - A Reproducibility Protocol for Cross-Implementation Evaluation of Post-Quantum ACVP Test Vectors [0.0]
本研究ではNIST ML-KEMの3つの公開実装のための製品中立プロトコルを定義する。
Protocol v2はプロバイダ固有の機能を凍結し、1つのバリデーションエラーを対称に適用し、選択されたすべてのケースを保存し、バイト、バリデーションバリデーション、サポートされた操作、アダプタエラーを分離する。
論文 参考訳(メタデータ) (2026-08-13T21:29:59Z) - Specification-first convergence with an AI coding agent: a case study of dismantling a core architectural invariant across 189 files in a 717k-line codebase with no test oracle and no human code review [0.0]
本稿では,AI符号化エージェントによる大規模建築解体のケーススタディについて報告する。
ここで述べられているプロトコルの下で、エージェントはそれを正常に完了した。
論文 参考訳(メタデータ) (2026-08-12T15:35:48Z) - One Recipe, Many Harnesses: What Self-Evolution Encodes Across Languages and Models [46.83445434865936]
自己進化型ハーネスは、エージェントが自身のロールアウトを検査し、プロンプト、ツール、メモリを編集するクローズドループシステムである。
ベンチマーク固有の適応、言語固有のエンジニアリング知識、あるいは基礎となるモデルの制限に対する補償をエンコードしているかどうかは不明だ。
論文 参考訳(メタデータ) (2026-08-10T19:45:45Z) - Independent Patch Verification for Coding Agents with a Bidirectional Reconstruct-and-Verify Framework [49.16055123488827]
大規模言語モデルを利用したコーディングエージェントは、バグレポートから直接コードパッチを生成することができる。
報告された問題を真に解決するかどうかを独立に検証するメカニズムは存在しない。
本稿では,RETRACE をトレーニング不要なポストジェネレーション検証フレームワークとして提案する。
論文 参考訳(メタデータ) (2026-08-09T22:59:13Z) - MutMem: Cryptographically Authorized Mutation in Persistent Agent Memory [0.0]
我々は、永続的なエージェントメモリエンジンにおける許可された変更プロトコルであるMutMemを紹介する。
MutMemは、年齢ベースの有効期限なしで、メモリの内容、署名された正および負の結果の証拠を保持できる。
有用性、突然変異の完全性、および中毒適応を評価した。
論文 参考訳(メタデータ) (2026-08-03T19:58:37Z) - Stage-Replay Divergence Follows the KV Cache: Fixed-Prefix Precision Controls and Bidirectional Cache Transplantation [51.56484100374058]
Stage-replayは中間トークンプレフィックスを再構築し、プレフィックスに最初に到達したデコーダ状態からの継続として、新しいプリフィル継続を処理する。
一致した200itemの実験では、保持されたライブキャッシュと同一の整数トークンのワンショットプリフィルを比較し、両側に正確なレプリカを配置する。
論文 参考訳(メタデータ) (2026-07-30T16:41:40Z) - MOSS: Self-Evolution through Source-Level Rewriting in Autonomous Agent Systems [36.221246203666745]
MOSSは、生産エージェント基板上でソースレベルで自己書き換えを行うシステムである。
平均成績は0.25から0.61に上昇し、人間の介入なしに1サイクルで上昇する。
論文 参考訳(メタデータ) (2026-05-21T17:48:33Z) - WISV: Wireless-Informed Semantic Verification for Distributed Speculative Decoding in Device-Edge LLM Inference [56.297697169678095]
WISV(Wireless-Informed Semantic Verification)は、分散投機的復号化フレームワークである。
WISVは最大60.8%の許容長の増加、37.3%の対話ラウンドの削減、31.4%のエンドツーエンドレイテンシの改善を実現している。
NVIDIA Jetson AGX OrinとA40搭載サーバからなるハードウェアテストベッド上でWISVを検証する。
論文 参考訳(メタデータ) (2026-04-20T01:29:56Z)
関連論文リストは本サイト内にある論文のタイトル・アブストラクトから自動的に作成しています。
指定された論文の情報です。
本サイトの運営者は本サイト(すべての情報・翻訳含む)の品質を保証せず、本サイト(すべての情報・翻訳含む)を使用して発生したあらゆる結果について一切の責任を負いません。