2607.23929v1 Jul 27, 2026 cs.AI

MemTX: 상태 기반 에이전트 메모리를 위한 트랜잭셔널 Belief Commit

MemTX: Transactional Belief Commit for Stateful Agent Memory

Xiaoyang Li
Xiaoyang Li
Citations: 0
h-index: 0
Pingan Song
Pingan Song
Citations: 0
h-index: 0
Taotao Cai
Taotao Cai
Citations: 47
h-index: 2
Mingkai Zheng
Mingkai Zheng
Citations: 0
h-index: 0
Yiqi Wang
Yiqi Wang
Citations: 2
h-index: 1
Haohui Lu
Haohui Lu
Citations: 10
h-index: 2
Z. Chen
Z. Chen
Citations: 0
h-index: 0
Mo Li
Mo Li
Citations: 0
h-index: 0

LLM (Large Language Model) 에이전트는 점점 더 지속적인 공유 메모리를 통해 협력합니다. 한 에이전트의 쓰기 작업은 다른 에이전트에게 전제가 되며, 결국 실제 부작용을 가진 도구 호출로 이어질 수 있습니다. 현재 에이전트 메모리 시스템에서는 모든 허용된 쓰기 작업을 즉시 실행 가능한 진실로 취급하므로, 오염된 도구 결과, 오래된 업데이트 또는 동료의 미완성 메모가 조용히 되돌릴 수 없는 동작을 유발할 수 있습니다. 우리는 메모리 쓰기 작업이 단순한 Belief Commit (신념 약속)이 아니라고 주장합니다. 본 논문에서는 MemTX라는 트랜잭셔널 Belief-Commit 프로토콜을 제시합니다. 각 레코드는 증거, 권한, 출처 및 유효성 정보를 포함합니다. 쓰기 작업은 스냅샷 격리된 트랜잭션 내에서 이루어지며, 검증 및 커밋 파이프라인을 거쳐 허용됩니다. 되돌릴 수 없는 도구 호출은 현재 진행 중인 Belief 상태에 의해 제어되며, Belief를 철회하면 관련 기록과 도구의 부작용에 대한 유형별 체계적인 복구가 트리거됩니다. 본 논문에서 제시하는 두 가지 불변 조건 (액션 안전성 제어 및 Cascade-Repair 완전성)은 속성 기반 테스트와 550만 개의 프로토콜 상태에 대한 경계 지향적 완전 탐색을 통해 기계적으로 검증되었으며, 위반 사례는 전혀 발견되지 않았습니다. 세 가지 모델 패밀리에서 파생된 다섯 가지 백본 환경에서 MemTX는 네 가지 백본에서 8개의 기준 성능보다 우수한 결과를 보였으며, 통계적으로 유의미하게 가장 좋은 기준 성능과 동률을 보였습니다. 또한 모든 백본에서 다운스트림 피해가 전혀 없는 유일한 방법입니다. 백본의 기능은 Commit 규율을 대체할 수 없습니다.

Original Abstract

LLM agents increasingly coordinate through persistent shared memory: one agent's write becomes another agent's premise, and eventually a tool call with real side effects. Current agent memory systems treat every accepted write as immediately actionable truth, so a polluted tool result, a stale update, or a teammate's half-finished note can silently drive an irreversible action. We argue that a memory write is not a belief commit. We present MemTX, a transactional belief-commit protocol. Each record carries evidence, permissions, provenance, and validity. Writes are staged inside snapshot-isolated transactions and admitted by a validate-and-commit pipeline, irreversible tool calls are gated on in-flight belief state, and retracting a belief triggers typed cascading repair of its derived records and tool side effects. Two invariants, action-safety gating and cascade-repair completeness, are machine-checked by property-based testing and bounded exhaustive enumeration of 5.5 million protocol states, with zero violations. Across five backbones from three model families, MemTX leads all eight baselines with paired-McNemar significance on four backbones and statistically ties the best baseline on the fifth and strongest, while remaining the only method with zero downstream harm on every backbone. Backbone capability does not substitute for commit discipline.

0 Citations
0 Influential
1 Altmetric
5.0 Score
Original PDF

No Analysis Report Yet

This paper hasn't been analyzed by Gemini yet.

Log in to request an AI analysis.

댓글

댓글을 작성하려면 로그인하세요.

아직 댓글이 없습니다. 첫 번째 댓글을 남겨보세요!