2608.06953v1 Aug 07, 2026 cs.CL

명시적으로 표현하되, 더 길게 하지 마세요: 기억 압축 과정에서 지식적 태도가 유지되는 이유

Explicit, Not Longer: What Makes Epistemic Stance Survive Memory Compression

Alex Kwon
Alex Kwon
Citations: 0
h-index: 0

에이전트의 메모리 시스템은 저장된 정보를 압축하며, 이 과정에서 수식어구는 제거되기 쉽습니다. 따라서 주장의 지식적 의미(epistemic standing)가 메모리에 기록될 때 제대로 보존되지 않는 경향이 있습니다. 본 연구에서는 이러한 현상이 발생하는 원인을 분석합니다. 동일한 주장과 지식적 태도를 가진 두 가지 형태의 정보를 비교하여, 압축 과정에서 한 모델은 다른 모델보다 더 많은 정보를 유지하도록 설정했습니다. 평가자는 어떤 조건 하에서 정보가 더 잘 보존되는지 판단합니다. 7가지 유형의 60개 주장을 대상으로 분석한 결과, 지식적 태도를 괄호 안에 넣는 방식 대신 명시적인 레이블로 표현하는 것이 두 모델 모두에서 약 15% 정도의 유지율을 높였습니다 (모델 1에서는 37개의 주장이 보존되고 2개만 손실된 반면, 모델 2에서는 30개가 보존되고 8개가 손실됨; permutation p=0.00005). 미리 등록된 복제 실험에서도 유사한 결과 (+15.6%)가 나타났으며, 38개의 주장이 보존되고 1개만 손실되었습니다. 두 모델 모두에서 형식(format)을 제거했을 때 동일한 효과가 나타났습니다. 레이블은 양쪽 모델 모두에서 도움이 되었고(+9.7 및 +12.8), 길이가 도움이 되지 않았습니다. 하지만 한 모델에서는 지식적 태도를 완전한 문장으로 표현하는 것이 가장 큰 영향을 미쳤지만 (+12.5), 다른 모델에서는 거의 영향이 없었습니다 (+0.6). 각 모델만으로는 서로 다른 메커니즘을 설명할 수 있었기 때문에, 우리는 두 모델의 공통점만을 강조합니다. 즉, 지식적 태도를 단순히 더 길게 표현하는 것이 아니라 명시적으로 표현해야 하며, 가장 효과적인 명시적인 표현 방식은 모델에 따라 달라질 수 있습니다. 모델 없이 결정론적인 방식으로 결과를 분석한 결과, 원래 관찰된 경향과 7가지 형식 변경 비교 중 5가지가 재현되었지만, 길이와 레이블은 재현되지 않았습니다. 50개의 수동으로 생성된 레이블(kappa=0.75)이 방향성에 대해 일치했으며, 7가지 불일치는 전체적으로 제시되어 있습니다. 또한, 본 논문의 원래 제목 후보였던 9개의 주장이 삭제되었습니다.

Original Abstract

Agent memory systems compress what they store, and compression is built to drop qualifiers, so a claim's epistemic standing tends not to survive being written to memory. We ask what governs whether it does. Matched notes carry the identical claim and identical stance and differ only in where that stance sits; one model compresses both under the same budget among the same filler notes, and a blind reader that never sees the condition scores the result. Across 60 claims in seven registers, writing the stance as a labelled field rather than a bracketed aside raises retention by about 15 points on two models (37 claims to 2 on one, 30 to 8 on the other; permutation p=0.00005), and a pre-registered replication on Haiku, its prediction and decision rule committed before the run, gives +15.6 points, 38 claims to 1. Ablating the format on both models gives the same net effect from different parts: labels help on both (+9.7 and +12.8) and length helps on neither, but wording the stance as a full sentence is the largest component on one model (+12.5) and worth nothing on the other (+0.6). Either model alone would have licensed a confident and different mechanism, so we claim only the intersection: make the stance explicit, not merely longer, and expect the best way of being explicit to depend on the model. A deterministic readout with no model reproduces the two-cell direction and five of seven ablation contrasts, but not length or labels, which we therefore do not claim on one instrument. Fifty hand labels (kappa=0.75) agree on direction; we print their seven disagreements in full. We also report nine withdrawn claims, three of them former title claims of this paper.

0 Citations
0 Influential
0 Altmetric
0.0 Score
Original PDF

No Analysis Report Yet

This paper hasn't been analyzed by Gemini yet.

Log in to request an AI analysis.

댓글

댓글을 작성하려면 로그인하세요.

아직 댓글이 없습니다. 첫 번째 댓글을 남겨보세요!