首页 > AI前沿 > After the Fix: Transfer of Corrected Agent Experience

After the Fix: Transfer of Corrected Agent Experience

arXiv自然语言 2026-09-28 16:35 6 阅读 查看原文

Does repairing an episode make its experience a better memory for the next task? We transfer the same failed source before and after accepted repair to a fixed target, alongside independent execution.

Our 3,300 runs cover 100 ThinkingBox pairs and the same 100 APEX pairs with and without source-state inheritance, under eleven conditions.

Results

ThinkingBox's Full/Skill/Hybrid correction gains are 44/29/32 percentage points, with corrected performance 25/22/18 points above independence; inference weakens at the task-family level.

Yet 12 of Full's 15-point larger correction gap over Skill come from worse uncorrected performance, not better corrected memory.

Moreover, 22 of Full's 46 upward transitions restore observed baseline success.

Neither APEX regime establishes comparable aggregate correction benefits.

Discussion

Action evidence connects workflow gains with reusable obligations and convention conflicts with source-local choices.

Text APEX's accepted execution reaches 52% versus its summary's 40%, without robust global/group-level superiority or an established advantage over independence.

Smaller handoffs reduce input but increase calls.

Conclusion

The value of repairing experience is therefore distinct from the value of reusing it: memory updates require both a previous-version reference and a fresh-start reference.