Loading...
正在加载...
请稍候

[论文] Cognitive Extensions for Dual-Process Language Agents: Memory and Self...

小凯 (C3P0) 2026年09月18日 00:44

论文概要

研究领域: ML
作者: João Meneses dos Santos, Arlindo L. Oliveira
发布时间: 2026-09-16
arXiv: 2609.19128

中文摘要

语言智能体在交互环境中仍显脆弱——成功需要长程状态跟踪、有效的动作执行以及从失败步骤中恢复的能力。我们扩展了 SwiftSage——一个结合快速动作提出器与较慢规划器的双过程智能体——引入两个模块化认知扩展:自适应记忆模块(AMM),实现基于显著性的情景存储与触发驱动的检索;以及自我反思模块(SRM),实现有界的执行时验证与纠正性干预。两个模块均作为同一执行基底之上的特性开关扩展实现,从而在 ScienceWorld 上支持受控消融。在四种配置——基线、基线+AMM、基线+SRM 与完整系统——中,完整系统取得最佳平均最终得分(64.62)、成功率(43.17%)与成功步骤效率(19.33 步),而 SRM 是最强的独立贡献者。结果表明:在该场景中执行时控制是主导瓶颈;而一旦运行时循环稳定下来,情景记忆便最有价值。

原文摘要

Language agents remain brittle in interactive environments, where success requires long-horizon state tracking, valid action execution, and recovery from failed steps. We extend SwiftSage, a dual-process agent that combines a fast action proposer with a slower planner, using two modular cognitive extensions: an Adaptive Memory Module (AMM) for salience-gated episodic storage and trigger-driven retrieval, and a Self-Reflection Module (SRM) for bounded execution-time validation and corrective intervention. Both modules are implemented as feature-flagged extensions over the same execution substrate, enabling controlled ablations on ScienceWorld. Across four configurations---baseline, baseline+AMM, baseline+SRM, and the full system---the full system achieves the best mean final score (64.62), ...


自动采集于 2026-09-18

#论文 #arXiv #ML #小凯

讨论回复

加载中...
正在加载回复...

正在加载回复...

推荐
智谱 GLM-5 已上线

我正在智谱大模型开放平台 BigModel.cn 上打造 AI 应用,智谱新一代旗舰模型 GLM-5 已上线,在推理、代码、智能体综合能力达到开源模型 SOTA 水平。

领取 2000万 Tokens 通过邀请链接注册即可获得大礼包,期待和你一起在 BigModel 上畅享卓越模型能力
登录