基本信息

要点解读

这是什么

提出一种名为人工 id 的内部驱动机制,使自主智能体能够根据行为在不同情境中的持续性自动决定继续、停止或切换,从而实现无需外部明确目标的适应性控制。

用在哪里

适用于设计长期运行、跨任务保持状态的自主智能体系统,特别是需要在内置驱动层面实现安全约束和持续对齐的研究与实践。

可以推断的

  • 推测:实现上可能采用轻量级状态模块,跟踪行为的历史持续性并生成驱动力。
  • 推测:在实际部署时,需要额外的边界机制限制持久化状态的范围,以防止错误行为长期累积。

来源摘要/节选

Agentic AI is moving from bounded task execution toward systems that retain consequential state, continue operating and adapt across task boundaries. That shift creates a control problem that current harnesses largely solve by hand: objectives, retries, verification, stopping rules and other behavioral transitions are specified externally. We propose an artificial id, an adaptive internal drive for determining whether behavior should continue, stop or change. In a minimal virtual Petri-dish experiment, a controller too small to perform general-purpose reasoning and receiving no task-specific behavioral objective develops useful control through differential persistence. The same mechanism selects an unintended physical strategy when that behavior persists better and later replaces a learned sensor mapping when its environmental meaning changes. These results show that adaptive direction can emerge without being explicitly specified as a behavioral objective. The same persistence that makes such adaptive agency useful can also allow misalignment, corrupted state and unintended behavior to persist across task boundaries. A scalable artificial id would carry consequential state and adaptive drive across those boundaries, making alignment a property of the continuing agentic system rather than of a model response or single trajectory. Such systems require a persistent alignment boundary over trusted observations, consequence channels, persistent state, authority, identity, provenance and hard constraints.

来源说明

当前保存的是来源摘要,不代表论文全文。请以原始来源为准。

「要点解读」由 AI Stack 依据上方已保存内容整理,不代表来源的完整表述;标注「推测:」的判断来自编辑,不是来源陈述。