基本信息

要点解读

这是什么

该模型把对已部署 AI 主体的治理抽象为一种机制设计,利用计算资源的分配实现自我执行的授权,并以硬件签名的计算许可证形式兑现决策。

用在哪里

适用于希望在部署阶段引入持续人类监督的 AI 系统,尤其是公共平台或合规层的治理框架;对机制设计、AI 安全与治理研究具有参考价值。

可以推断的

推测:该模型目前以理论建模为主,尚未进行大规模实证验证。
推测:防止被治理主体操纵治理选民是核心难题,说明方案对策略性行为较为敏感,需要额外的防御手段。

来源摘要/节选

We give a formal mechanism design model for the continuous participatory governance of a deployed AI agent. The mechanism is built on the principle that governance should control an AI agent through resource allocation so as to make authorization self enforcing via compute budgets. The mechanism seeks to establish the Safe AI paradigm that compute is an effective governance lever. We situate our work as a compliance or commons overlay on a deployer. One governance period is an extensive form game in which verified human stakeholders arrive sequentially and contribute, on a provision or a rejection market, in a governance currency that is deliberately distinct from the agents compute. A funding aggregator turns raw contributions into breadth weighted effective supports - a two threshold gate with hysteresis converts net support into a binary authorization that, through a coupling map bounded by an exogenously certified safety ceiling, releases a metered compute budget - realized in hardware as a signed compute license so that the decision is self-enforcing. We characterize the class of agents the mechanism can govern and isolate manipulation of the governing electorate by the governed agent as the central open problem. We also introduce several challenges addressing manipulation of governing electorate by the governed agents.

来源说明

当前保存的是来源摘要,不代表论文全文。请以原始来源为准。

「要点解读」由 AI Stack 依据上方已保存内容整理,不代表来源的完整表述;标注「推测:」的判断来自编辑,不是来源陈述。