基本信息

要点解读

这是什么

ScienceBuddy 是一个交互式科研工作空间,利用递归‑递归的自改进机制,使科学智能体在执行任务时不断收集反馈并更新训练评估框架,从而提升模型表现。

用在哪里

适用于需要频繁开展科学实验并希望自动积累实验证据、持续优化模型的科研团队,或在 AI 研究中探索自我改进式智能体研发的开发者。

可以推断的

推测:递归的自改进方法可能在长期使用中逐步提升智能体的任务成功率。
推测:该系统在需要快速构建并迭代评估标准的研究项目中具有实际应用价值。

来源摘要/节选

We introduce and release ScienceBuddy, an interactive scientific research workspace that brings continually improving scientific agents into researchers’ everyday workflows. ScienceBuddy supports researchers in carrying out scientific tasks while transforming their requests, feedback, and execution evidence into tasks and evaluation rubrics for continual learning. At its core is recursive-in-recursive self-improvement, a paradigm that couples harness evolution with model reinforcement learning: the inner recursion improves the harness with the model fixed, while the outer recursion trains the model under the improved harness. Harness evolution shapes training experience, and model learning creates new opportunities for harness adaptation. We present case studies of researcher interaction, harness refinement, and model learning, with the benchmark cases spanning four scientific task families. By releasing ScienceBuddy as a research product, we make this paradigm available to the scientific community and take a step toward discovery intelligence: scientific AI that advances through sustained collaboration with researchers and evolves alongside the research it supports. Website: http://science-buddy.io

来源说明

当前保存的是来源摘要,不代表论文全文。请以原始来源为准。

「要点解读」由 AI Stack 依据上方已保存内容整理,不代表来源的完整表述;标注「推测:」的判断来自编辑,不是来源陈述。