独立研究智能体系统
负责任智能体系统的独立研究。
Harness the Agents 是一家独立、循证的出版物与研究实验室。我们研究 AI 模型如何演化为智能体系统,以及这些系统如何被负责任地运行——涵盖 Harness、运行时、控制平面、治理、权限、记忆、执行、可观测性与恢复。

精选研究
查看全部研究Inside
DeepSeek Harness 不是 Claude Code 杀手,它是更有趣的东西
DeepSeek 的开源智能体 harness 上线第一周 GitHub 星数就突破了 168,000。我装了它,读了仓库,还亲手跑了一遍。以下是我验证下来站得住的东西。
Above
Amodei's Embedded Evaluators: Access Is the Whole Game
Amodei pledges embedded evaluators with employee-like access and publication rights. What the pledge promises, what it leaves open, and what is on paper today.
Inside
An Agent That Works for Weeks Needs More Than Memory
Salesforce's long-horizon runtime lets an Agentforce agent pursue a goal for weeks. Memory, durable execution and dynamic steering are documented. The authority questions are not.
可复现的实验
所有实验都全程记录,结果可验证、可复用。
方法论
设计与协议
版本
代码与依赖
模型
提供方与配置
指标
我们测量什么
执行轨迹
运行日志
结果
分析与统计
局限
威胁与注意事项
原始数据
下载
Agent Field Notes
一份定期简报,汇总我们最新的文章,并链接回原始研究。
我们尊重你的收件箱,可随时退订。
