← 返回时间线

paper

Scores Alone Do Not Prove Discovery: The Discovery Certification Protocol for Auditing AI Research Agents

arXiv ↗
ID
2609.09219
分类
首次捕获
2026-09-10
状态
unread
作者
Jingjie Ning, Shanshan Zhong, Xiaochuan Li, Ji Zeng
信号
🔥 11

信号历史

  • 2026-09-10HF Daily Papers · 🔥11
暂无信号数据

摘要

AI research agents combine prior knowledge, public sources, and experimental feedback to produce useful results. The Discovery Certification Protocol (DCP) turns claims about these results into executable recovery and feedback tests. Gate 1 validates useful improvement on sealed evaluation. Gate 2 gives matched agents the registered starting information and observed Web content while withholding the target research history. Every valid method reaching the numerical target supplies a recovery witness and triggers the Core veto. DCP Core requires adequate controls, zero observed recoveries, and a finite-sample bound on recovery in one fresh registered episode. Optional Gate 3 measures the average effect of truthful feedback relative to a specified neutral policy from a shared checkpoint. DCP Evidence adds this effect after independent null calibration and a registered effect margin. Two controlled audits exercise the complete protocol in SQLite optimization and virtual catalyst control under different models. Each produced zero recoveries in 96 episodes, with an upper bound of 0.0468. Each paired study yielded 30 truthful recoveries and zero neutral recoveries, with passing 60-pair null studies. Additional cases exercise Core, recovered, and audit-incomplete decisions. A deterministic, LLM-free verifier reproduces the decisions from frozen evidence. DCP provides a common evidence language for useful outcomes, alternative routes, and feedback effects across AI research.

我的笔记

还没有笔记。

在 GitHub 上写笔记 ↗(新建 content/notes/2609.09219.md,PR 合并后本页自动更新)