发布于 更新于 Published Updated

You found the paper. Does it support the claim?

A writing assistant returns a valid paper link. Before keeping the sentence beside that citation, ask what a reader could verify by following it. A bibliography entry identifies a source; an evidence record explains why that source belongs beside this particular sentence.

I coauthored ScholarCopilot. This AI-assisted note offers an editorial checking workflow, with three handwritten claims. They are not model outputs, a new benchmark, or experimental results. The downloads contain records for human review; they do not automatically decide whether evidence supports a statement.

Where retrieval fits

ScholarCopilot joins scholarly writing with learned citation retrieval. Its [RET] token requests references during generation, and generation and retrieval are trained jointly. That makes it relevant to research on writing assistants and retrieval timing. The mechanism alone does not certify every sentence. See the original paper and official implementation.

A complementary evaluation perspective appears in ALCE, which studies generated text with citations and distinguishes fluency, correctness and citation quality. These works address different evaluation settings; this note does not compare their scores.

Three claims, three decisions

The examples below inspect the ScholarCopilot v2 abstract, not the whole paper. This deliberately small scope makes the evidence boundary visible. Open the source before interpreting a verdict.

Handwritten claim Check against the cited abstract Editorial decision
Generation and citation retrieval are optimized jointly. The abstract states this training choice. Keep, with the source.
The reported 40.1% means 40.1% of generated claims are factually correct. The number is identified as top-1 retrieval accuracy on the authors’ evaluation dataset. It is not a claim-level factuality rate. Replace the metric description; do not relabel the number.
Every generated sentence is guaranteed to have complete evidence. This guarantee is not established by the inspected abstract. Leave unsupported by this source; inspect further evidence or remove the guarantee.

The third decision is limited: “not established by the source I checked” does not prove the opposite, nor does it claim the entire paper was searched. In the second row, making the sentence sound less certain would not repair the change in metric. The measured object itself must match.

Use the worksheet on one paragraph

Download the blank record and three completed examples. Each example is explicitly marked as handwritten teaching material. No account or API key is needed to read or edit the JSON files.

  1. Copy one substantive sentence exactly. Split it if it contains several independently checkable propositions.
  2. Identify the original source and version. A search snippet is a lead; record where you actually inspected evidence.
  3. Write the relevant passage location and a short paraphrase. Record the task, population, metric and conditions needed to interpret it.
  4. Separate source identity from support. A reachable URL does not settle either question by itself.
  5. Choose a revision, record unresolved questions and leave reviewer/time fields empty until a real review occurs.

For a number, check the denominator, measured object, dataset and comparison. For a causal statement, inspect the design supporting that interpretation. If evidence comes from several sources, retain separate records so each source’s contribution stays visible. These are proposed editorial steps, not a validated automatic scoring system.

A useful first exercise is to audit three adjacent sentences in your own related-work paragraph. Keep the original wording beside the revision. The result should let another reader inspect the decision without having to trust the assistant that drafted it.

ScholarCopilot research guide and citation files · All publications

检索到了论文,下一步怎么核查主张?

写作助手给出了一个能打开的论文链接。保留旁边那句话之前,先问:读者沿着这个链接,究竟能核实什么?书目条目标识来源;证据记录解释这份来源为什么能放在这句话旁边。

我是 ScholarCopilot 的合著者。这篇笔记由 AI 辅助准备,提供的是编辑核查流程。下面三个主张是手写教学材料,不是模型输出、新基准或实验结果。下载文件供人审阅,不会自动判定证据是否支持主张。

检索在流程中的位置

ScholarCopilot 联合训练学术写作与引用检索,并用 [RET] 在生成中请求文献;适合研究写作助手与检索时机。机制本身不能认证每句话。原论文官方实现提供具体定义。

ALCE从带引用文本的评测出发,区分流畅性、正确性与引用质量。两项工作的设置不同,这里不比较分数。

三个主张,三个处理决定

这次只核查 ScholarCopilot v2 的摘要。范围写清楚,才能知道哪些问题还没有回答。

手写主张 与所查摘要对照 编辑决定
生成和引用检索接受联合优化。 摘要说明了这项训练选择。 保留并引用来源。
40.1% 表示生成主张中事实正确的比例。 原文数字是评测集上的 top-1 检索准确率。 修正指标描述,不能偷换统计对象。
每个生成句子都保证有完整证据。 所查摘要没有建立这项保证。 保留未证实状态;继续核查或删除保证。

第三行的结论仅是“所查来源没有证实”,不证明相反命题,也不意味着全文已经查完。第二行需要修正被测对象;把语气改得更谨慎,仍然不能把检索指标变成主张事实性指标。

拿一段自己的文章试用

下载空白记录三个填写示例。示例显式标记为手写教学材料,使用统一的英文 JSON 字段;无需登录或 API key,即可阅读和编辑。

  1. 原样抄录一个实质性句子;含多个可独立核查命题时先拆开。
  2. 标识原始来源和版本。搜索摘要可作线索,但要记录实际检查到哪里。
  3. 填写证据位置和简短转述,保留任务、对象、指标和适用条件。
  4. 将来源身份与主张支持分开判断。链接能打开,不能独自解决这两个问题。
  5. 记录修改决定、未解决的问题;未实际审阅时,审阅人和时间保持空白。

遇到数字,核对分母、测量对象、数据集和比较条件;遇到因果表述,核对支持这种解释的研究设计。需要多份来源时分别记录,让每份证据承担的部分清楚可见。这些是建议的编辑步骤,不是经过验证的自动评分系统。

今天就可以检查 related work 里相邻的三个句子,把原句与修订并列留下。目标是让下一位读者能检查你的决定,而不必直接相信生成草稿的助手。

ScholarCopilot 说明与引用文件 · 全部论文