Turnitin AI 检测跳过哪些内容?非散文、代码等

你的文档不是每一部分都被 Turnitin 的 AI 检测器分析。FAQ 说模型只分析符合检测条件的正文句子。一篇帮助文章列出了诗歌、剧本和注释书目等它不能可靠检测的类型。FAQ 还说 Turnitin 不追求代码检测。以下是什么算符合条件的文本、什么不算。

HumanPen 团队

· 4 分钟

简短回答

Turnitin 的 AI 检测器不分析你文档的每一部分。FAQ 声明:"This qualifying text includes only prose sentences, meaning that we only analyze blocks of text that are written in standard grammatical sentences and do not include other types of writing such as lists, bullet points (short non-sentence structures), or other non-sentence structures." 模型还"does not reliably detect AI-generated text in the form of non-prose, or code." FAQ 补充:"In addition, we are not pursuing ChatGPT code detection at this time." 一篇 Turnitin 帮助文章提供了更广的列表:"The model does not reliably detect AI-generated text in the form of non-prose, such as poetry, scripts, or code, nor does it detect short-form/unconventional writing such as bullet points, tables, or annotated bibliographies." 这意味着你的 AI 分只反映文档中的正文中的完整句子,不是全部内容。

什么算符合条件的文本

FAQ 定义了模型检查的范围:

"This qualifying text includes only prose sentences, meaning that we only analyze blocks of text that are written in standard grammatical sentences and do not include other types of writing such as lists, bullet points (short non-sentence structures), or other non-sentence structures."(此限定文本仅包含正文中的完整句子,这意味着我们仅分析以标准语法句子编写的文本块,不包含其他类型的写作,如列表、项目符号(简短的非句子结构)或其他非句子结构。)

下一句:"This means that a document containing several different writing types would result in a disparity between the percentage and the highlights."

你看到的百分比不是整个提交的百分比。它只是符合检测条件的正文文本的百分比。如果你的文档 50% 是公式、20% 是项目符号、30% 是散文,AI 分只反映那 30%。这道差在报告页面上长什么样,见怎么读一份 Turnitin AI 检测报告

代码和非散文内容

FAQ 对非散文内容说得很直接:

"The model does not reliably detect AI-generated text in the form of non-prose, or code, nor does it detect short-form/unconventional writing such as bullet points (short non-sentence structures)."(该模型无法可靠地检测非散文或代码形式的AI生成文本,也无法检测诸如项目符号(简短的非句子结构)等短篇/非传统写作。)

FAQ 还声明:"In addition, we are not pursuing ChatGPT code detection at this time."

这不能证明每一个代码块都会被解析器完全排除。官方说法更窄:模型对代码的 AI 检测不可靠,而且 Turnitin 目前不打算做 ChatGPT 代码检测。周围的长篇连续正文仍可能属于合格文本。改写工具也得守住同一份文档里的边界:改写工具会把引用、表格和公式改成什么样

诗歌、剧本和注释书目

一篇 Turnitin 帮助文章提供了模型不能可靠检测的内容类型的更完整列表:

"The model does not reliably detect AI-generated text in the form of non-prose, such as poetry, scripts, or code, nor does it detect short-form/unconventional writing such as bullet points, tables, or annotated bibliographies."(该模型无法可靠地检测非散文形式(如诗歌、剧本或代码)的AI生成文本,也无法检测诸如项目符号、表格或带注释参考书目等短篇/非传统写作。)

下一句:"This means that a document containing several different writing types would result in a disparity between the percentage and the highlights."

这份列表来自一篇帮助文章,不是 FAQ。对诗歌、剧本式对话和注释书目条目,官方公开的是“不能可靠检测”,不是承诺每一段都会被完全排除。

表格和参考书目

2023 年 8 月的 release note 描述了更新:

"We are now able to process long-form prose text in tables."(我们现在能够处理表格中的长篇幅散文文本。)

下一句:"Resubmit to reprocess existing submissions that contain tables."

同一 release note 还说:"Bibliographies are now excluded when processing the AI writing report." 下一句同样要求重新提交已有提交才能生效。

这意味着表格里的长篇连续正文可以被处理。短标签和原始数值不会因为放进表格就自动变成符合检测条件的正文。那次模型更新后,参考书目被排除在 AI 检测报告处理之外;旧提交需要重新提交才能应用这项变化。为什么 Turnitin 看起来把参考文献标红了会把这种情况与 Similarity Report 的匹配分开。

排除、处理与检测不可靠,是三种不同状态

官方页面用了三组不同动词。把它们分开,才能避免把“不能可靠检测”改写成“永远不处理”。

内容形态现行官方表述能安全得出的结论来源日期 / 核验日期
符合标准语法的正文中的完整句子检测能力 FAQ说,合格文本“includes only prose sentences”,也就是只包含标准语法块中的正文中的完整句子。这些文本单位会进入 AI 百分比现行页面 / 2026-08-28
列表、项目符号和短的非句子结构同一 FAQ 说,合格文本不包含这些写作类型。它们不进入 FAQ 所描述的合格文本分母现行页面 / 2026-08-28
代码FAQ 说模型“does not reliably detect”AI 生成代码,并说 Turnitin 目前“not pursuing ChatGPT code detection”。代码检测不可靠;页面没有承诺解析器会把代码一律排除现行页面 / 2026-08-28
诗歌、剧本和注释书目AI Writing Report 指南把这些形态列进模型不能可靠检测的清单。不能把 AI 结果当成对这些形态的可靠覆盖现行页面 / 2026-08-28
表格中的长篇连续正文模型变更日志写道:“We are now able to process long-form prose text in tables.”它可以被处理;要让 2023 年 8 月的变化作用于旧提交,需要重新提交2023-08-09 / 2026-08-28
参考书目同一变更日志写道:“Bibliographies are now excluded when processing the AI writing report.”这是明确的 AI Writing Report 排除项;旧的参考文献高亮段需要重新提交2023-08-09 / 2026-08-28

哪些排除项属于哪份报告

Turnitin 的两份报告使用不同的排除规则。出处里写的报告名称就是边界:

项目官方页面怎么写报告与控制方式
参考书目AI 模型变更日志原文是 "Bibliographies are now excluded when processing the AI writing report.",并要求重新提交旧文件才能重新处理。AI Writing Report;这是本次所查页面里唯一明确写出的 AI 报告排除项
引号内文字与块引用Similarity Report 过滤器指南说,它的过滤器会忽略引号内文字和设置成块引用的文字。Similarity Report 过滤器;不是 AI Writing Report 排除项
文内引用及其所在句同一份 Similarity Report 指南说,cited-text 过滤器会移除识别到的文内引用及其关联句子,并注明经典版报告没有这个设置。新版 Similarity Report 过滤器;能不能用取决于报告版本
脚注与尾注本次查的三份 AI Writing Report 官方页面都没有把这两个词写成排除项;现行 Similarity Report 过滤器指南也没有列出独立的脚注或尾注开关。没有公开依据可把它们视为自动排除在 AI Writing Report 之外

这项否定结论也做了灵敏度对照。2026 年 8 月 28 日,固定语料是 Turnitin 的「Using the AI Writing Report」指南、「Turnitin's AI writing detection capabilities FAQs」和「AI writing detection model」变更日志。每页都取 `document.querySelector('article').innerText`,不归一化空白,逐页量完再相加,不先拼接:6,492 + 30,987 + 9,233 = 46,712 个字符。不区分大小写的子串计数里,quotation、quote、footnote、endnote、citation、verbatim、statute、law、legal 全部是 0。相同提取同时数到 prose 13 次、qualifying 15 次、bibliograph 4 次,所以零不是空页面造成的。它不能证明未公开的解析器会怎样处理每一种排版,只能证明公开的 AI 报告页面没有承诺这些排除项。

机制如何在符合条件的文本上工作

检测模型通过把符合检测条件的正文切成重叠片段来处理:

"When a paper is submitted to Turnitin, sentences from the submission are extracted and segmented into overlapping sections for prediction analysis. Each segment is classified by the AI detection model and given a value between 0 and 1, denoting the probability of the text being likely human or AI-generated."(当论文提交至 Turnitin 时,系统会从提交的内容中提取句子,并将其分割成重叠的片段以进行预测分析。每个片段都会由 AI 检测模型进行分类,并赋予一个介于 0 到 1 之间的值,表示该文本可能由人类撰写或由 AI 生成的概率。)

只有符合检测条件的正文句子被提取和切分。非散文内容不包含在此过程中。文档级百分比只从符合检测条件的正文片段计算。

这对你意味着什么

总结一下我们讲的内容:

  • AI 检测器只分析用标准语法句子写的符合检测条件的正文句子。
  • 列表、项目符号和短的非句子结构不在合格文本定义里。
  • 对代码,官方说法是检测不可靠、目前不做 ChatGPT 代码检测,不是承诺一律排除。
  • FAQ 说 Turnitin 不追求代码检测。
  • 一篇帮助文章列出诗歌、剧本和注释书目为不可靠检测。
  • 表格中的长篇连续正文可以被处理;短标签和原始数值仍要满足符合检测条件的正文规则。
  • 参考书目在 AI 检测处理中被排除。
  • 百分比只反映符合检测条件的正文,不是整个提交。

如果收到 Turnitin AI 报告并想处理被标记的段落,可以导入报告处理。符合条件时可以免费继续降 AI。

继续阅读