好词汇被标记为 AI?原因在这里
很多同学反映自己最好的写作——词汇强、句子润色过的那种——被标记为 AI。原因不是好写作像 AI。而是检测器读词概率模式,某些正式学术散文可能和 AI 生成文本共享统计特征。
HumanPen 团队
· 9 分钟
简短回答
用好词汇本身不会导致 AI 标记。检测器不评估你选词的质量。它评估的是你词序列的统计概率。使用一致句子结构、标准转折词和可预测词汇模式的正式学术散文,可能产出和模型与 AI 生成文本关联的词概率特征有重叠的模式。问题不是你的词汇太好。而是你的文本统计特征,由统一结构和正式词汇模式塑造,可能像机器生成的写作。Turnitin 的 FAQ 把"content without a lot of structural variation"列为误判易发特征。高级词汇配统一结构恰好匹配这个描述。
检测器怎么读你的词
Turnitin 的 FAQ 描述了过程:
"When a paper is submitted to Turnitin, sentences from the submission are extracted and segmented into overlapping sections for prediction analysis. Each segment is classified by the AI detection model and given a value between 0 and 1, denoting the probability of the text being likely human or AI-generated. Each qualifying sentence within these segments inherits the segment's score. Since segments overlap, some sentences may have multiple scores, which are then pooled into a single score. These sentence scores are further aggregated and used to compute the overall document AI writing score."
模型不是在读你的词汇然后判断它对一个人来说太高级了。它是在按词概率模式分类片段。FAQ 上的两段话澄清了模型测量什么:"Our model is not explicitly programmed to evaluate specific signals such as "burstiness," "perplexity," or other individual metrics sometimes referenced in public discussions." 以及:"Our classifiers are trained to detect these differences in word probability and are adept at the particular word probability sequences of human writers."
模型不计算一个叫 perplexity 的指标。但它在词概率数据上训练。持续使用相同转折词、相同句首和相同词汇模式的正式学术写作,可能产出比自然人类写作的多变、不可预测模式更像 AI 生成文本的词概率序列。困惑度与突发性之外,AI 检测器还在测什么是这件事的长版本。
为什么润色过的写作匹配误判特征
FAQ 列出了误判易发文本:
"Sometimes false positives (incorrectly flagging human-written text as AI-generated), can include content without a lot of structural variation, text that literally repeats itself, or text that has been paraphrased without developing new ideas."
下一句:"If our indicator shows a higher amount of AI writing in such text, we advise you to take that into consideration when looking at the percentage indicated."
编辑良好的学术散文可能匹配第一条。当你润色写作时,你可能标准化句子结构,使用一致的转折词,全文使用正式词汇。这产出了"a lot of structural variation"低的文本。从文学角度让你的写作更好的润色,在统计角度可能让文本更统一,而统一正是误判描述指向的。为什么写得好的文章反而被判成 AI整篇讲的就是这件事。
分数不是对质量的判断
FAQ 还声明:
"Hence, we must emphasize that the percentage on the AI writing indicator should not be used as the sole basis for action or a definitive grading measure by instructors."
AI 分不是质量评估。它不意味着你的写作太好、太润色或太正式。它意味着模型对你词概率模式的分类产出了某个百分比。在写得好、词汇丰富的文本上的高分并不否定你作品的质量。它意味着你文本的统计特征和模型与 AI 关联的模式有重叠。要不要因此改自己的写法是另一个问题:为了不被标红,要不要改自己的写法。
怎么办
总结一下我们讲的内容:
- 好词汇本身不导致 AI 标记。检测器读词概率模式,不是词汇质量。
- 结构统一的正式学术散文可能和 AI 生成文本共享统计特征。
- Turnitin 把"content without a lot of structural variation"列为误判易发特征。
- 润色可能减少结构变化,让文本在统计上更统一。
- 分数不是对写作质量的判断,不应作为唯一依据。
如果你有 Turnitin 报告显示哪些段落被标记了,可以导入报告专门处理那些段落。符合条件时可以免费继续降 AI。
继续阅读