A Ship of Theseus: Curious Cases of Paraphrasing in LLM-Generated Texts
Nafis Irtiza Tripto, Saranya Venkatraman, Dominik Macko, Robert Moro, Ivan Srba, Adaku Uchendu, Thai Le, Dongwon Lee
TL;DR
This work reframes authorship in the age of LLM paraphrasing through the Ship of Theseus lens, asking whether original authorship endures when texts are sequentially rewritten as $T^0 \to T^1 \to T^2 \to T^3$. It combines seven datasets, seven authors (six LLMs plus human) and four paraphrasers (including ChatGPT and PaLM2) to study traditional versus alternative ground-truth attributions and AI-detection under varying stylistic and content changes, using both a transparent stylometry pipeline and content embeddings. The key finding is that paraphrasing induces substantial style drift, especially for LLM paraphrasers, which shifts text toward the paraphraser’s own style and degrades attribution performance more than content similarity—yet adopting an alternative ground-truth where authorship follows the paraphraser yields substantially milder declines for many cases. The results highlight that ground truth in authorship and detection is task-dependent, with important implications for plagiarism policies, copyright disputes, and the design of robust AI-text detectors that must account for stylistic transformations introduced by paraphrasing tools. Overall, the study provides empirical and philosophical grounding for nuanced attribution frameworks in an era where paraphrasing and AI-generated writing increasingly blur the boundaries of authorship.
Abstract
In the realm of text manipulation and linguistic transformation, the question of authorship has been a subject of fascination and philosophical inquiry. Much like the Ship of Theseus paradox, which ponders whether a ship remains the same when each of its original planks is replaced, our research delves into an intriguing question: Does a text retain its original authorship when it undergoes numerous paraphrasing iterations? Specifically, since Large Language Models (LLMs) have demonstrated remarkable proficiency in both the generation of original content and the modification of human-authored texts, a pivotal question emerges concerning the determination of authorship in instances where LLMs or similar paraphrasing tools are employed to rephrase the text--i.e., whether authorship should be attributed to the original human author or the AI-powered tool. Therefore, we embark on a philosophical voyage through the seas of language and authorship to unravel this intricate puzzle. Using a computational approach, we discover that the diminishing performance in text classification models, with each successive paraphrasing iteration, is closely associated with the extent of deviation from the original author's style, thus provoking a reconsideration of the current notion of authorship.
