AI Argues Differently: Distinct Argumentative and Linguistic Patterns of LLMs in Persuasive Contexts
Distinguishing LLM-generated text from human-written is a key challenge for safe and ethical NLP, particularly in high-stake settings such as persuasive online discourse. While recent work focuses on detection, real-world use cases also demand interpretable tools to help humans understand and distinguish LLM-generated texts. To this end, we present an analysis framework comparing human- and LLM-authored arguments using two easily-interpretable feature sets: general-purpose linguistic features (e.g., lexical richness, syntactic complexity) and domain-specific features related to argument quality (e.g., logical soundness, engagement strategies). Applied to /r/ChangeMyView arguments by humans and three LLMs, our method reveals clear patterns: LLM-generated counter-arguments show lower type-token and lemma-token ratios but higher emotional intensity – particularly in anticipation and trust. They more closely resemble textbook-quality arguments – cogent, justified, explicitly respectful toward others, and positive in tone. Moreover, counter-arguments generated by LLMs converge more closely with the original post's style and quality than those written by humans.
Analyzing distributional differences between human and LLM arguments in persuasive discourse, we find substantial differences both in style and argument quality: LLM arguments show higher emotional positivity, stronger convergence with original posts (especially in named entities and psycholinguistic features), and greater alignment with argument quality markers. In contrast, human arguments display more negative emotion, greater lexical and syntactic creativity, and stronger use of interactive discourse.
Moreover, we show that linguistic and argument quality features enable nearly 99% accurate detection of LLM-generated comments to CMV posts from human-written ones. Our approach thus offers a practical safeguard against unethical uses of LLMs in online discussions. Furthermore, tests on an external benchmark show that our lightweight and interpretable method performs comparably to computationally intensive detectors in generalized detection scenarios, highlighting the viability of low-resource, transparent detection methods.
These results prompt important questions for future research: Under what conditions are LLM-generated texts harder to detect? How do the prompt design and task objective influence detectability? How do the convergence patterns of humans and LLMs align with social theories of communication, such as communication accommodation theory? Our framework provides a straightforward and interpretable approach to assess such questions, thereby facilitating future investigations into the nuances of LLM-generated content.
Lines of inquiry this paper opens 24
Research framings built by reading the notes related to this paper — the questions it feeds into.
Why can't humans reliably detect AI-generated text despite measurable linguistic signatures?- Can AI text detectors reliably identify AI-generated websites?
- Why do human judges fail to detect systematic linguistic differences that classifiers easily identify?
- Can readers detect when text was written or heavily influenced by AI?
- What linguistic markers reveal AI text lacks embodied authorship?
- Why does lexical difference fail to trigger reader suspicion of artificial origin?
- What linguistic cues help humans detect whether moral arguments come from AI?
- What properties of natural text does artificial text actually eliminate?
- Why do human judges fail to detect AI text consistently?
- Is statistical analysis the only reliable way to detect modern AI writing?
- Does higher lexical density in fewer tokens indicate systematic AI signature?
- Why do AI signatures exist statistically but remain imperceptible to human judges?
- Can AI arguments participate in discourse without temporal grounding?
- Does conversational format make AI arguments more persuasive than static text?
- Can audiences learn to recognize and resist moralized AI rhetoric?
- Can persuasive equivalence exist without process equivalence in other domains?
- Can readers distinguish between AI and human persuasion on textual surface alone?
- Can probing methods detect RLHF-induced persuasion in the same way they catch backdoors?
- How well can platforms detect AI-generated personalized persuasion attempts?
- Can current AI safety defenses actually stop semantic-level persuasion attacks?