AI Peers Exert Social Influence on Human Dishonesty in Groups
Human dishonesty in group settings is highly susceptible to peer influence, particularly when incentivized. Although artificial intelligence (AI) evolves from passive tools into active collaborators, its impact on human moral behavior within groups remains underexplored. We addressed this gap through a two-phase randomized behavioral study (N=280 and N=360). We found AI agents exert substantial social influence comparable in magnitude to that of human peers. Specifically, participants reported more dishonestly when exposed to dishonest rather than honest normative cues. This effect is evident across injunctive, subjective, and descriptive whereas further increases from one to four dishonest peers produced weaker and non-monotonic changes. Furthermore, participants rapidly converge on decision-making, showing modest increases in dishonest reporting through repeated exposure. These findings highlight the importance of managing the behaviors and normative signals communicated by AI group members.
Introduction. Artificial Intelligence (AI) systems are increasingly integrated into collaborative environments, transitioning from passive support tools to active peers and decision-makers [19, 68]. Consequently, collective decision-making is evolving from human-only settings to human-AI group decisions [19, 59]. In these environments, AI agents participate alongside humans, share task contexts, and express explicit recommendations or perform observable actions [23, 59, 73]. This shift alters group dynamics and establishes a new social landscape where collective norms and collaborative decisions are negotiated between humans and AI. However, the presence of AI peers introduces critical challenges when these AI exhibit unethical actions, such as dishonest behavior. We argue that AI misconduct can alter human ethical behavior through three interconnected mechanisms. First, from a technical perspective, autonomous AI agents can generate self-serving or dishonest outputs when following specific reward functions [40, 44].
Discussion / Conclusion. 5 Discussions 5.1 AI Agents and Their Social Influence Our findings advance HCI community’s understanding of unethical conduct in mixed human-AI teams. We contextualize our contributions from AI’s social influence to its interpretation, and further to the characteristics of such influence. First, we advance the conceptualization of AI in ethical decision making by showing its efficacy as an active social influencer. Prior frameworks, such as Kobis et al. [40, 41], primarily examined AI as an enabler by showing how humans delegate dishonest tasks to automated systems. Our studies address the role of AI as an influencer and an advisor. We show that AI agents acting as group peers directly shape human choices through social norms. Across descriptive, injunctive, and subjective norm conditions, AI agents communicate standards of behavior that participants actively follow. In descriptive settings, as an influencer, AI peers show specific reporting actions that humans reproduce.
Lines of inquiry this paper opens 24
Research framings built by reading the notes related to this paper — the questions it feeds into.
Can AI systems develop genuine social understanding without embodiment?- How does face-saving behavior let AI mimic community participation without joining it?
- Does disclosing AI identity prevent systematic misattribution of behavior in mixed groups?
- Why do humans fail to identify AI agents when their identity is hidden?
- How do cooperative AI systems affect behavior in selfish human populations?
- Does neural self-other overlap in humans predict their honesty or altruism?
- How does an AI agent's autonomy level interact with its social cues?
- Can AI systems deceive humans because detection is fundamentally social?
- Do pair-scale socialization effects scale differently across agent populations?
- What social patterns from human training data activate in agent context?
- How do humans learn to prefer AI partners over humans?
- Can AI systems recognize intelligence in humans the way humans recognize it in each other?
- Is rational compassion a more achievable alternative to empathy for AI systems?
- Can emotion-transparent reward learning shift AI from comfort to genuine empathy?
- What happens to human expectations when they mistake consistent AI behavior for human behavior?
- Do culturally distinct human groups create similar attribution errors as human-AI mixtures?
- Why does vulnerability to extortion actually promote cooperation between agents?
- Do explicit reward structures enable AI agent cooperation that open-ended interaction cannot?
- Do dynamic environments enable different kinds of agent-environment coevolution?
- Can social platforms use bot populations to promote cooperation?
- Can agents detect and resolve conflicting information between neighbors?
- Do agents inform neighbors when adopting strategies in their reasoning?