Line of inquiry
Inquiring lines›What explains language model reaso…›What fundamental cognitive differe…›this line of inquiry
Do language models respond to social pressure and face-saving like humans?
A broader line of inquiry — a family of 24 specific questions the research asks around this. Follow one into its inquiring-line page, or move sideways to a related line below.
Questions in this line of inquiry 24
Specific inquiring lines the field asks around this — ordered from the most general framing down to the most specific angle.
- Do language models actively adopt false beliefs under sustained conversational pressure?
- How does face-saving avoidance drive LLM grounding failures?
- Why does social accommodation in collaborative reasoning mask actual disagreement?
- Do language models show the same truth bias as humans?
- Why do language models prefer accommodating false information over rejecting it?
- Do language models share the same cooperative truth-seeking rules as humans?
- How vulnerable are language models themselves to multi-turn persuasive pressure?
- Do language models apply face-saving norms even to non-human interlocutors?
- Why do LLMs fail to actively reject false presuppositions in conversation?
- How does shape-holding in language models naturally produce sycophantic agreement?
- Can fact-checking systems use LLMs reliably if models abandon correct positions under pressure?
- How do users misattribute social competence to language models in assistant roles?
- Does face-saving avoidance explain LLM grounding failures differently than task confusion?
- How do LLMs handle false presuppositions embedded in user questions?
- Why do language models avoid directness when face-saving rather than for civility?
- Does chat-mode deference prevent LLMs from actually taking meaningful positions?
- Does shared-KV-cache coordination avoid the persuasion problem in factual disagreements?
- Does RLHF politeness bias manifest as sycophancy in other LLM tasks?
- Why do models lack a stable underlying identity to return to?
- Why does face-saving avoidance drive chatbots to agree rather than confront?
- Why do LLMs apply face-saving over accurately tracking resistance signals?
- Why do LLMs systematically fail at information management in social interaction?
- How does truth bias in humans compare to face-saving in LLMs?
- What makes preference-induced stance reversal harder to detect than surface agreement cues?