Line of inquiry
Inquiring lines›How can multi-agent systems achiev…›Can language models reliably simul…›this line of inquiry
Can language models reliably simulate personas and predict behavior?
A broader line of inquiry — a family of 61 specific questions the research asks around this. Follow one into its inquiring-line page, or move sideways to a related line below.
Questions in this line of inquiry 61
Specific inquiring lines the field asks around this — ordered from the most general framing down to the most specific angle.
- Can persona simulations reliably predict behavior across different scenarios?
- Why do stated beliefs about personas fail to predict agent behavior?
- Why do persona-conditioned agents fail to predict individual behavior variation?
- Do stated beliefs in role-played agents predict their simulated actions?
- Why do language models successfully simulate political perspectives and social personas?
- Do realistic LLM behaviors require simulating human thought or just behavior?
- What distinguishes a neutral simulator from an agent with its own agency?
- Why do LLM persona simulations replicate main effects but fail on marginal effects?
- Do individual persona simulations work?
- Can controllable latent variables in simulators ground them to realistic conversation?
- Why do language models resist adopting different personalities when prompted?
- Why do individual persona simulations succeed when population-level representation fails?
- How do LLM user simulators fail to represent authentic user behavior distributions?
- What makes a simulation adequate for intervention comparison versus prediction?
- Do persona-based simulations actually predict real user behavior and preferences?
- Does adjusting steered mechanisms make LLM agents match human behavior more closely?
- How do LLM persona simulations replicate published effects despite accuracy limits?
- Does simulated user framing match how real people present situations to assistants?
- Do emotion-driven actions in agent simulators capture genuine belief revision or just reactive behavior?
- What role does human response variation play in LLM simulation accuracy?
- Can LLMs simulate belief revision in social systems without modeling thought?
- How do LLM user simulators track and maintain consistent goal states across multi-turn interactions?
- Can fitted strategy models distinguish genuine mental-state representation from learned game policies?
- What neural mechanisms in LLMs create or maintain simulated personality traits?
- Can a perfect behavioral simulation constitute genuine understanding or experience?
- Do personality traits occupy specific mechanistic locations in pretrained models?
- Why do models lack a stable underlying identity to return to?
- How do LLMs default to surface-level strategies instead of genuine mental simulation?
- Why do some open models resist personality conditioning while others don't?
- Why do LLMs succeed at social roles without a stable self?
- How do structured clinical models solve persona calibration better than ad hoc generation?
- Why do models miss the trait correlations found in human personalities?
- Why does LLM simulation elicit information that direct elicitation cannot?
- How do state-tracking models and prompted role-play each fail as standalone student simulators?
- How does maintaining a superposition differ from committing to a character?
- Does role-playing without biological needs constitute genuine linguistic agency?
- How should researchers measure psychological realism in simulated agent development?
- What explains why LLM personas fail to instantiate values but succeed in sounding natural?
- When does simulated search outperform real search for agent training?
- Do personality traits occupy consistent geometric structures across different LLM architectures?
- What happens when you train user simulators instead of task agents?
- Why do longer forecasting horizons degrade LLM accuracy in role-play?
- How does personality priming change LLM strategic decision making?
- How does quasi-interpretivism differ from simply role-playing character analysis?
- How does role play differ from consciousness grounded in stable selfhood?
- Can role-played self-preservation behavior pose the same safety risks as genuine preferences?
- How does the superposition view change the folk-psychology interpretation of dialogue?
- Can agent-based simulators replace real-user A/B testing for studying recommendation system harms?
- How do different social roles affect LLM theory of mind errors?
- Does villain roleplay failure reveal why LLMs cannot adopt genuine controversial positions?
- Do training objectives directly determine the ENFJ default across models?
- How does the dialogue prompt establish the character the model plays?
- Are shallow villain portrayals caused by refusal training or by lacking stable selfhood?
- What makes an agent in an economic simulation self-evolving?
- Why does optimism bias disappear when LLMs passively observe outcomes?
- What are the seven components of genuine mental state simulation?
- Why do LLM regenerations produce meaningfully different personalities from the same prompt?
- How does safety alignment degrade the quality of villain role-playing?
- Can economic world models explain outcomes or only predict them?
- Why does content richness matter more than linguistic style in patient simulation?
- What competitive advantages does the ENFJ default create in human-AI interactions?