Line of inquiry
Inquiring lines›How does AI assistance reshape hum…›Do persona models reliably predict…›this line of inquiry
Why do language models resist personality conditioning through prompts?
A broader line of inquiry — a family of 19 specific questions the research asks around this. Follow one into its inquiring-line page, or move sideways to a related line below.
Questions in this line of inquiry 19
Specific inquiring lines the field asks around this — ordered from the most general framing down to the most specific angle.
- Why do language models resist adopting different personalities when prompted?
- Can persona prompting overcome the default ENFJ personality in language models?
- Why do models resist personality change despite sophisticated prompting techniques?
- Do open-source LLMs show different resistance patterns to persona prompting than closed models?
- Why do most open language models resist personality conditioning via prompts?
- Why do some open models resist personality conditioning while others don't?
- Why do personas in language models resist correction through prompting alone?
- Can persona framing reduce refusal by providing representational scaffolding?
- Why does model uncertainty dominate persona-specific knowledge in annotation tasks?
- Why do language models prefer certain response styles regardless of what the prompt asks?
- Why do LLM persona annotations become unstable when run multiple times?
- Does combining role and personality prompts produce stable behavioral changes?
- Can prompt engineering fully prevent role flipping in LLM agents?
- What distinguishes personality resistance from persona instability in LLMs?
- What explains why LLM personas fail to instantiate values but succeed in sounding natural?
- How does the dialogue prompt establish the character the model plays?
- How does persona instability in annotation compare to LLM overconfidence in low-resource domains?
- Why do LLM regenerations produce meaningfully different personalities from the same prompt?
- What does the 20-questions test reveal about LLM character consistency?