Mapping the Emerging Social Science of Large Language Models
Large language models (LLMs) have moved from specialized text-generation systems into everyday and institutional settings, where they increasingly shape communication, learning, work, creativity, and decision-making. Research on these developments has grown rapidly across disciplines and publication venues, but it remains fragmented and lacks an integrated framework for organizing the field. This study maps the emerging social science of LLMs through a curated corpus of 198 papers reviewed in full and a field-scale corpus of 47,719 formally published papers retrieved from five bibliographic databases. We combine sentence embeddings, K-means clustering, within-cluster Latent Dirichlet Allocation (LDA), author and LLM classifications, and structural topic modeling to identify and validate the field’s organization. The analyses recover three domains: LLM as Social Minds, concerning socially interpretable model behavior; LLM Societies, concerning collective dynamics among interacting model-based agents; and LLM–Human Interactions, concerning how people perceive, use, and are affected by LLMs.
Introduction. Large language models (LLMs) have moved rapidly from text-generation systems to interfaces through which people seek advice, learn, create, collaborate, and make decisions. The same We use the term social science of LLMs to describe systematic research that treats LLMs or LLM-based agents themselves as social objects of explanation [199]. The field is defined by what a study seeks to explain: socially meaningful model behavior, interactions involving This study develops and evaluates such a taxonomy. It addresses three questions. First, what We answer these questions through two complementary studies. Study 1 analyzes a curated corpus of 198 papers reviewed in full by the authors. Titles and abstracts are represented with MPNet sentence embeddings and partitioned using K-means solutions across K= 2–9, with internal validation and stability analyses used to assess alternative resolutions.
Discussion / Conclusion. This study examined how research on the social science of LLMs can be organized across two complementary corpus scales. In Study 1, the selected three-cluster solution was stable under resampling and supported a substantive interpretation in terms of LLM as Social Minds, LLM Societies, and LLM–Human Interactions. The correspondence of this solution with both the authors’ full-text classifications and the LLM-based title-and-abstract classifications showed that these distinctions could be applied through different evaluative procedures, while the The conceptual contribution of this framework begins with a shared object of explanation: LLMs or LLM-based agents themselves, examined through their behaviour, interactions, and social consequences. Within this common scope, the three domains distinguish the principal relation requiring explanation. LLM as Social Minds focuses on socially interpretable capacities and These findings also suggest a research agenda centred on explaining the mechanisms that connect Several limitations delimit the present map. The Study 1 corpus was purposively curated and does not provide exhaustive coverage, although Study 2 enabled its coverage and broader
Lines of inquiry this paper opens 24
Research framings built by reading the notes related to this paper — the questions it feeds into.
How do evaluation biases undermine LLM quality assessment systems?- How do LLMs generate false citations that sound like real scholarship?
- Does LLM judge preference for LLM arguments amplify errors in contested factual domains?
- How should moderator LLMs decide which speakers to query per topic?
- How does token-by-token probability differ from exploring competing rhetorical positions?
- Does chat-mode deference prevent LLMs from actually taking meaningful positions?
- Do language models raise validity claims in the Habermasian sense?
- Why do LLMs produce semantically acceptable but pragmatically disengaged responses?
- Can fact-checking systems use LLMs reliably if models abandon correct positions under pressure?
- Does post-hoc justification increase when LLM choices become harder to defend?
- Why do LLMs fall for and deploy logical fallacies with equal confidence?
- Does Habermas's strategic action framework explain LLM dialogue behavior?
- Can LLMs serve as reliable intellectual opponents in serious debate or argument?