102
Studies
Evaluation and psychometric validity
Instruments and benchmarks examining what is measured, how reliably, and whether inferred or declared traits are stable and valid.
Explore trackCritical review · literature cutoff July 18, 2026
An academic map of 385 unique studies, organized into 6 research tracks. All 385 records were written after reading the full source and distinguish method, results, limitations, and what each study does not establish.
102
Studies
Instruments and benchmarks examining what is measured, how reliably, and whether inferred or declared traits are stable and valid.
Explore track65
Studies
Prompting, fine-tuning, and editing activations or weights to elicit, represent, modulate, or suppress traits and behaviors.
Explore track68
Studies
Design and evaluation of characters, profiles, and persistent agents, including role-play, personalization, memory, identity, and consistency.
Explore track45
Studies
Use of LLMs to study or approximate human group opinions, values, and decisions, as well as multi-agent interaction, networks, and collective phenomena.
Explore track88
Studies
Applications in health, education, recommendation, finance, and design, together with risks involving bias, discrimination, privacy, manipulation, and harm.
Explore track17
Studies
Field reviews, conceptual frameworks, methodological requirements, and proposals for responsible research and governance.
Explore trackA 3x3x5x3 design with ChatGPT-5.4, Claude 4.6 Sonnet, and Gemini 3.1 Pro; three names, five domains, and three registers produce 135 responses through public interfaces. Grounded theory generates ten phenomena; a second coder reviews the corpus. A 136-feature pipeline applies ANOVA, chi-square, regression, a 1,000-bootstrap, K-means, and HDBSCAN.
Track: Society, culture, and collective behavior
Track: Evaluation and psychometric validity
Track: Trait induction and control
Track: Reviews, theory, and governance
Track: Applications, bias, and safety