Psychometric Validation

Momentum

3 papers in the last four weeks, against 1 the four weeks before. 0.0% of all new papers.

Jul 13Week of Sep 28

Latest papers 48

All topics
CardsList
  1. COMPASS 2.0: psychometric representational similarity analysis distinguishes symptom structure from personal signal

    Oct 5, 2026Baihan LinRepresentational Similarity AnalysisPsychometric Validation

  2. Artificial Societies Benchmark: A Validation Framework for Synthetic Research

    Sep 24, 2026Edoardo Chidichimo, Min Jun Jung, Felix P. S. Wallis +1Benchmark DesignComputational Social Science

  3. Measuring AI Leadership: Development and Validation of a Multidimensional Measure for AI-Native Organizations

    Sep 16, 2026Mustafa Akben, Leslie CoynePsychometric Validation

  4. Automated Generation of Complexity-Validated Decision Scenarios Using Large Language Models

    Aug 9, 2026Abdalla Doleh, Toni Somers, Ratna Babu ChinnamSynthetic Data GenerationPsychometric Validation

  5. Natural Language Processing Psychometrics

    Aug 7, 2026Edoardo Sebastiano De Duro, Emma Franchino, Massimo StellaClinical Outcome PredictionPsychometric Validation

  6. Every Wrong Answer Counts: Option-Level Psychometrics for LLM Multiple-Choice Benchmarks

    Aug 3, 2026Xiao Fei, Yang Zhang, Sarah Almeida Carneiro +1LLM EvaluationPsychometric Validation

  7. Comparative Validation of GPT-4o-mini and Teacher Mean Scores for Automated Scoring of Music Analysis Responses: Single-Pass Deployment, Repeatability, and Strategy-Specific Bias

    Aug 3, 2026Baicheng Lin, Lingxi Jin, Kyung-Seok MinLLM-as-a-JudgeEducational Technology

  8. Dimensionality and Measurement Precision in HLE's Multiple-Choice Subset

    Jul 29, 2026Mayank Sharma, Savira Nadela, Tyler MattesonLLM EvaluationPsychometric Validation

  9. Developing and Validating the Spanish Version of the Large Language Models Dependency Scale (LLM-D12-SP)

    Jul 24, 2026Tran Gia Bao, Mo El-Haj, Sameha Al-Shakhsi +3Human-AI InteractionPsychometric Validation

  10. Bringing Back Rule Induction to Fluid Intelligence Research? An Initial Validation of the ARC-AGI Benchmark in Humans

    Jul 13, 2026Jasmin Thelen, Oliver WilhelmPsychometric Validation

  11. Validating LLMs in social science: Epistemic threats and emerging norms

    Jul 8, 2026Meera Desai, Dallas Card, Abigail Z. JacobsLLM EvaluationComputational Social Science

  12. Measuring Intelligence Beyond Human Scale

    Jul 8, 2026Jerry Han, Rafael Moschopoulos, Ella Colby +6LLM EvaluationAI Agent Evaluation

  13. Correct codes for the wrong reasons? validating LLMs as measurement instruments for theoretical constructs

    Jun 26, 2026Manuel PitaLLM EvaluationPsychometric Validation

  14. Apparent Psychological Profiles of Large Language Models are Largely a Measurement Artifact

    Jun 18, 2026Jelena Meyer, David Garcia, Dirk U. WulffLLM EvaluationPersonality Modeling in Language Models

  15. The Unsampled Truth: Quantifying Prompt Artifacts in LM Psychometrics

    Jun 2, 2026Nils Schwager, Christoph Hau, Simon Münker +1Prompt SensitivityLLM Prompting

  16. GenPT: Beyond Self-Report for Reliable LLM Psychometrics via Generative Projective Testing

    May 30, 2026Ming Wang, Shuang Wu, Bixuan Wang +7Psychometric Validation

  17. AI Cartography: Mapping the Latent Landscape of AI Benchmark Ecosystems

    May 24, 2026Michael Hardy, Anka Reuel, Lijin Zhang +6LLM EvaluationBenchmark Design

  18. Generative-Evaluative Agreement: A Necessary Validity Criterion for LLM-Enabled Adaptive Assessment

    May 19, 2026Grandee Lee, Yue Wang, Che Yee Lye +1LLM EvaluationEducational Assessment

  19. The Association of Transformer-based Sentiment Analysis with Symptom Distress and Deterioration in Routine Psychotherapy Care

    May 11, 2026Douglas K. Faust, Peter Awad, Alexandre Vaz +1Sentiment AnalysisClinical Outcome Prediction

  20. Machine Psychometrics: A Mathematical Psychology of Artificial Intelligence

    May 10, 2026Alex Bogdan, Adrian de Valois-FranklinAI Agent EvaluationAI Agent Monitoring

  21. The Proxy Presumption: From Semantic Embeddings to Valid Social Measures

    May 8, 2026Baishi Li, Ta Yu, Kelvin J. L. Koa +1Text EmbeddingsText Embedding Evaluation

  22. The Pinocchio Dimension: Phenomenality of Experience as the Primary Axis of LLM Psychometric Differences

    May 6, 2026Hubert Plisiecki, Sabina Siudaj, Kacper Dudzic +4LLM EvaluationPersonality Modeling in Language Models