SEMbeddings: how to evaluate model misfit before data collection using large-language models
preprint
OA: closed
CC-BY-4.0
Abstract
Recent developments suggest that Large Language Models (LLMs) offer a viable approach to approximate empirical correlation matrices of item responses through the utilization of item embeddings and their cosine similarities. In this paper, we introduce a novel tool, that we label SEMbeddings, which integrates prior findings with the application of latent measurement models to assess model fit or misfit prior to data collection. To support our statement, we use the 96 items of the VIA-IS-P which measures 24 different character strengths and responses from 31,697 participants to those items. Our analysis reveals a significant correlation (r = .56) between cosine similarities of embeddings and empirical correlations among items. Moreover, we demonstrate the feasibility of fitting confirmatory factor analyses on cosine similarity matrices and interpreting their outcomes using modification indices: We found consistency of the results obtained with SEMbeddings procedures and classical procedures on empirical data. While not always identical, these results suggest that SEMbeddings could serve as a screening tool to inspect model fit and to make informed decisions about items before data collection. With the increasing precision of LLMs, these procedures will produce increasingly reliable results, potentially transforming the way we develop new questionnaires.
My notes (saved in your browser only)
Citation neighborhood (no data yet)
We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2024) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.
Source provenance
- europepmc
- last seen: 2026-05-20T01:45:00.602351+00:00
- unpaywall
- last seen: 2026-05-27T02:00:06.600101+00:00
License: CC-BY-4.0