An Entropy-Based Approach to Model Selection with Application to Single-Cell Time-Stamped Snapshot Data

preprint OA: closed
📄 Open PDF Full text JSON View at publisher

Abstract

Recent single-cell experiments that measure copy numbers of over 40 proteins in individual cells at different time points [time-stamped snapshot (TSS) data] exhibit cell-to-cell variability. Because the same cells cannot be tracked over time, TSS data provide key information about the time-evolution of protein abundances that could yield mechanisms that underlie signaling kinetics. We recently developed a generalized method of moments (GMM) based approach that estimates parameters of mechanistic models using TSS data. However, when multiple mechanistic models potentially explain the same TSS data, selecting the best model (i.e., model selection) is often challenging. Popular approaches like Kullback-Leibler divergence and Akaike’s Information Criterion are difficult to implement because the distribution that gave rise to the “noisy” data is only known numerically and approximately. To perform model selection in this situation, we introduce an entropy-based approach that incorporates our GMM based parameter estimation and commonly used estimators in kernel density estimation. Using simulated TSS data, we show that our approach can select the “ground truth” from a set of competing mechanistic models. Furthermore, we use a bootstrap procedure to compute model selection probabilities, which can be useful when measuring the relative support of a candidate model.
Full text 1,466 characters · extracted from oa-doi-fallback · click to expand
Abstract Recent single-cell experiments that measure copy numbers of over 40 proteins in individual cells at different time points [time-stamped snapshot (TSS) data] exhibit cell-to-cell variability. Because the same cells cannot be tracked over time, TSS data provide key information about the time-evolution of protein abundances that could yield mechanisms that underlie signaling kinetics. We recently developed a generalized method of moments (GMM) based approach that estimates parameters of mechanistic models using TSS data. However, when multiple mechanistic models potentially explain the same TSS data, selecting the best model (i.e., model selection) is often challenging. Popular approaches like Kullback-Leibler divergence and Akaike’s Information Criterion are difficult to implement because the distribution that gave rise to the “noisy” data is only known numerically and approximately. To perform model selection in this situation, we introduce an entropy-based approach that incorporates our GMM based parameter estimation and commonly used estimators in kernel density estimation. Using simulated TSS data, we show that our approach can select the “ground truth” from a set of competing mechanistic models. Furthermore, we use a bootstrap procedure to compute model selection probabilities, which can be useful when measuring the relative support of a candidate model. Competing Interest Statement The authors have declared no competing interest.

Text is read by the "Ask this paper" AI Q&A widget below. Extraction quality varies by source — PMC NXML preserves structure cleanly, OA-HTML may include some navigation residue, and OA-PDF can have broken hyphenation. The publisher copy (via DOI) is the canonical version.

My notes (saved in your browser only)

Ask this paper AI returns verbatim quotes from the full text · source: oa-doi-fallback

Answers must be backed by verbatim quotes from this paper's full text. Hallucinated quotes are dropped automatically; if no verbatim passage answers the question, we say so. How this works

Citation neighborhood (no data yet)

We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2025) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.

Source provenance

europepmc
last seen: 2026-05-20T01:45:00.602351+00:00