Skip to main content

Lund University Publications

LUND UNIVERSITY LIBRARIES

Ranking group-level outcomes with multilevel models : An information-theoretic measure of statistical separation

Bashir, Nasir Z. ; Merlo, Juan LU orcid and Leckie, George LU (2026) In Social Science and Medicine 403.
Abstract

Social epidemiologists frequently aim to quantify how social, spatial, or organizational contexts shape individual outcomes, an aim commonly addressed through the use of multilevel models. These models readily estimate the magnitude of group-level differences, but it is more difficult to formalize the uncertainty in the relative ordering of predicted group-level outcomes, which is often considered qualitatively in practice. We propose an entropy-based coefficient, grounded in information theory, which quantifies the stability of predicted group-level rankings. This metric, termed the separation statistic (S), integrates both the magnitude of group-level differences and their statistical uncertainty, providing a principled summary of how... (More)

Social epidemiologists frequently aim to quantify how social, spatial, or organizational contexts shape individual outcomes, an aim commonly addressed through the use of multilevel models. These models readily estimate the magnitude of group-level differences, but it is more difficult to formalize the uncertainty in the relative ordering of predicted group-level outcomes, which is often considered qualitatively in practice. We propose an entropy-based coefficient, grounded in information theory, which quantifies the stability of predicted group-level rankings. This metric, termed the separation statistic (S), integrates both the magnitude of group-level differences and their statistical uncertainty, providing a principled summary of how well groups are separated in terms of their predicted outcomes. Our motivation is drawn from Multilevel Analysis of Individual Heterogeneity and Discriminatory Accuracy (MAIHDA), a widely used approach in social epidemiology for assessing group-level heterogeneity with multilevel models. We show how the separation statistic can be applied to group-level predictions derived from MAIHDA models and is compatible with both Bayesian and frequentist estimation approaches. The metric can be computed globally, across all groups, or locally, within specific subsets of interest. We demonstrate its utility using applied examples from intersectional MAIHDA and provide accompanying code to facilitate its use in future studies. By quantifying the stability of group-level predictions, the separation statistic offers a broadly applicable tool for describing certainty in the relative ordering of outcomes from multilevel models. Importantly, it should be interpreted as a descriptive measure of ranking uncertainty rather than a prescriptive target, with its limitations carefully considered in applied settings.

(Less)
Please use this url to cite or link to this publication:
author
; and
organization
publishing date
type
Contribution to journal
publication status
published
subject
keywords
Entropy, Epidemiology, Information theory, MAIHDA, Multilevel model, Statistics
in
Social Science and Medicine
volume
403
article number
119391
publisher
Elsevier
external identifiers
  • pmid:42142492
  • scopus:105039066736
ISSN
0277-9536
DOI
10.1016/j.socscimed.2026.119391
language
English
LU publication?
yes
id
3122d510-65dd-445b-a53a-61fe8ebee0e0
date added to LUP
2026-08-27 11:47:56
date last changed
2026-09-10 12:38:40
@article{3122d510-65dd-445b-a53a-61fe8ebee0e0,
  abstract     = {{<p>Social epidemiologists frequently aim to quantify how social, spatial, or organizational contexts shape individual outcomes, an aim commonly addressed through the use of multilevel models. These models readily estimate the magnitude of group-level differences, but it is more difficult to formalize the uncertainty in the relative ordering of predicted group-level outcomes, which is often considered qualitatively in practice. We propose an entropy-based coefficient, grounded in information theory, which quantifies the stability of predicted group-level rankings. This metric, termed the separation statistic (S), integrates both the magnitude of group-level differences and their statistical uncertainty, providing a principled summary of how well groups are separated in terms of their predicted outcomes. Our motivation is drawn from Multilevel Analysis of Individual Heterogeneity and Discriminatory Accuracy (MAIHDA), a widely used approach in social epidemiology for assessing group-level heterogeneity with multilevel models. We show how the separation statistic can be applied to group-level predictions derived from MAIHDA models and is compatible with both Bayesian and frequentist estimation approaches. The metric can be computed globally, across all groups, or locally, within specific subsets of interest. We demonstrate its utility using applied examples from intersectional MAIHDA and provide accompanying code to facilitate its use in future studies. By quantifying the stability of group-level predictions, the separation statistic offers a broadly applicable tool for describing certainty in the relative ordering of outcomes from multilevel models. Importantly, it should be interpreted as a descriptive measure of ranking uncertainty rather than a prescriptive target, with its limitations carefully considered in applied settings.</p>}},
  author       = {{Bashir, Nasir Z. and Merlo, Juan and Leckie, George}},
  issn         = {{0277-9536}},
  keywords     = {{Entropy; Epidemiology; Information theory; MAIHDA; Multilevel model; Statistics}},
  language     = {{eng}},
  publisher    = {{Elsevier}},
  series       = {{Social Science and Medicine}},
  title        = {{Ranking group-level outcomes with multilevel models : An information-theoretic measure of statistical separation}},
  url          = {{http://dx.doi.org/10.1016/j.socscimed.2026.119391}},
  doi          = {{10.1016/j.socscimed.2026.119391}},
  volume       = {{403}},
  year         = {{2026}},
}