The use of SOFA scores as an outcome measure in RCTs currently lacks sufficient reproducibility and methodological robustness.
There is major variability in the choice of summary statistic for SOFA score analysis and assessment timepoints, when using it as outcome measure in RCTs. There was either no information or great variability in the handling of missing values, use of imputation, and accounting for competing risk. The current use of SOFA scores in RCTs lacks sufficient reproducibility and statistical and methodological robustness.
Marmiere et al. (Wed,) studied this question.