Science1 distinct publisher3 min readUpdated
A Scientific Reports workflow reports a held-out AUC of 0.91 against 0.81 for PCA and 0.75 for mRMR on the UMMC benchmark. The authors call the findings preliminary and externally unvalidated.
The Scientist · Science desk
Compiled by The ScientistSomething wrong?How this is made
A group publishing in Scientific Reports has put forward a radiomics workflow for predicting brain tumor recurrence after radiotherapy whose central argument is about the feature space rather than the classifier [1][2]. On the University of Mississippi Medical Center benchmark dataset, the pipeline reports a held-out test AUC of 0.91 against 0.81 for a principal component analysis baseline and 0.75 for mRMR feature selection [5].
The stated diagnosis is familiar to anyone who has tried to fit imaging features to a clinical endpoint: the high-dimensional radiomic feature space raises overfitting risk and cuts generalizability, and it does so worst on small samples [6]. The two standard escapes each cost something. PCA compresses the space but makes the resulting features hard to interpret, and direct feature selection is distorted by multicollinearity among radiomic variables, which are often near-duplicates of one another [7].
The proposed answer is structural discipline instead of a new model. The workflow picks the single most predictive feature family, adds a categorical variable for primary tumor location, and then applies rank-correlation-based hierarchical clustering inside the selected family to collapse redundant features [3]. Restricting the clustering to one family is the load-bearing choice: it keeps features interpretable in the way PCA components are not, while attacking the collinearity that trips up ranking-based selection [3][7].
The second contribution is smaller and more practical. The authors introduce a categorical scaling parameter specifically to stop one-hot encoded clinical variables from dominating the scaled numerical radiomic features [4]. That is an honest admission about preprocessing that most papers skip. A binary indicator column and a standardized texture feature do not carry the same effective weight in a distance-based or regularized model, and with a handful of clinical dummies against dozens of radiomic columns, the encoding choice can quietly decide which signal the model sees.
Now the discount rate. The margins are 0.10 AUC over PCA and 0.16 over mRMR, on one benchmark dataset, at a single held-out evaluation [5][10]. The published abstract does not report how many patients the UMMC dataset contains, how large the held-out split was, confidence intervals on any of the three AUC figures, or whether the categorical scaling parameter was tuned strictly inside cross-validation [11]. In a small-sample setting, which is exactly the setting the paper is designed for, a 0.10 AUC gap can turn on a few patients changing sides, and a newly introduced scaling hyperparameter is precisely the kind of knob that can absorb information from the test split if the protocol is loose. The authors do not oversell: they state the findings are preliminary and require external validation before clinical adoption [8]. The work reports no specific grant funding and no competing interests [9].
What to watch: whether the full text, which is open access under a Creative Commons NonCommercial NoDerivatives license, reports the cohort size and the tuning protocol for the scaling parameter [12][11]. Then whether the pipeline survives a second institution's scans, where scanner and protocol differences usually destroy radiomic feature stability. Until an external cohort exists, the transferable result here is not 0.91; it is the two design habits, family-restricted clustering and explicit categorical scaling, which cost nothing to test against your own baseline [3][4].
Follow any of these and your For You feed starts watching them — no settings page required.
Ranked by verification strength, evidence, and original report placement.
The paper "A new radiomics-based approach for predicting brain tumor recurrence" by Olatunde, Oyetunde, Khasawneh et al. was published in Scientific Reports (2026).
The study combined MRI-derived radiomic features with clinical variables to predict brain tumor recurrence after radiotherapy.
The evaluated workflow, designed for limited-sample settings, identifies the most predictive feature family, incorporates a categorical feature based on primary tumor location, and applies rank-correlation-based hierarchical clustering within the selected feature family.
A categorical scaling parameter was introduced to reduce the dominance of one-hot encoded clinical variables over scaled numerical radiomic features.
On the University of Mississippi Medical Center (UMMC) benchmark dataset, the proposed workflow achieved a held-out test ROC AUC of 0.91, compared with 0.81 for PCA and 0.75 for mRMR baselines, and outperformed previously reported methods on the same dataset.
The authors state that radiomics' high-dimensional feature space increases the risk of overfitting and reduces generalizability, particularly in small-sample datasets.
Evidence-backed comparisons of source perspectives and observed adoption signals. Read the methodology
Which Builder, Operator, and Investor concerns the observed source mix emphasized—not a truth score.
Evidence, demonstrated adoption, hype gap, incentives, and confidence are assessed independently, each on its own current evidence. How these are measured.
Single-dataset abstract, key details unreported
The evidence base is one peer-reviewed abstract reporting one held-out comparison on one institutional benchmark. The metric comparison is specific and internally consistent, and the authors describe their own method clearly, which counts for something. Against that, the supplied text gives no cohort size, no held-out split size, no confidence intervals, and no cross-validation protocol for the newly introduced scaling parameter, and there is no external validation cohort. That leaves a plausible but unverifiable result.
No adoption signal in supplied sources
The cluster contains no release, deployment, benchmark-suite entry, licensing, or usage disclosure indicating anyone has taken up this workflow. The authors condition clinical adoption on future external validation, but an absence of adoption evidence is not itself a measurement, so this dimension is left unmeasured.
Headline metric outruns documented detail, but authors hedge
The single reported figure of 0.91 AUC, with a 0.10 margin over PCA and 0.16 over mRMR, is doing more rhetorical work than the supplied documentation supports: no interval estimates, no cohort or split sizes, and no assurance the new scaling parameter was tuned inside cross-validation rather than against the held-out split. That pushes the gap positive. It stays small because the authors explicitly label the findings preliminary and requiring external validation, and make no clinical or commercial promise.
No grant, no declared competing interests, non-commercial license
Incentive pressure looks low: the authors report no specific public, commercial, or not-for-profit grant and declare no competing interests, and the article carries a non-commercial CC BY-NC-ND license. Residual incentive is the ordinary academic one, since the paper is published by the journal reporting it and its value rests on beating named baselines on the dataset the authors selected.
Clear primary source, but one publisher and abstract-only
Confidence in this assessment is moderate. The provenance is unambiguous - a named peer-reviewed article with DOI, funding and interest declarations, and licence terms - so the factual claims are secure. But the cluster has exactly one publisher, that publisher is the journal itself, and only the abstract is supplied, so judgements about the strength of the result, its generalisability, and any uptake rest on material that was never in evidence.
science
Two Neanderthal pelvises suggest the strange hip belongs to the modern human male1 distinct publisher
science
The ocean is the coolant, and the coolant is warming1 distinct publisher
science
21 language models, one habit: tell them your politics and they adopt them1 distinct publisher
product
The most expensive part of your EV may depend on the driver's right foot1 distinct publisher
Distinct publishers with included, body-backed reporting in this cluster.
1 article · August 19, 2026