遇见数据集

Results Accuracies from Constraining classifiers in molecular analysis: invariance and robustness

收藏
DataCite Commons2020-08-26 更新2024-07-28 收录
官方服务:

资源简介:

Analysing molecular profiles requires the selection of classification models that can cope with the high dimensionality and variability of this data. Also, improper reference point choice and scaling pose additional challenges. Often model selection is somewhat guided by <i>ad hoc</i> simulations rather than by sophisticated considerations on the properties of a categorization model. Here, we derive and report four linked linear concept classes/models with distinct invariance properties for high-dimensional molecular classification. We can further show that these concept classes also form a half-order of complexity classes in terms of Vapnik–Chervonenkis dimensions, which also implies increased generalization abilities. We implemented support vector machines with these properties. Surprisingly, we were able to attain comparable or even superior generalization abilities to the standard linear one on the 27 investigated RNA-Seq and microarray datasets. Our results indicate that <i>a priori</i> chosen invariant models can replace <i>ad hoc</i> robustness analysis by interpretable and theoretically guaranteed properties in molecular categorization.

提供机构:
The Royal Society
创建时间:
2020-01-21
二维码
社区交流群
二维码
科研交流群
商业服务