Model assessment with renormalization group in statistical learning
Qingguo Wang, Chao Yu, Yong Zhang
Abstract
Qingguo Wang, Chao Yu, Yong Zhang
Abstract
This paper proposes a new method for model assessment based on Renormalization Group. Renormalization Group is applied to the original data set to obtain the transformed data set with the majority rule to set its labels. The assessment is first performed on the data level without invoking any learning method, and the consistency and nonrandomness indices are defined by comparing two data sets to reveal informative content of the data. When the indices indicate informative data, the next assessment is carried out at the model level, and the predictions are compared between two models learnt from the original and transformed data sets, respectively. The model consistency and reliability indices are introduced accordingly. Unlike cross-validation and other standard methods in the literature, the proposed method creates a new data set and data assessment. Besides, it requires only two models and thus less computational burden for model assessment. The proposed method is illustrated with academic and practical examples.
A significance statement is not available in the OpenAlex record.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
This paper proposes a new method for model assessment based on Renormalization Group. Renormalization Group is applied to the original data set to obtain the transformed data set with the majority rule to set its labels. The assessment is first performed on the data level without invoking any learning method, and the consistency and nonrandomness indices are defined by comparing two data sets to reveal informative content of the data. When the indices indicate informative data, the next assessment is carried out at the model level, and the predictions are compared between two models learnt from the original and transformed data sets, respectively. The model consistency and reliability indices are introduced accordingly. Unlike cross-validation and other standard methods in the literature, the proposed method creates a new data set and data assessment. Besides, it requires only two models and thus less computational burden for model assessment. The proposed method is illustrated with academic and practical examples.
Key concepts: Consistency (knowledge bases), Data set, Set (abstract data type), Computer science, Reliability (semiconductor), Data mining, Data modeling, Group (periodic table)