International Journal of Engineering، جلد ۳۳، شماره ۲، صفحات ۲۱۳-۲۲۰

عنوان فارسی
چکیده فارسی مقاله
کلیدواژه‌های فارسی مقاله

عنوان انگلیسی Feature Selection for Small Sample Sets with High Dimensional Data Using Heuristic Hybrid Approach
چکیده انگلیسی مقاله Feature selection can significantly be decisive when analyzing high dimensional data, especially with a small number of samples. Feature extraction methods do not have decent performance in these conditions. With small sample sets and high dimensional data, exploring a large search space and learning from insufficient samples becomes extremely hard. As a result, neural networks and clustering algorithms perform poorly on this kind of data. In this paper, a novel hybrid feature selection technique is proposed, which can reduce drastically the number of features with an acceptable loss of prediction accuracy. The proposed approach operates in multiple stages, starting by removing irrelevant features with a low discrimination power, and then eliminating the ones with low variation range. Afterward, among each set of features with high cross-correlation, a single feature that is strongly correlated with the output is kept. Finally, a Genetic Algorithm with a customized cost function is provided to select a small subset of the remainder of features. To show the effectiveness of the proposed approach, we investigated two challenging case studies with sample set sizes of about 100 and the number of features larger than 1000. The experimental results look promising as they showed a percentage decrease of more than 99% in the number of features, with a prediction accuracy of more than 92%.
کلیدواژه‌های انگلیسی مقاله

نویسندگان مقاله M. Biglari |
Computer Engineering and IT Department, Shahrood University of Technology, Shahrood, Iran

F. Mirzaei |
Computer Engineering and IT Department, Shahrood University of Technology, Shahrood, Iran

H. Hassanpour |
Computer Engineering and IT Department, Shahrood University of Technology, Shahrood, Iran


نشانی اینترنتی http://www.ije.ir/article_103369_1f0e61951d122be9176a407f43dfd32d.pdf
فایل مقاله اشکال در دسترسی به فایل - ./files/site1/rds_journals/409/article-409-2284819.pdf
کد مقاله (doi)
زبان مقاله منتشر شده en
موضوعات مقاله منتشر شده
نوع مقاله منتشر شده
برگشت به: صفحه اول پایگاه   |   نسخه مرتبط   |   نشریه مرتبط   |   فهرست نشریات