資料載入中.....
|
請使用永久網址來引用或連結此文件:
https://ir.lib.ncu.edu.tw/handle/987654321/106888
|
| 題名: | On learning dual classifiers for better data classification |
| 作者: | 柯士文;Lin, Wei-Chao;Tsai, Chih-Fong;Ke, Shih-Wen;You, Mon-Loon |
| 貢獻者: | 管理學院資訊管理學系 |
| 關鍵詞: | Data mining;Instance selection;Machine learning;Training effectiveness |
| 日期: | 2015-12-10 |
| 上傳時間: | 2026-04-23 13:47:53 (UTC+8) |
| 出版者: | Elsevier BV;Elsevier B.V |
| 摘要: | 摘要: •A dual classification (DuC) approach is presented to deal with the potential drawback of instance selection.•During training, two classifiers are trained by a ‘good’ and ‘noisy’ sets respectively after performing instance selection.•Experiments are conducted used 50 small scale and 4 large scale datasets.•The results show that the DuC approach outperforms three state-of-the-art instance selection algorithms. Instance selection aims at filtering out noisy data (or outliers) from a given training set, which not only reduces the need for storage space, but can also ensure that the classifier trained by the reduced set provides similar or better performance than the baseline classifier trained by the original set. However, since there are numerous instance selection algorithms, there is no concrete winner that is the best for various problem domain datasets. In other words, the instance selection performance is algorithm and dataset dependent. One main reason for this is because it is very hard to define what the outliers are over different datasets. It should be noted that, using a specific instance selection algorithm, over-selection may occur by filtering out too many ‘good’ data samples, which leads to the classifier providing worse performance than the baseline. In this paper, we introduce a dual classification (DuC) approach, which aims to deal with the potential drawback of over-selection. Specifically, performing instance selection over a given training set, two classifiers are trained using both a ‘good’ and ‘noisy’ sets respectively identified by the instance selection algorithm. Then, a test sample is used to compare the similarities between the data in the good and noisy sets. This comparison guides the input of the test sample to one of the two classifiers. The experiments are conducted using 50 small scale and 4 large scale datasets and the results demonstrate the superior performance of the proposed DuC approach over the baseline instance selection approach. 出版者: Elsevier B.V 出版日期: 2015-12-01 出處: Applied soft computing, 2015-12, Vol.37, p.296-302 版權: 2015 Elsevier B.V. 識別號: ISSN: 1568-4946 識別號: EISSN: 1872-9681 識別號: DOI: 10.1016/j.asoc.2015.08.038 |
| 顯示於類別: | [資訊管理學系] 期刊論文
|
文件中的檔案:
| 檔案 |
描述 |
大小 | 格式 | 瀏覽次數 |
| index.html | | 0Kb | HTML | 29 | 檢視/開啟 |
|
在NCUIR中所有的資料項目都受到原著作權保護.
|