English  |  正體中文  |  简体中文  |  全文筆數/總筆數 : 94274/94274 (100%)
造訪人次 : 82906540      線上人數 : 2385
RC Version 7.0 © Powered By DSPACE, MIT. Enhanced by NTU Library IR team.
搜尋範圍 查詢小技巧:
  • 您可在西文檢索詞彙前後加上"雙引號",以獲取較精準的檢索結果
  • 若欲以作者姓名搜尋,建議至進階搜尋限定作者欄位,可獲得較完整資料
  • 進階搜尋


    請使用永久網址來引用或連結此文件: https://ir.lib.ncu.edu.tw/handle/987654321/106888


    題名: On learning dual classifiers for better data classification
    作者: 柯士文;Lin, Wei-Chao;Tsai, Chih-Fong;Ke, Shih-Wen;You, Mon-Loon
    貢獻者: 管理學院資訊管理學系
    關鍵詞: Data mining;Instance selection;Machine learning;Training effectiveness
    日期: 2015-12-10
    上傳時間: 2026-04-23 13:47:53 (UTC+8)
    出版者: Elsevier BV;Elsevier B.V
    摘要: 摘要: •A dual classification (DuC) approach is presented to deal with the potential drawback of instance selection.•During training, two classifiers are trained by a ‘good’ and ‘noisy’ sets respectively after performing instance selection.•Experiments are conducted used 50 small scale and 4 large scale datasets.•The results show that the DuC approach outperforms three state-of-the-art instance selection algorithms. Instance selection aims at filtering out noisy data (or outliers) from a given training set, which not only reduces the need for storage space, but can also ensure that the classifier trained by the reduced set provides similar or better performance than the baseline classifier trained by the original set. However, since there are numerous instance selection algorithms, there is no concrete winner that is the best for various problem domain datasets. In other words, the instance selection performance is algorithm and dataset dependent. One main reason for this is because it is very hard to define what the outliers are over different datasets. It should be noted that, using a specific instance selection algorithm, over-selection may occur by filtering out too many ‘good’ data samples, which leads to the classifier providing worse performance than the baseline. In this paper, we introduce a dual classification (DuC) approach, which aims to deal with the potential drawback of over-selection. Specifically, performing instance selection over a given training set, two classifiers are trained using both a ‘good’ and ‘noisy’ sets respectively identified by the instance selection algorithm. Then, a test sample is used to compare the similarities between the data in the good and noisy sets. This comparison guides the input of the test sample to one of the two classifiers. The experiments are conducted using 50 small scale and 4 large scale datasets and the results demonstrate the superior performance of the proposed DuC approach over the baseline instance selection approach.
    出版者: Elsevier B.V
    出版日期: 2015-12-01
    出處: Applied soft computing, 2015-12, Vol.37, p.296-302
    版權: 2015 Elsevier B.V.
    識別號: ISSN: 1568-4946
    識別號: EISSN: 1872-9681
    識別號: DOI: 10.1016/j.asoc.2015.08.038
    顯示於類別:[資訊管理學系] 期刊論文

    文件中的檔案:

    檔案 描述 大小格式瀏覽次數
    index.html0KbHTML29檢視/開啟


    在NCUIR中所有的資料項目都受到原著作權保護.

    社群 sharing

    ::: Copyright National Central University. | 國立中央大學圖書館版權所有 | 收藏本站 | 設為首頁 | 最佳瀏覽畫面: 1024*768 | 建站日期:8-24-2009 :::
    DSpace Software Copyright © 2002-2004  MIT &  Hewlett-Packard  /   Enhanced by   NTU Library IR team Copyright ©   - 隱私權政策聲明