中大學術數位典藏-NCU Institutional Repository-提供博碩士論文、考古題、期刊論文、研究計畫等下載:Item 987654321/106888
English  |  正體中文  |  简体中文  |  Items with full text/Total items : 94274/94274 (100%)
Visitors : 82915322      Online Users : 2297
RC Version 7.0 © Powered By DSPACE, MIT. Enhanced by NTU Library IR team.
Scope Tips:
  • please add "double quotation mark" for query phrases to get precise results
  • please goto advance search for comprehansive author search
  • Adv. Search
    HomeLoginUploadHelpAboutAdminister Goto mobile version


    Please use this identifier to cite or link to this item: https://ir.lib.ncu.edu.tw/handle/987654321/106888


    Title: On learning dual classifiers for better data classification
    Authors: 柯士文;Lin, Wei-Chao;Tsai, Chih-Fong;Ke, Shih-Wen;You, Mon-Loon
    Contributors: 管理學院資訊管理學系
    Keywords: Data mining;Instance selection;Machine learning;Training effectiveness
    Date: 2015-12-10
    Issue Date: 2026-04-23 13:47:53 (UTC+8)
    Publisher: Elsevier BV;Elsevier B.V
    Abstract: 摘要: •A dual classification (DuC) approach is presented to deal with the potential drawback of instance selection.•During training, two classifiers are trained by a ‘good’ and ‘noisy’ sets respectively after performing instance selection.•Experiments are conducted used 50 small scale and 4 large scale datasets.•The results show that the DuC approach outperforms three state-of-the-art instance selection algorithms. Instance selection aims at filtering out noisy data (or outliers) from a given training set, which not only reduces the need for storage space, but can also ensure that the classifier trained by the reduced set provides similar or better performance than the baseline classifier trained by the original set. However, since there are numerous instance selection algorithms, there is no concrete winner that is the best for various problem domain datasets. In other words, the instance selection performance is algorithm and dataset dependent. One main reason for this is because it is very hard to define what the outliers are over different datasets. It should be noted that, using a specific instance selection algorithm, over-selection may occur by filtering out too many ‘good’ data samples, which leads to the classifier providing worse performance than the baseline. In this paper, we introduce a dual classification (DuC) approach, which aims to deal with the potential drawback of over-selection. Specifically, performing instance selection over a given training set, two classifiers are trained using both a ‘good’ and ‘noisy’ sets respectively identified by the instance selection algorithm. Then, a test sample is used to compare the similarities between the data in the good and noisy sets. This comparison guides the input of the test sample to one of the two classifiers. The experiments are conducted using 50 small scale and 4 large scale datasets and the results demonstrate the superior performance of the proposed DuC approach over the baseline instance selection approach.
    出版者: Elsevier B.V
    出版日期: 2015-12-01
    出處: Applied soft computing, 2015-12, Vol.37, p.296-302
    版權: 2015 Elsevier B.V.
    識別號: ISSN: 1568-4946
    識別號: EISSN: 1872-9681
    識別號: DOI: 10.1016/j.asoc.2015.08.038
    Appears in Collections:[Department of Information Management] journal & Dissertation

    Files in This Item:

    File Description SizeFormat
    index.html0KbHTML30View/Open


    All items in NCUIR are protected by copyright, with all rights reserved.

    社群 sharing

    ::: Copyright National Central University. | 國立中央大學圖書館版權所有 | 收藏本站 | 設為首頁 | 最佳瀏覽畫面: 1024*768 | 建站日期:8-24-2009 :::
    DSpace Software Copyright © 2002-2004  MIT &  Hewlett-Packard  /   Enhanced by   NTU Library IR team Copyright ©   - 隱私權政策聲明