中大學術數位典藏-NCU Institutional Repository-提供博碩士論文、考古題、期刊論文、研究計畫等下載:Item 987654321/106888
English  |  正體中文  |  简体中文  |  全文笔数/总笔数 : 94274/94274 (100%)
造访人次 : 82911171      在线人数 : 2241
RC Version 7.0 © Powered By DSPACE, MIT. Enhanced by NTU Library IR team.
搜寻范围 查询小技巧:
  • 您可在西文检索词汇前后加上"双引号",以获取较精准的检索结果
  • 若欲以作者姓名搜寻,建议至进阶搜寻限定作者字段,可获得较完整数据
  • 进阶搜寻


    jsp.display-item.identifier=請使用永久網址來引用或連結此文件: https://ir.lib.ncu.edu.tw/handle/987654321/106888


    题名: On learning dual classifiers for better data classification
    作者: 柯士文;Lin, Wei-Chao;Tsai, Chih-Fong;Ke, Shih-Wen;You, Mon-Loon
    贡献者: 管理學院資訊管理學系
    关键词: Data mining;Instance selection;Machine learning;Training effectiveness
    日期: 2015-12-10
    上传时间: 2026-04-23 13:47:53 (UTC+8)
    出版者: Elsevier BV;Elsevier B.V
    摘要: 摘要: •A dual classification (DuC) approach is presented to deal with the potential drawback of instance selection.•During training, two classifiers are trained by a ‘good’ and ‘noisy’ sets respectively after performing instance selection.•Experiments are conducted used 50 small scale and 4 large scale datasets.•The results show that the DuC approach outperforms three state-of-the-art instance selection algorithms. Instance selection aims at filtering out noisy data (or outliers) from a given training set, which not only reduces the need for storage space, but can also ensure that the classifier trained by the reduced set provides similar or better performance than the baseline classifier trained by the original set. However, since there are numerous instance selection algorithms, there is no concrete winner that is the best for various problem domain datasets. In other words, the instance selection performance is algorithm and dataset dependent. One main reason for this is because it is very hard to define what the outliers are over different datasets. It should be noted that, using a specific instance selection algorithm, over-selection may occur by filtering out too many ‘good’ data samples, which leads to the classifier providing worse performance than the baseline. In this paper, we introduce a dual classification (DuC) approach, which aims to deal with the potential drawback of over-selection. Specifically, performing instance selection over a given training set, two classifiers are trained using both a ‘good’ and ‘noisy’ sets respectively identified by the instance selection algorithm. Then, a test sample is used to compare the similarities between the data in the good and noisy sets. This comparison guides the input of the test sample to one of the two classifiers. The experiments are conducted using 50 small scale and 4 large scale datasets and the results demonstrate the superior performance of the proposed DuC approach over the baseline instance selection approach.
    出版者: Elsevier B.V
    出版日期: 2015-12-01
    出處: Applied soft computing, 2015-12, Vol.37, p.296-302
    版權: 2015 Elsevier B.V.
    識別號: ISSN: 1568-4946
    識別號: EISSN: 1872-9681
    識別號: DOI: 10.1016/j.asoc.2015.08.038
    显示于类别:[資訊管理學系] 期刊論文

    文件中的档案:

    档案 描述 大小格式浏览次数
    index.html0KbHTML29检视/开启


    在NCUIR中所有的数据项都受到原著作权保护.

    社群 sharing

    ::: Copyright National Central University. | 國立中央大學圖書館版權所有 | 收藏本站 | 設為首頁 | 最佳瀏覽畫面: 1024*768 | 建站日期:8-24-2009 :::
    DSpace Software Copyright © 2002-2004  MIT &  Hewlett-Packard  /   Enhanced by   NTU Library IR team Copyright ©   - 隱私權政策聲明