Genre as noise: noise in genre

被引:7
|
作者
Stubbe, Andrea [2 ]
Ringlstetter, Christoph [1 ]
Schulz, Klaus U. [2 ]
机构
[1] Univ Alberta, Dept Comp Sci, AICML, Edmonton, AB T6G 2E8, Canada
[2] Univ Munich, CIS, D-80538 Munich, Germany
关键词
genre hierarchies; features; genre classification; error dictionaries; noisy corpora;
D O I
10.1007/s10032-007-0060-2
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Given a specific information need, documents of the wrong genre can be considered as noise. From this perspective, genre classification helps to separate relevant documents from noise. Orthographic errors represent a second, finer notion of noise. Since specific genres often include documents with many errors, an interesting question is whether this "micro-noise" can help to classify genre. In this paper we consider both problems. After introducing a comprehensive hierarchy of genres, we present an intuitive method to build specialized and distinctive classifiers that also work for very small training corpora. Special emphasis is given to the selection of intelligent high-level features. We then investigate the correlation between genre and micro noise. Using special error dictionaries, we estimate the typical error rates for each genre. Finally, we test if the error rate of a document represents a useful feature for genre classification.
引用
收藏
页码:199 / 209
页数:11
相关论文
共 50 条
  • [1] Genre as noise: noise in genre
    Andrea Stubbe
    Christoph Ringlstetter
    Klaus U. Schulz
    [J]. International Journal of Document Analysis and Recognition (IJDAR), 2007, 10 : 199 - 209
  • [2] The Impact of Noise in Web Genre Identification
    Pritsos, Dimitrios
    Stamatatos, Efstathios
    [J]. EXPERIMENTAL IR MEETS MULTILINGUALITY, MULTIMODALITY, AND INTERACTION, 2015, 9283 : 268 - 273
  • [3] Fan Discourse and the Construction of Noise Music as a Genre
    Atton, Chris
    [J]. JOURNAL OF POPULAR MUSIC STUDIES, 2011, 23 (03) : 324 - 342
  • [4] Genre-based Decomposition of Email Class Noise
    Kolcz, Aleksander
    Cormack, Gordon V.
    [J]. KDD-09: 15TH ACM SIGKDD CONFERENCE ON KNOWLEDGE DISCOVERY AND DATA MINING, 2009, : 427 - 435
  • [5] 1/f Noise analysis of songs in various genre of music
    Ro, Wosuk
    Kwon, Younghun
    [J]. CHAOS SOLITONS & FRACTALS, 2009, 42 (04) : 2305 - 2311
  • [6] Television Genre/Musical Genre / Expressive Genre
    Rodman, Ron
    [J]. AMERICAN MUSIC, 2019, 37 (04) : 435 - 457
  • [7] Genre as an absence of genre
    Strajn, Jelka Kernev
    [J]. PRIMERJALNA KNJIZEVNOST, 2006, 29 : 265 - 274
  • [8] PARVENUE GENRE, GENRE OF PARVENUS
    FLETCHER, J
    [J]. COMPARATIVE LITERATURE STUDIES, 1978, 15 (02) : 193 - 202
  • [9] Genre Theory and Genre History
    Heinz, Sarah
    Horlacher, Stefan
    [J]. GERMANISCH-ROMANISCHE MONATSSCHRIFT, 2009, 59 (03): : 456 - 458
  • [10] Massive-Scale Genre Communities Learning Using a Noise-Tolerant Deep Architecture
    Zhang, Luming
    Ju, Xiaoming
    Yao, Yiyang
    Liu, Zhenguang
    [J]. IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 22 (09) : 2467 - 2478