Towards a Reference Corpus of Web Genres for the Evaluation of Genre Identification Systems (CROSBI ID 732238)
Prilog sa skupa u zborniku | izvorni znanstveni rad | međunarodna recenzija
Podaci o odgovornosti
Rehm, Georg ; Santini, Marina ; Mehler, Alexander ; Braslavski, Pavel ; Gleim, Rüdiger ; Stubbe, Andrea ; Symonenko, Svetlana ; Tavosanis, Mirko ; Vidulin, Vedrana
engleski
Towards a Reference Corpus of Web Genres for the Evaluation of Genre Identification Systems
We present initial results from an international and multi-disciplinary research collaboration that aims at the construction of a reference corpus of web genres. The primary application scenario for which we plan to build this resource is the automatic identification of web genres. Web genres are rather difficult to capture and to describe in their entirety, but we plan for the finished reference corpus to contain multi-level tags of the respective genre or genres a web document or a website instantiates. As the construction of such a corpus is by no means a trivial task, we discuss several alternatives that are, for the time being, mostly based on existing collections. Furthermore, we discuss a shared set of genre categories and a multi-purpose tool as two additional prerequisites for a reference corpus of web genres.
web genres, reference corpus
nije evidentirano
nije evidentirano
nije evidentirano
nije evidentirano
nije evidentirano
nije evidentirano
Podaci o prilogu
351-358.
2008.
objavljeno
Podaci o matičnoj publikaciji
Proceedings of the Sixth International Conference on Language Resources and Evaluation (LREC'08)
Calzolari, Nicoletta ; Choukri, Khalid ; Maegaard, Bente ; Mariani, Joseph ; Odijk, Jan ; Piperidis, Stelios ; Tapias, Daniel
2-9517408-4-0
Podaci o skupu
International Conference on Language Resources and Evaluation
predavanje
28.05.2008-30.05.2008
Marakeš, Maroko