2014
DOI: 10.1007/978-3-642-54903-8_35
|View full text |Cite
|
Sign up to set email alerts
|

Named Entities as New Features for Czech Document Classification

Abstract: Abstract. This paper is focused on automatic document classification. The results will be used to develop a real application for the Czech News Agency. The main goal of this work is to propose new features based on the Named Entities (NEs) for this task. Five different approaches to employ NEs are suggested and evaluated on a Czech newspaper corpus. We show that these features do not improve significantly the score over the baseline word-based features. The classification error rate improvement is only about 0… Show more

Help me understand this report

Search citation statements

Order By: Relevance

Paper Sections

Select...
1

Citation Types

0
1
0

Year Published

2015
2015
2020
2020

Publication Types

Select...
3

Relationship

1
2

Authors

Journals

citations
Cited by 3 publications
(1 citation statement)
references
References 27 publications
0
1
0
Order By: Relevance
“…In [9], three different multi-label classification approaches are compared and evaluated. Other recent works propose novel features based on named entities [14] or on unsupervised machine learning [3].…”
Section: Related Workmentioning
confidence: 99%
“…In [9], three different multi-label classification approaches are compared and evaluated. Other recent works propose novel features based on named entities [14] or on unsupervised machine learning [3].…”
Section: Related Workmentioning
confidence: 99%