Toward Network-based Keyword Extraction from Multitopic Web Documents (CROSBI ID 617203)
Prilog sa skupa u zborniku | izvorni znanstveni rad | međunarodna recenzija
Podaci o odgovornosti
Šišović, Sabina ; Martinčić-Ipšić, Sanda ; Meštrović, Ana
engleski
Toward Network-based Keyword Extraction from Multitopic Web Documents
In this paper we analyse the selectivity measure calculated from the complex network in the task of the automatic keyword extraction. Texts, collected from different web sources (portals, forums), are represented as directed and weighted co-occurrence complex networks of words. Words are nodes and links are established between two nodes if they are directly co-occurring within the sentence. We test different centrality measures for ranking nodes – keyword candidates. The promising results are achieved using the selectivity measure. Then we propose an approach which enables extracting word pairs according to the values of the in/out selectivity and weight measures combined with filtering.
keyword extraction; complex networks; co-occurrence language networks; Croatian texts; selectivity
nije evidentirano
nije evidentirano
nije evidentirano
nije evidentirano
nije evidentirano
nije evidentirano
Podaci o prilogu
18-27.
2014.
objavljeno
Podaci o matičnoj publikaciji
Proceedings of the 6th International Conference on Information Technologies and Information Society (ITIS 2014)
Levnajić, Zoran ; Boshkoska, Biljana Mileva
Novo Mesto: Fakulteta za informacijske študije v Novem mestu
978-961-93391-3-8
Podaci o skupu
International Conference on Information Technologies and Information Society (ITIS2014)
predavanje
05.11.2014-07.11.2014
Šmarješke toplice, Slovenija