Abstract
This paper presents L-KD, a tool that relies on available linguistic and knowledge resources to perform keyphrase clustering and labelling. The aim of L-KD is to help finding and tracing themes in English and Italian text data, represented by groups of keyphrases and associated domains. We perform an evaluation of the top-ranked domains using the 20 Newsgroup dataset, and we show that 8 domains out of 10 match with manually assigned labels. This confirms the good accuracy of this approach, which does not require supervision.
| Original language | English |
|---|---|
| Title of host publication | Proceedings of Third Italian Conference on Computational Linguistics (CLiC-it 2016) & Fifth Evaluation Campaign of Natural Language Processing and Speech Tools for Italian. Final Workshop (EVALITA 2016) |
| Pages | 216-221 |
| Number of pages | 6 |
| Volume | 1749 |
| DOIs | |
| Publication status | Published - 2016 |
| Event | Third Italian Conference on Computational Linguistics (CLiC-it 2016) - Napoli, Italia Duration: 5 Dec 2016 → 7 Dec 2016 |
Conference
| Conference | Third Italian Conference on Computational Linguistics (CLiC-it 2016) |
|---|---|
| City | Napoli, Italia |
| Period | 5/12/16 → 7/12/16 |
Keywords
- computational linguistics, keyphrase extraction, clustering, wordnet domains, conceptnet
Fingerprint
Dive into the research topics of 'KD Strikes Back: from Keyphrases to Labelled Domains Using External Knowledge Sources'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver