Abstract
Many academic journals and conferences require that each article include a list of keyphrases. These keyphrases should provide general information about the contents and the topics of the article. Keyphrases may save precious time for tasks such as filtering, summarization, and categorization. In this paper, we investigate automatic extraction and learning of keyphrases from scientific articles written in English. Firstly, we introduce various baseline extraction methods. Some of them, formalized by us, are very successful for academic papers. Then, we integrate these methods using different machine learning methods. The best results have been achieved by J48, an improved variant of C4.5. These results are significantly better than those achieved by previous extraction systems, regarded as the state of the art.
| Original language | English |
|---|---|
| Pages (from-to) | 657-669 |
| Number of pages | 13 |
| Journal | Lecture Notes in Computer Science |
| Volume | 3406 3406 LNCS |
| DOIs | |
| State | Published - 2005 |
| Externally published | Yes |
| Event | 6th International Conference on Computational Linguistics and Intelligent Text Processing, CICLing 2005 - Mexico City, Mexico Duration: 13 Feb 2005 → 19 Feb 2005 |
Fingerprint
Dive into the research topics of 'Automatic extraction and learning of keyphrases from scientific articles'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver