Abstract
We present a system that allows a user to search a large linguistically annotated corpus using syntactic patterns over dependency graphs. In contrast to previous attempts to this effect, we introduce a light-weight query language that does not require the user to know the details of the underlying syntactic representations, and instead to query the corpus by providing an example sentence coupled with simple markup. Search is performed at an interactive speed due to an efficient linguistic graph-indexing and retrieval engine. This allows for rapid exploration, development and refinement of syntax-based queries. We demonstrate the system using queries over two corpora: the English wikipedia, and a collection of English pubmed abstracts. A demo of the wikipedia system is avilable at: https://allenai.github.io/spike/ .
Original language | English |
---|---|
Title of host publication | ACL 2020 - 58th Annual Meeting of the Association for Computational Linguistics, Proceedings of the System Demonstrations |
Publisher | Association for Computational Linguistics (ACL) |
Pages | 17-23 |
Number of pages | 7 |
ISBN (Electronic) | 9781952148040 |
State | Published - 2020 |
Event | 58th Annual Meeting of the Association for Computational Linguistics, ACL 2020 - Virtual, Online, United States Duration: 5 Jul 2020 → 10 Jul 2020 |
Publication series
Name | Proceedings of the Annual Meeting of the Association for Computational Linguistics |
---|---|
ISSN (Print) | 0736-587X |
Conference
Conference | 58th Annual Meeting of the Association for Computational Linguistics, ACL 2020 |
---|---|
Country/Territory | United States |
City | Virtual, Online |
Period | 5/07/20 → 10/07/20 |
Bibliographical note
Publisher Copyright:© 2020 Association for Computational Linguistics
Funding
This project has received funding from the Eu-ropoean Research Council (ERC) under the Eu-ropoean Union’s Horizon 2020 research and innovation programme, grant agreement No. 802774 (iEXTRACT). This project has received funding from the Europoean Research Council (ERC) under the Europoean Union?s Horizon 2020 research and innovation programme, grant agreement No. 802774 (iEXTRACT).
Funders | Funder number |
---|---|
Eu-ropoean Research Council | |
Europoean Union?s Horizon 2020 research and innovation programme | |
Horizon 2020 Framework Programme | |
European Commission | |
Horizon 2020 | 802774 |