Two New Challenging Resources to Evaluate Natural Language Interfaces to Databases Generated Based on Geobase and Geoquery

Two New Challenging Resources to Evaluate Natural Language Interfaces to Databases Generated Based on Geobase and Geoquery

Juan Javier González-Barbosa, Juan Frausto Solís, Juan Paulo Sánchez-Hernández, Julia Patricia Sanchez-Solís
ISBN13: 9781799847304|ISBN10: 1799847306|EISBN13: 9781799847311
DOI: 10.4018/978-1-7998-4730-4.ch004
Cite Chapter Cite Chapter

MLA

González-Barbosa, Juan Javier, et al. "Two New Challenging Resources to Evaluate Natural Language Interfaces to Databases Generated Based on Geobase and Geoquery." Handbook of Research on Natural Language Processing and Smart Service Systems, edited by Rodolfo Abraham Pazos-Rangel, et al., IGI Global, 2021, pp. 70-100. https://doi.org/10.4018/978-1-7998-4730-4.ch004

APA

González-Barbosa, J. J., Frausto Solís, J., Sánchez-Hernández, J. P., & Sanchez-Solís, J. P. (2021). Two New Challenging Resources to Evaluate Natural Language Interfaces to Databases Generated Based on Geobase and Geoquery. In R. Pazos-Rangel, R. Florencia-Juarez, M. Paredes-Valverde, & G. Rivera (Eds.), Handbook of Research on Natural Language Processing and Smart Service Systems (pp. 70-100). IGI Global. https://doi.org/10.4018/978-1-7998-4730-4.ch004

Chicago

González-Barbosa, Juan Javier, et al. "Two New Challenging Resources to Evaluate Natural Language Interfaces to Databases Generated Based on Geobase and Geoquery." In Handbook of Research on Natural Language Processing and Smart Service Systems, edited by Rodolfo Abraham Pazos-Rangel, et al., 70-100. Hershey, PA: IGI Global, 2021. https://doi.org/10.4018/978-1-7998-4730-4.ch004

Export Reference

Mendeley
Favorite

Abstract

Databases and corpora are essential resources to evaluate the performance of Natural Language Interfaces to Databases (NLIDB). The Geobase database and the Geoquery corpus (Geoquery250 and Geoquery880) are among the most commonly used. In this chapter, the authors analyze both resources to offer two elaborate resources: 1) N-Geobase, which is a relational database, and 2) the corpus Geoquery270. The former follows the standard normalization procedure, then N-Geobase has a schema similar to enterprise databases. Geoquery270 consists of 270 queries selected from Geoquery880, preserving the same kind of natural language problems as Geoquery880, but with more challenging issues for an NLIDB than Geoquery250. To evaluate the new resources, they compared the performance of the NLIDB using Geoquery270 and Geoquery250. The results indicated that Geoquery270 was the harder corpus, while Geoquery250 is the easier one. Consequently, this chapter offers a broader range of resources to NLIDB designers.

Request Access

You do not own this content. Please login to recommend this title to your institution's librarian or purchase it from the IGI Global bookstore.