Publications des agents du Cirad

Cirad

Automatic Identification of Research Fields in Scientific Papers

Kergosien E., Farvardin M.A., Teisseire M., Bessagnet M.N., Schöpfel J., Chaudiron S., Jacquemin B., Lacayrelle A., Roche M., Sallaberry C., Tonneau J.P.. 2018. In : Calzolari Nicoletta (ed.), Choukri Khalid (ed.), Cieri Christopher (ed.), Declerck Thierry (ed.), Goggi Sara (ed.), Hasida Koiti (ed.), Isahara Hitoshi (ed.), Maegaard Bente (ed.), Mariani Joseph (ed.), Mazo Hélène (ed.), Moreno Asuncion (ed.), Odijk Jan. Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018). Miyazaki : ELRA, p. 1902-1907. International Conference on Language Resources and Evaluation. 11, 2018-05-07/2018-05-12, Miyazaki (Japon).

The TERRE-ISTEX project aims to identify scientific research dealing with specific geographical territories areas based on heterogeneous digital content available in scientific papers. The project is divided into three main work packages: (1) identification of the periods and places of empirical studies, and which reflect the publications resulting from the analyzed text samples, (2) identification of the themes which appear in these documents, and (3) development of a web-based geographical information retrieval tool (GIR). The first two actions combine Natural Language Processing patterns with text mining methods. The integration of the spatial, thematic and temporal dimensions in a GIR contributes to a better understanding of what kind of research has been carried out, of its topics and its geographical and historical coverage. Another originality of the TERRE-ISTEX project is the heterogeneous character of the corpus, including PhD theses and scientific articles from the ISTEX digital libraries and the CIRAD research center.

Documents associés

Communication de congrès

Agents Cirad, auteurs de cette publication :