Fine-tuning BERT-based models for Plant Health Bulletin Classification - Archive ouverte HAL Accéder directement au contenu
Communication Dans Un Congrès Année : 2021

Fine-tuning BERT-based models for Plant Health Bulletin Classification

Résumé

In the era of digitization, different actors in agriculture produce numerous data. Such data contains already latent historical knowledge in the domain. This knowledge enables us to precisely study natural hazards within global or local aspects, and then improve the risk prevention tasks and augment the yield, which helps to tackle the challenge of growing population and changing alimentary habits. In particular, French Plants Health Bulletins (BSV, for its name in French Bulletin de Santé du Végétal) give information about the development stages of phytosanitary risks in agricultural production. However, they are written in natural language, thus, machines and human cannot exploit them as efficiently as it could be. Natural language processing (NLP) technologies aim to automatically process and analyze large amounts of natural language data. Since the 2010s, with the increases in computational power and parallelization, representation learning and deep learning methods became widespread in NLP. Recent advancements Bidirectional Encoder Representations from Transformers (BERT) inspire us to rethink of knowledge representation and natural language understanding in plant health management domain. The goal in this work is to propose a BERT-based approach to automatically classify the BSV to make their data easily indexable. We sampled 200 BSV to finetune the pretrained BERT language models and classify them as pest or/and disease and we show preliminary results.
Fichier principal
Vignette du fichier
Fine_tuning_BERT_based_models_for_Plant_Health_Bulletin_Classification.pdf (112.72 Ko) Télécharger le fichier
Origine : Fichiers produits par l'(les) auteur(s)

Dates et versions

hal-03122939 , version 1 (28-01-2021)

Identifiants

  • HAL Id : hal-03122939 , version 1

Citer

Shufan Jiang, Rafael Angarita, Stephane Cormier, Francis Rousseaux. Fine-tuning BERT-based models for Plant Health Bulletin Classification. Technology and Environment Workshop, 2021, Montpellier, France. ⟨hal-03122939⟩

Collections

URCA CRESTIC ISEP
92 Consultations
167 Téléchargements

Partager

Gmail Facebook X LinkedIn More