Javascript must be enabled to continue!
Automatic Indexing for Agriculture: Designing a Framework by Deploying Agrovoc, Agris and Annif
View through CrossRef
There are several ways to employ machine learning for automating subject indexing. One popular strategy is to utilize a supervised learning algorithm to train a model on a set of documents that have been manually indexed by subject matter using a standard vocabulary. The resulting model can then predict the subject of new and previously unseen documents by identifying patterns learned from the training data. To do this, the first step is to gather a large dataset of documents and manually assign each document a set of subject keywords/descriptors from a controlled vocabulary (e.g., from Agrovoc). Next, the dataset (obtained from Agris) can be divided into – i) a training dataset, and ii) a test dataset. The training dataset is used to train the model, while the test dataset is used to evaluate the model's performance. Machine learning can be a powerful tool for automating the process of subject indexing. This research is an attempt to apply Annif (http://annif. org/), an open-source AI/ML framework, to autogenerate subject keywords/descriptors for documentary resources in the domain of agriculture. The training dataset is obtained from Agris, which applies the Agrovoc thesaurus as a vocabulary tool (https://www.fao.org/agris/download).
Sarada Ranganathan Endowment for Library Science
Title: Automatic Indexing for Agriculture: Designing a Framework by Deploying Agrovoc, Agris and Annif
Description:
There are several ways to employ machine learning for automating subject indexing.
One popular strategy is to utilize a supervised learning algorithm to train a model on a set of documents that have been manually indexed by subject matter using a standard vocabulary.
The resulting model can then predict the subject of new and previously unseen documents by identifying patterns learned from the training data.
To do this, the first step is to gather a large dataset of documents and manually assign each document a set of subject keywords/descriptors from a controlled vocabulary (e.
g.
, from Agrovoc).
Next, the dataset (obtained from Agris) can be divided into – i) a training dataset, and ii) a test dataset.
The training dataset is used to train the model, while the test dataset is used to evaluate the model's performance.
Machine learning can be a powerful tool for automating the process of subject indexing.
This research is an attempt to apply Annif (http://annif.
org/), an open-source AI/ML framework, to autogenerate subject keywords/descriptors for documentary resources in the domain of agriculture.
The training dataset is obtained from Agris, which applies the Agrovoc thesaurus as a vocabulary tool (https://www.
fao.
org/agris/download).
Related Results
A Review on Indexing Techniques and its application in Multilingual Information Retrieval System
A Review on Indexing Techniques and its application in Multilingual Information Retrieval System
To implement the indexing in multilingual dataset, the indexing process must know. This paper gives the brief about indexing and presents role of indexing, logical view of indexing...
Avaliação do sistema de indexação automática Annif em artigos de periódicos em Ciência da Informação
Avaliação do sistema de indexação automática Annif em artigos de periódicos em Ciência da Informação
Objetivo: Avaliar o desempenho da ferramenta de indexação automática Annif na atribuição de descritores a artigos científicos na área de Ciência da Informação.
Método: Adota uma ab...
Applying (semi-)automatic metadata to early modern normative texts. Annif and
Policeygesetzgebung
from the City-State of Bern (1528–1798)
Applying (semi-)automatic metadata to early modern normative texts. Annif and
Policeygesetzgebung
from the City-State of Bern (1528–1798)
Abstract
This study investigates the application of modern digital tools to analyze handwritten normative texts from the City-State of Bern (1528-1798). By levera...
LLM-Generated RDF Triples for Agricultural Species: A Comparative Evaluation Using AGROVOC Grounding
LLM-Generated RDF Triples for Agricultural Species: A Comparative Evaluation Using AGROVOC Grounding
The construction of RDF knowledge bases for specialized domains is costly and requires close collaboration between domain experts and knowledge engineers. This paper evaluates thre...
Non-Recommended Publishing Lists: Strategies for Detecting Deceitful Journals
Non-Recommended Publishing Lists: Strategies for Detecting Deceitful Journals
Abstract
The rapid growth of open access publishing (OAP) has significantly improved the accessibility and dissemination of scientific knowledge. However, this expansion has also c...
Archives of Pediatric Neurosurgery is now indexed on Scopus !
Archives of Pediatric Neurosurgery is now indexed on Scopus !
I have great news to share! The Archives of Pediatric Neurosurgery is now indexed on Scopus. This is a significant achievement that will enhance the visibility and accessibility of...
Indexing policies for knowledge organization and representation
Indexing policies for knowledge organization and representation
The aim of this research was to investigate the elements of the indexing policy used by professional cataloguers in a university library system in the Amazon region. The methodolog...
What is NMC Indexed or NMC Approved Journals?
What is NMC Indexed or NMC Approved Journals?
The National Medical Commission (NMC) released guidelines in 2021 regarding faculty eligibility and journal indexing for research publications in medical institutions. To address ...

