Textual analysis of scientific articles published on Colombian fossils

Objective: Identify the lexical proximities in a corpus of texts of scientific articles published in academic journals indexed in the Scopus database on Colombian fossils. Method:  This work applies textual analysis to five paleontological articles on Colombian fossils to identify lexical proximity...

Descripción completa

Detalles Bibliográficos
Autores: Restrepo-Arango, Cristina, Cárdenas-Rozo, Andrés L.
Tipo de recurso: artículo
Estado:Versión publicada
Fecha de publicación:2022
País:Brasil
Institución:Universidade Federal de Santa Catarina (UFSC)
Repositorio:Encontros Bibli
Idioma:portugués
OAI Identifier:oai:periodicos.ufsc.br:article/83470
Acceso en línea:https://periodicos.ufsc.br/index.php/eb/article/view/83470
Access Level:acceso abierto
Palabra clave:Colômbia
Iramuteq
Lexicon
Paleontologia
Colombia
Léxico
Paleontología
Paleontology
Descripción
Sumario:Objective: Identify the lexical proximities in a corpus of texts of scientific articles published in academic journals indexed in the Scopus database on Colombian fossils. Method:  This work applies textual analysis to five paleontological articles on Colombian fossils to identify lexical proximity in a corpus of texts.  This work allowed us to determine:  the grammatical categories, the proximity between categories of words and variables with the analysis of specificities (AE), the grouping of the words with the study of the descending hierarchical classification (CJD) and the graphic presentation of the words. Results: The documentary corpus comprises 31,319- word occurrences, 1,450 active forms or specific words and 303 complimentary forms or common words. The grammatical category of nouns predominates (24%) and words not recognized in the dictionary (17%). The familiar words with the highest frequencies are articles, conjugations, propositions, and pronouns. Conclusions: It was found that there is linguistic proximity between article 1 and the active forms of “Colombia” and article 2 and the active forms of “fossil”. The words were grouped into five classes, and the word cloud was created with 1271 words.