Services on Demand
Journal
Article
Indicators
Cited by SciELO
Access statistics
Related links
Similars in SciELO
Share
Computación y Sistemas
On-line version ISSN 2007-9737Print version ISSN 1405-5546
Abstract
ALEMAN, Yuridiana; SOMODEVILLA, María and VILARINO, Darnes. An Analysis of Variance Method for Detection of Collocations in a Pedagogical Domain Corpus. Comp. y Sist. [online]. 2020, vol.24, n.2, pp.739-743. Epub Oct 04, 2021. ISSN 2007-9737. https://doi.org/10.13053/cys-24-2-3411.
In this paper, an exploratory experiment, based on analysis of variance, was carried out in order to get collocations in a pedagogical domain corpus. A semi-automatic corpus containing learning styles papers in Spanish was built. Afterwards, the corpus was lemmatized and a bigrams representation was extracted. The proposed method consists on divide the list of bigrams in quartiles, and analyzing the variance on each one of them. A list of collocations, which was evaluated using a gold standard built by an expert in the domain, was retrieved from each experiment according to established thresholds for the method. Results showed a retrieved list with important collocation in the selected domain.
Keywords : Pedagogical domain; variance; collocations; ontology; important concepts.