An Analysis Of Variance Method For Detection Of Collocations In A Pedagogical Domain Corpus

COMPUTACION Y SISTEMAS(2020)

引用 0|浏览6
暂无评分
摘要
In this paper, an exploratory experiment, based on analysis of variance, was carried out in order to get collocations in a pedagogical domain corpus. A semi-automatic corpus containing learning styles papers in Spanish was built. Afterwards, the corpus was lemmatized and a bigrams representation was extracted. The proposed method consists on divide the list of bigrams in quartiles, and analyzing the variance on each one of them. A list of collocations, which was evaluated using a gold standard built by an expert in the domain, was retrieved from each experiment according to established thresholds for the method. Results showed a retrieved list with important collocation in the selected domain.
更多
查看译文
关键词
Pedagogical domain, variance, collocations, ontology, important concepts
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要