SciELO - Scientific Electronic Library Online

 
vol.26 número3Distributional Word Vectors as Semantic Maps FrameworkMachine Translation for Low-Resource English-Mizo Pair Encountering Tonal Words índice de autoresíndice de assuntospesquisa de artigos
Home Pagelista alfabética de periódicos  

Serviços Personalizados

Journal

Artigo

Indicadores

Links relacionados

  • Não possue artigos similaresSimilares em SciELO

Compartilhar


Computación y Sistemas

versão On-line ISSN 2007-9737versão impressa ISSN 1405-5546

Resumo

YASHOTHARA, S.; UTHAYASANKER, R. T.  e  DIAS, G. V.. Semi-Automatic Alignment of Multilingual Parts of Speech Tagsets. Comp. y Sist. [online]. 2022, vol.26, n.3, pp.1365-1375.  Epub 02-Dez-2022. ISSN 2007-9737.  https://doi.org/10.13053/cys-26-3-4357.

We cast the problem of mapping a pair of Parts of Speech (POS) tagsets as a labelled tree mapping problem and present a general-purpose semi-automatic POS tree alignment algorithm to solve the alignment. This algorithm can be used to align two POS tagsets of different languages or the same language. We evaluate its usefulness using POS tagsets of two languages: Tamil and Sinhala. The proposed approach shows that manual effort in prior approaches is drastically reduced due to the proposed algorithm and eliminates the need to create new POS tagsets.

Palavras-chave : Parts of speech; POS tagset mapping; POS tagset alignment; semi-automatic approach; BIS tagset; UOM tagset; Tamil NLP; Sinhala NLP.

        · texto em Inglês