pith. sign in

arxiv: cmp-lg/9503020 · v1 · submitted 1995-03-20 · cmp-lg · cs.CL

Different Issues in the Design of a Lemmatizer/Tagger for Basque

classification cmp-lg cs.CL
keywords lemmatizertaggerbasquedesignissuesmorphologicaltagsetwill
0
0 comments X
read the original abstract

This paper presents relevant issues that have been considered in the design of a general purpose lemmatizer/tagger for Basque (EUSLEM). The lemmatizer/tagger is conceived as a basic tool necessary for other linguistic applications. It uses the lexical data base and the morphological analyzer previously developed and implemented. Due to the characteristics of the language, the tagset here proposed in structured in for levels, so that each level is a refinement of the previous one in the sense that it adds more detailed information. We will focus on the problems found in designing this tagset and on the strategies for morphological disambiguation that will be used.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.