A token centric part-of-speech tagger for biomedical text

Conclusion Our analysis of tagger performance suggests that lexical differences between corpora have more effect on tagging accuracy than originally considered by previous research work. Biomedical POS tagging algorithms may be modified to improve their cross-domain tagging accuracy without requiring extra training or large training data sets. Future work should reexamine POS tagging methods for biomedical text. This differs from the work to date that has focused on retraining existing POS taggers.
Source: Artificial Intelligence in Medicine - Category: Bioinformatics Source Type: research