György Orosz


An efficient language independent toolkit for complete morphological disambiguation
László Laki | György Orosz
Proceedings of the Ninth International Conference on Language Resources and Evaluation (LREC'14)

In this paper a Moses SMT toolkit-based language-independent complete morphological annotation tool is presented called HuLaPos2. Our system performs PoS tagging and lemmatization simultaneously. Amongst others, the algorithm used is able to handle phrases instead of unigrams, and can perform the tagging in a not strictly left-to-right order. With utilizing these gains, our system outperforms the HMM-based ones. In order to handle the unknown words, a suffix-tree based guesser was integrated into HuLaPos2. To demonstrate the performance of our system it was compared with several systems in different languages and PoS tag sets. In general, it can be concluded that the quality of HuLaPos2 is comparable with the state-of-the-art systems, and in the case of PoS tagging it outperformed many available systems.


Morphological annotation of Old and Middle Hungarian corpora
Attila Novák | György Orosz | Nóra Wenszky
Proceedings of the 7th Workshop on Language Technology for Cultural Heritage, Social Sciences, and Humanities

PurePos 2.0: a hybrid tool for morphological disambiguation
György Orosz | Attila Novák
Proceedings of the International Conference Recent Advances in Natural Language Processing RANLP 2013