arXiv · 1611.03057
When silver glitters more than gold: Bootstrapping an Italian part-of-speech tagger for Twitter
Abstract
We bootstrap a state-of-the-art part-of-speech tagger to tag Italian Twitter data, in the context of the Evalita 2016 PoSTWITA shared task. We show that training the tagger on native Twitter data enriched with little amounts of specifically selected gold data and additional silver-labelled data scraped from Facebook, yields better results than using large amounts of manually annotated data from a mix of genres.
Explore related subjects
Keep this discovery
Barbara Plank, Malvina Nissim. 2016-11-09. When silver glitters more than gold: Bootstrapping an Italian part-of-speech tagger for Twitter. https://arxiv.org/abs/1611.03057
Cite the original work for its findings. Save a collection to share your selection of sources.