Part-of-Speech Tagging for Twitter : Word Clusters and Other Advances

  title={Part-of-Speech Tagging for Twitter : Word Clusters and Other Advances},
  author={Olutobi Owoputi and Brendan F O’CONNOR and Chris Dyer and Kevin Gimpel and Nathan Schneider},
We present improvements to a Twitter part-of-speech tagger, making use of several new features and largescale word clustering. With these changes, the tagging accuracy increased from 89.2% to 92.8% and the tagging speed is 40 times faster. In addition, we expanded our Twitter tokenizer to support a broader range of Unicode characters, emoticons, and URLs. Finally, we annotate and evaluate on a new tweet dataset, DAILYTWEET547, that is more statistically representative of English-language… CONTINUE READING
Highly Cited
This paper has 65 citations. REVIEW CITATIONS

From This Paper

Figures, tables, results, connections, and topics extracted from this paper.
42 Extracted Citations
15 Extracted References
Similar Papers

Citing Papers

Publications influenced by this paper.
Showing 1-10 of 42 extracted citations

66 Citations

Citations per Year
Semantic Scholar estimates that this publication has 66 citations based on the available data.

See our FAQ for additional information.

Similar Papers

Loading similar papers…