In this post, I want to show how I use NLTK for preprocessing and tokenization, but then apply machine learning techniques (e.g. building a linear SVM using stochastic gradient descent) using Scikit-Learn.
F. Karimi, C. Wagner, F. Lemmerich, M. Jadidi, and M. Strohmaier. Proceedings of the 25th International Conference Companion on World Wide Web, page 53--54. Republic and Canton of Geneva, Switzerland, International World Wide Web Conferences Steering Committee, (2016)
X. Zhang, and Y. LeCun. (2015)cite arxiv:1502.01710Comment: This technical report is superseded by a paper entitled "Character-level Convolutional Networks for Text Classification", arXiv:1509.01626. It has considerably more experimental results and a rewritten introduction.