A Corpus of Quotes
NLTK Source
When bored, code.
Tabular ML
Yet Another Neural Machine Translation Toolkit
The dataset contains 3 million attribute-value annotations across 1257 unique categories on 2.2 million cleaned Amazon product profiles. It is a large, multi-sourced, diverse dataset for product attribute extraction study.
The dataset contains 3 million attribute-value annotations across 1257 unique categories on 2.2 million cleaned Amazon product profiles. It is a la...
"Ale is my bear necessity."
Resources and tools for Indian language Natural Language Processing
Japanese Address Munger
A little word cloud generator in Python
Significant Machine Translation news/gossips...
WMT data in Python
Hack and Tell @ Saarland University
Unsupervised Neural Machine Translation
USAAR participation in SemEval2015
A re-implementation of redpony/cdec's tokenize-anything.pl script in python