Deep learning for NLP resources
github.com
github.com
http://www.cs.berkeley.edu/~istoica/classes/cs294/15/class.h...
I think it makes more sense when CNNs are applied at the character level. The filter banks then activate for specific n-gram patterns of characters, like certain prefixes, suffixes, and root words. The higher level LSTMs are then relieved of having to understand that level of structure. Also, tokenization is hard, and might be especially wrong for media with grammatical abuse like Twitter, and this avoids that janky preprocessing. See: http://arxiv.org/abs/1508.06615
even in images, you're able to zero pad the input.