N-gram embeddings (e.g., bigrams) seem to improve performance on many tasks compared to simple word embeddings. Again, fastText could be a role model?
N-gram embeddings (e.g., bigrams) seem to improve performance on many tasks compared to simple word embeddings. Again, fastText could be a role model?