cs.CL 1301.3781

Efficient Estimation of Word Representations in Vector Space

Proposes two efficient models—CBOW and Skip-gram—for large-scale word embedding training, achieving 53.3% accuracy on semantic-syntactic tasks within 3 days on 1.6B words.

Tomas Mikolov, Kai Chen, Greg Corrado et al.

2013-01-17 45