讨论:2015-08-17

来自cslt Wiki
2015年8月19日 (三) 04:52Luoling讨论 | 贡献的版本

(差异) ←上一版本 | 最后版本 (差异) | 下一版本→ (差异)
跳转至: 导航搜索

Works in This week:

1.Finish training word embeddings via 5 models : using EnWiki dataset(953M): CBOW,Skip-Gram(SG) using text8 dataset(95.3M): CBOW,Skip-Gram(SG),C&W,GloVe,LBL and Order(count-based)

2.Use tasks to measure quality of the word vectors with various dimensions: word similarity(ws) the TOEFL set analogy task text classification named entity recognition(ner) sentence-level sentiment classification (based on convolutional neural networks),just call it 'cnn' part-of-speech tagging(pos)