利用gensim 直接生成文档向量

    
    def gen_d2v_corpus(self, lines):

        with open("./data/ques2_result.txt", "wb") as fw:
            for line in lines:
                fw.write(" ".join(jieba.lcut(line)) + "\n")

        sents = doc2vec.TaggedLineDocument("./data/ques2_result.txt")
        model = doc2vec.Doc2Vec(sents, size = 50, window = 5, alpha = 0.015)
        model.train(sents)

        corpus = model.docvecs
        np.save("./output/d2v.corpus.npy", corpus)

        return np.asarray(corpus)

 

相关文章:

  • 2021-11-10
  • 2021-10-31
  • 2021-11-01
  • 2022-02-14
  • 2022-01-28
  • 2021-10-08
  • 2021-11-17
  • 2021-11-04
猜你喜欢
  • 2022-12-23
  • 2022-12-23
  • 2018-04-14
  • 2022-12-23
  • 2022-12-23
  • 2022-12-23
  • 2021-12-06
相关资源
相似解决方案