AbrahamSanders / AbrahamSanders/seq2seq-chatbot

embeddings

未关闭
#27 1 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看
主要语言
Python
星标
99
派生
54
PR 合并指标
30 天内没有已合并 PR

描述

hi Abraham! i ran your algorithm with some series subtitles (in italian), but the result is quite akward... surely due to the small dimension of my dataset (60k lines)... i was wondering, maybe i could try helping the training process with embeddings...
i found some embeddings file in italian... but most of it are not in the form you required (TF checkpoints)
they come with .m extensions or .npy ect... nothing seems to fit the one your algorithm can process
do you think is possibile, in a few lines (i don't want to excessively bother you) to explain to me how to create a brand new embedding checkpoint (from scratch or converting one already built), or tell me where i can check in github projects?
thank you!!
Matteo

贡献指南

这个仓库没有索引到贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。