ai-forever / ai-forever/ru-prompts

Controlling Text Generation with Sentiment as Attribute

未关闭
#2 2 条评论 0 个 reaction 已指派 1 人 已被 @konodyuk 认领 在 GitHub 查看
主要语言
Python
星标
58
派生
5
PR 合并指标
30 天内没有已合并 PR

描述

Hello! Thank you for such an amazing tool.

I am trying to implement it for my task: i want model to generate texts according to one of three sentiments: positive, neutral, negative. I am not sure it is a text to text task but it is not a simple style transfer either... I tried preprocessing my df with customised preprocessing function:
![2023-01-30 09 56 39](https://user-images.githubusercontent.com/121341170/215408475-263f189f-efdd-4b02-af26-f0c2e92e5f63.jpg)

The traget field is a text which corresponds to a text column in my dataset (there is also a target column in a dataset which is a sentiment). The prompt has a format of "{target}". I used a model ruGPT3large but the results were not very good n the context of language and the attribute usage (however, the text would differ with different sentiments). Important to mention, the loss does not fall, it stays stable all training process:
![2023-01-30 10 04 59](https://user-images.githubusercontent.com/121341170/215410014-9fbffdd8-1f60-4e37-930c-e1d6a89a18f8.jpg)

What can be a problem here?

So, my task of controllable text generation does not probably suit text2text. However, in your notebooks there is no other preprocessing pipeline, only for text2text. I have read the article on Habr and you had there a generation pipeline 'text-generation-with-prompt'. How should I preprocess data then for this task? Or maybe I should use completely different strategy for the task?

贡献指南

这个仓库没有索引到贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。