Implementation of the training framework proposed in Self-Rewarding Language Model, from MetaAI
暂无评论,来聊聊你的看法吧