< 返回工具列表
A

Alpaca-LoRA-RLHF-PyTorch

> AI 编程
开源

A full pipeline to finetune Alpaca LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of th

60 stars0 点赞0 次浏览
访问官网GitHub

工具介绍

A full pipeline to finetune Alpaca LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the Alpaca architecture. Basically ChatGPT but with Alpaca

> 标签

Pythonalpacachatgptdeepspeedfinetune

暂无评论,来聊聊你的看法吧

> 工具信息

发布日期2026年8月1日
最后更新2026年8月1日
分类AI 编程
定价开源