【Dora’s 学习】用LoRA和p-tuning微调chatglm6b|low rank adaptation |prompt tuning

5999
3
2023-05-27 09:00:00
119
46
403
50
基于清华开源的chatglm6b分别用lora和ptuning两种方式微调,没有使用量化的的情况下,lora需要29G显存,ptuning需要24G显存,最后用微调后的模型做推理需要13G显存(和原chatglm6b一样),供参考~ 参考这位大佬的帖子(感谢分享~):https://github.com/HarderThenHarder/transformers_tasks/tree/main/LLM/finetune
客服
顶部
赛事库 课堂 2021拜年纪