Comments (10)
遇到了同样的问题,解决了么?大佬
from zero_nlp.
注意使用我发布的最新版本的代码!!!!!!!!!
data2
数据已经是老版本的数据了。03-28日最新代码
from zero_nlp.
大概是用的旧的代码,之前也有人碰到过这个问题,是由于数据集被过滤后,变成空的了。用新的code02_训练模型全部流程.ipynb 不会碰到这个问题。
from zero_nlp.
我用新的 code02_训练模型全部流程.ipynb 可以跑通
from zero_nlp.
你们训练后有效果吗?我问模型属性相关的问题,还是之前chatglm模型的回答,没有纠正过来。
from zero_nlp.
参考作者的data2数据集里,问:你是谁?,我是良XXXX程序员训练的一个AI模型。。
看它用了多少数据量。 讲一次两次。它听不进去。。思维太固执了,据作者说要1600次。我估计 得准备几十上百轮吧。
from zero_nlp.
from zero_nlp.
from zero_nlp.
是的 发自我的 iPhone 在 2023年3月31日,11:45,chenyiwan @.> 写道: 参考作者的data2数据集里,问:你是信,我是良XXXX程序员训练的一个AI模型。。 看它用了多少数据量。 讲一次两次。它听不进去。。思维太固执了,据作者说要1600次。我估计 得准备几十上百轮吧。 — Reply to this email directly, view it on GitHub<#38 (comment)>, or unsubscribehttps://github.com/notifications/unsubscribe-auth/AHJRI6PESFC3GNR55RDV3ETW6ZHLVANCNFSM6AAAAAAWMXAT24. You are receiving this because you commented.Message ID: @.>
这个1600次是啥意思,就是说要准备1600次问你是谁,然后回答是xxx,就可以转变过来了?
from zero_nlp.
我一行行的运行的,到最后就报错了RuntimeError: self and mat2 must have the same dtype。
运行下面的代码:
20 trainer = MyTrainer(
21 model=model,
22 tokenizer=tokenizer,
(...)
26 eval_dataset=tokenized_datasets["valid"],
27 )
---> 28 trainer.train()的时候
from zero_nlp.
Related Issues (20)
- 求助:chatglm2 lora训练error:RuntimeError: Expected is_sm80 to be true, but got false. HOT 2
- 训练的时候报错ValueError: The current `device_map` had weights offloaded to the disk. HOT 11
- 训练出错
- 两张4090单机多卡跑,咋感觉越跑越慢了,比单卡慢 HOT 2
- 请问有部署或者运行的文档吗?在哪里可以看?
- 实时微调可以通过加入传统RL实现吗
- 请问如果单纯使用zeroth-order向前优化少量batch(只要体现出一定的优化效果)的话要怎么实现 HOT 2
- lora推理中只能指定一个输入吗?有办法实现batch_size的推理吗
- 救命!!ChatGlm-v2-6b_Lora该怎么设置epoch?? HOT 1
- 大佬,可以多个多个lora叠加使用吗?
- chatglm_v2_6b_lora多卡如何设置,没有找到 HOT 2
- 能出一个ChatGLM
- 能出一个ChatGLM的教程吗
- Segment Fault 是哪的问题?
- 大佬 chinese_llama 还可以用吗 HOT 1
- 出个chatglm3的吧 微调后 推理老是出问题 HOT 1
- internlm-sft 单机多卡微调 GPU 利用率低 HOT 5
- 大佬出个教程把 HOT 2
- 4/8bit量化的问题及源码阅读的问题 HOT 1
- 关于流水线并行的一个问题
Recommend Projects
-
React
A declarative, efficient, and flexible JavaScript library for building user interfaces.
-
Vue.js
🖖 Vue.js is a progressive, incrementally-adoptable JavaScript framework for building UI on the web.
-
Typescript
TypeScript is a superset of JavaScript that compiles to clean JavaScript output.
-
TensorFlow
An Open Source Machine Learning Framework for Everyone
-
Django
The Web framework for perfectionists with deadlines.
-
Laravel
A PHP framework for web artisans
-
D3
Bring data to life with SVG, Canvas and HTML. 📊📈🎉
-
Recommend Topics
-
javascript
JavaScript (JS) is a lightweight interpreted programming language with first-class functions.
-
web
Some thing interesting about web. New door for the world.
-
server
A server is a program made to process requests and deliver data to clients.
-
Machine learning
Machine learning is a way of modeling and interpreting data that allows a piece of software to respond intelligently.
-
Visualization
Some thing interesting about visualization, use data art
-
Game
Some thing interesting about game, make everyone happy.
Recommend Org
-
Facebook
We are working to build community through open source technology. NB: members must have two-factor auth.
-
Microsoft
Open source projects and samples from Microsoft.
-
Google
Google ❤️ Open Source for everyone.
-
Alibaba
Alibaba Open Source for everyone
-
D3
Data-Driven Documents codes.
-
Tencent
China tencent open source team.
from zero_nlp.