We read every piece of feedback, and take your input very seriously.
To see all available qualifiers, see our documentation.
Have a question about this project? Sign up for a free GitHub account to open an issue and contact its maintainers and the community.
By clicking “Sign up for GitHub”, you agree to our terms of service and privacy statement. We’ll occasionally send you account related emails.
Already on GitHub? Sign in to your account
#3510 看下面的图,反向是用的fp16的,计算之后才后cast为32的。 从下面微软介绍deepspeed的视频也可以看到,反向的时候,用的也是16。 https://www.microsoft.com/en-us/research/blog/zero-deepspeed-new-system-optimizations-enable-training-models-with-over-100-billion-parameters/
这里提到过,但是我不知道如何reopen issue,所以新提了一个。
No response
The text was updated successfully, but these errors were encountered:
我今天又看了下,我测试的是14B,如果是72B full的话,即使是80G的显卡也撑不过model.float()
Sorry, something went wrong.
44cfa9a
fixed
No branches or pull requests
Reminder
Reproduction
#3510
看下面的图,反向是用的fp16的,计算之后才后cast为32的。
从下面微软介绍deepspeed的视频也可以看到,反向的时候,用的也是16。
https://www.microsoft.com/en-us/research/blog/zero-deepspeed-new-system-optimizations-enable-training-models-with-over-100-billion-parameters/
这里提到过,但是我不知道如何reopen issue,所以新提了一个。
Expected behavior
No response
System Info
No response
Others
No response
The text was updated successfully, but these errors were encountered: