利用多張GPU訓練大型語言模型—從零開始介紹DeepSpeed、Liger Kernel、Flash Attention及Quantization —— 【生成式AI時代下的機器學習(2025)】助教課

【生成式AI時代下的機器學習(2025)】助教課:利用多張GPU訓練大型語言模型—從零開始介紹DeepSpeed、Liger Kernel、Flash Attention及Quantization




image




地址:


https://www.youtube.com/watch?v=mpuRca2UZtI



















posted on 2026-04-21 18:12  Angry_Panda  阅读(11)  评论(0)    收藏  举报

导航