NVIDIA/TransformerEngine

在 GitHub 查看

A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.

星标
3.5k
派生
831
开放的新手 issue
2
已索引 issue
163
平均合并
3 天 11 小时
30 天内合并 PR
65
主要语言
Python
许可证
Apache-2.0
最近 GitHub push
2026年9月19日
最近索引
2026年9月19日
贡献指南
贡献指南
行为准则
没有行为准则
新手标签
没有已索引的新手标签
目前已索引 163 个未关闭的 Issue 正在加载 Issue

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。