← 开源
NVIDIA
TransformerEngine
A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.
Model DevelopmentDeep Learning FrameworksPython
在 GitHub 打开 增长势头
+024 小时新增 Star0.0%
3.57k
Star
849
Fork
+11
本周
100
贡献者
创建于 2022-09-20 · 更新于 2026-10-06 · 今日第 7802 名
主要开发者