AI PILLED

A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.

Read
#1288
0%Merged PRs by AI agents

0 / 272 merged PRs · 90 days

AI agents
0
User accounts
272
Other bots
0

Agents

No AI-agent authors in this read.