FP8
Fp8
SkyRL Adopts FP8 for RL, Cuts Rollout Time by 23%
SkyRL's FP8 reinforcement learning stack matches BF16 convergence while reducing rollout step time by up to 23%, boosting efficiency on NVIDIA GPUs.
Fp8
NVIDIA TensorRT Brings FP8 Quantization to AI Deployment
NVIDIA TensorRT optimizes AI inference with FP8 quantization, offering faster performance and smaller models for scalable deployment.