NVIDIAHow-ToAI ModelsSeptember 24, 2026

NVIDIA details MoE training recipe for biological foundation models

Read original source →developer.nvidia.com

NVIDIA's BioNeMo MoE recipe pairs Transformer Engine primitives — GroupedLinear, MXFP8, and a fused GroupedMLP kernel — to cut memory use and improve GPU efficiency in expert-parallel training. The fused MXFP8 GroupedMLP kernel requires NVIDIA Blackwell GPUs, and expert parallelism needs at least two GPUs.

1 source

More stories today

Open the live feed