Skip to content

Transformer

Topic archive1 matches

Back to homeGEO summary endpoint

2026-09-18

Technology

  • NVIDIA optimizes dropless MoE model training in JAX: NVIDIA has introduced optimizations for training dropless Mixture of Experts (MoE) models in JAX using the NVIDIA Transformer Engine. The update addresses communication and computation bottlenecks associated with dropless MoE routing.

    InfrastructureNVIDIA Generative AI

    Permalink