Semiconductor News & Analysis Feed

2 articles
2026-07-11
developer.nvidia.com 2026-07-11
This article explores how host offloading techniques can alleviate high-bandwidth memory (HBM) bottlenecks in large language model (LLM) training using the JAX framework. As model size, sequence length, and batch size increase, GPU memory becomes a critical constraint. The study highlights NVIDIA's
2026-07-07
developer.nvidia.com 2026-07-07
NVIDIA's technical blog highlights a novel approach to enhancing Goodput in large-scale LLM training through Nonuniform Tensor Parallelism (NTP). As AI model training increasingly relies on thousands of GPUs, interruptions and resource fluctuations pose significant challenges. NTP addresses these by