2026.02.11

Training a Model on Multiple GPUs with Data Parallelism