nemo automodel speeds moe fine-tuning 3.7x over transformers v5
nvidia nemo automodel builds on huggingface transformers v5 to deliver 3.4-3.7x higher training throughput and up to 32% less gpu memory for mixture-of-experts models with a single import change.