MLCommons Releases MLPerf Training 6.0 Benchmarks

MLPerf Training 6.0 introduces DeepSeek-V3 and GPT-OSS benchmarks, highlighting performance gains for NVIDIA and AMD hardware systems.

MLCommons has launched MLPerf Training 6.0, featuring new benchmarks for the DeepSeek-V3 671B and GPT-OSS 20B Mixture-of-Experts models. DeepSeek-V3 671B utilizes 671 billion total parameters, while GPT-OSS 20B uses 21 billion. NVIDIA systems demonstrated significant scalability, utilizing NVLink to connect 72 GPUs per rack and scale-out networks for up to 8,192 GPUs. A cluster of 8,192 NVIDIA GB300 GPUs finished the DeepSeek-V3 training in 2,021 seconds, while 8,192 GB200 GPUs took 3,340 seconds. In Llama 2 70B tests, eight NVIDIA GB300 GPUs finished in 5,613 seconds, outperforming eight AMD MI355X GPUs at 8,271 seconds. AMD systems, limited to eight GPUs via Infinity Fabric, saw the MI355X complete Llama 3.1 8B training in 91,145 seconds. Results also detailed precision formats including FP4, NVFP4, and MXFP4.


Tensordyne Tape-Outs 138B Transistor AI Chip

Chinese Brands Eye Custom LLW DRAM for Phones

Intel Arctic Sound 2T GPU Sample Surfaces Online

Qualcomm Eyes $10B Acquisition of Tenstorrent

Kioxia Debuts Exceria G3 PCIe 5.0 SSD Series

Thermal Grizzly Debuts DeltaMate Mycro Pro II

AMD Acquires AI Startup MEXT for Server Memory

AMD to Relaunch Picasso APUs for Mobile in 2026

AMD Launches $3,999 Ryzen AI Halo Mini PC

MSI Unveils MPG 271KRAW18 5K Mini LED Monitor

Intel to Launch CPUs with NVIDIA Graphics by 2028

MSI Claw EX AI+ Handheld Debuts at $1,699

NVIDIA RTX PRO 6000 Blackwell Hits $13,250

Intel Raptor Lake Next Coming to LGA1700 in 2027

NVIDIA Blackwell GB300 Sets Agentic AI Records

Razer Debuts Seiren V3 Pro Dynamic Microphone

Samsung Denies SSD Warranty to Louis Rossmann

iFixit Teardown Reveals Trump Mobile T1 Origins

Airra to Launch 59g Rotary Mouse on Kickstarter

NVIDIA Adds Day-1 Support for DiffusionGemma AI