the tech buzz

SUBSCRIBE
AIEnterpriseDealsSecurityCrypto
Newsletter

the tech buzz

Your premier source for technology news, insights, and analysis. Covering the latest in AI, startups, cybersecurity, and innovation.

FOLLOW US

THE DAILY

Get the latest technology updates delivered straight to your inbox.

Company

  • About Us
  • Editorial Team
  • Write For Usnew
  • Contact Us
  • Advertisenew

Legal

  • Privacy Policy
  • Terms of Service
  • Cookie Policy
  • Disclaimer
  • EULA
  • AI Code of Conduct

Resources

  • Newsletters
  • RSS Feeds
  • Subscribe
  • Pricing & Packages
  • Sitemap
  • Archives
  • TechBuzz Pressnew

PUBLISH WITH US

Reach 1.1M+ subscribers via TechBuzz Press.

TechBuzz Press

HAVE A TIP?

Send us a tip using our anonymous form.

Send a tip

HAVE QUESTIONS?

Reach out to us on any subject.

Ask Now

Browse by Category

AIBlockchainCloudSecurityDataDealsInvestmentsEnterpriseVenturesIoTMobileRoboticsSoftwareStartupsAppleMetaMicrosoftOpenAiGoogleTesla

© 2026 The Tech Buzz. All rights reserved.

the tech buzz

NVIDIA Blackwell Ultra Crushes MLPerf Benchmarks

ArticlesNewsletters
ArticlesNewsletters
Launch

NVIDIA Blackwell Ultra Crushes MLPerf Benchmarks

NVIDIA's new Blackwell Ultra chips deliver 1.4x better AI inference performance

by The Tech Buzz

PUBLISHED: Tue, Sep 9, 2025, 3:45 PM UTC | UPDATED: Sun, Aug 30, 2026, 2:39 PM UTC

Add as a preferred source on Google
NVIDIA Blackwell Ultra Crushes MLPerf Benchmarks

NVIDIA's latest Blackwell Ultra architecture just rewrote the AI inference playbook. The GB300 NVL72 systems powered by these new chips crushed every benchmark in MLPerf Inference v5.1, delivering up to 1.4x better performance than the previous generation. This isn't just about bragging rights - it's about making AI factories dramatically more profitable.

NVIDIA just dropped a bombshell that's going to reshape the AI infrastructure landscape. The company's brand-new Blackwell Ultra architecture is absolutely demolishing performance benchmarks, and the numbers are staggering enough to make every data center operator take notice.

Less than six months after debuting at NVIDIA GTC, the GB300 NVL72 rack-scale systems powered by Blackwell Ultra have set new records across every single benchmark in the latest MLPerf Inference v5.1 suite. We're talking about 1.4x better DeepSeek-R1 inference throughput compared to the already impressive Blackwell-based GB200 systems.

But here's what makes this really interesting - this isn't just about raw speed. "Inference performance is critical, as it directly influences the economics of an AI factory," NVIDIA explains in their announcement. The higher the throughput, the more tokens these systems can pump out, which translates directly to increased revenue and lower total cost of ownership.

The technical specs behind these results are genuinely impressive. Blackwell Ultra packs 1.5x more NVFP4 AI compute than its predecessor, along with 2x better attention-layer acceleration and up to 288GB of HBM3e memory per GPU. That's not incremental improvement - that's a generational leap.

Advertisement

NVIDIA swept the board on all the new data center benchmarks added to MLPerf v5.1, including DeepSeek-R1, Llama 3.1 405B Interactive, Llama 3.1 8B, and Whisper. They're also maintaining their stranglehold on per-GPU records across every MLPerf data center benchmark.

The secret sauce here isn't just better silicon. NVIDIA's full-stack approach is paying dividends through hardware acceleration for their proprietary NVFP4 data format - a 4-bit floating point format that delivers better accuracy than other FP4 formats while matching higher-precision alternatives. Their TensorRT Model Optimizer software quantized major models like DeepSeek-R1 and Llama 3.1 405B to this format, squeezing out every bit of performance while meeting strict accuracy requirements.

One particularly clever optimization caught our attention: disaggregated serving. This technique splits large language model inference into separate context and generation tasks, allowing each to be optimized independently. The result? A nearly 50% performance boost per GPU on the Llama 3.1 405B Interactive benchmark when using GB200 NVL72 systems versus traditional serving approaches.

NVIDIA also made their debut with the NVIDIA Dynamo inference framework in these submissions, signaling their continued push into software differentiation alongside hardware dominance.

Advertisement

The ecosystem response has been immediate and overwhelming. Major partners including Azure, Broadcom, Cisco, CoreWeave, Dell Technologies, HPE, Lenovo, Oracle, and Supermicro have already submitted impressive results using NVIDIA's Blackwell and Hopper platforms.

This timing couldn't be better for NVIDIA. As enterprises grapple with the economics of deploying sophisticated AI applications at scale, these performance improvements translate directly to enhanced return on investment. The company is making it clear that their platform isn't just about cutting-edge performance - it's about making AI deployment financially sustainable.

The market implications are significant. Cloud providers can now offer more competitive pricing while maintaining margins, and enterprises can justify larger AI investments with clearer ROI projections. It's exactly the kind of breakthrough that could accelerate enterprise AI adoption across industries that have been sitting on the sidelines.

NVIDIA's Blackwell Ultra launch represents more than just another chip upgrade - it's a fundamental shift in AI infrastructure economics. With 1.4x performance improvements and comprehensive ecosystem support, these systems are poised to accelerate enterprise AI adoption by making deployment both more powerful and more profitable. The real test will be how quickly competitors can respond to this new performance bar.

Advertisement

Advertisement

Trending Now

1

GoPro CEO Vows Cameras Stay Core After Starman Deal

2

Judge Splits Ruling in X vs. Twitter Rival Fight

3

Tim Cook Steps Down, Ternus Takes Apple's Helm

4

Google's Lyria 3.5 Brings AI Music to Gemini

5

Google Translate Gets Listening Mode, Live Background Mode

More in Launch

Korg Kaoss Pad V: First Major Upgrade in 13 Years

Korg Kaoss Pad V: First Major Upgrade in 13 Years

Bluesky Launches Cashtags, LIVE Badges to Capitalize on X Exodus

Bluesky Launches Cashtags, LIVE Badges to Capitalize on X Exodus

Google Ads Campaign Total Budgets Launch

Google Ads Campaign Total Budgets Launch

Google Adds Gemini AI to Trends Explorer Tool

Google Adds Gemini AI to Trends Explorer Tool

Samsung Launches Galaxy Book6 at CES 2026 With Intel Core Ultra

Samsung Launches Galaxy Book6 at CES 2026 With Intel Core Ultra

Nvidia Teaches Autonomous Vehicles to 'Think' with Alpamayo Launch

Nvidia Teaches Autonomous Vehicles to 'Think' with Alpamayo Launch

More Articles

Google Lights Up Vegas with Android XR Sphere Display

Google Lights Up Vegas with Android XR Sphere Display

Jan 5

Samsung Embraces 'AI Living' Vision at CES 2026 Press Conference

Samsung Embraces 'AI Living' Vision at CES 2026 Press Conference

Jan 5

Subtle's AI Earbuds Challenge AirPods With Voice Dictation

Subtle's AI Earbuds Challenge AirPods With Voice Dictation

Jan 4

Samsung Teases 'Your Companion to AI Living' for CES 2026

Samsung Teases 'Your Companion to AI Living' for CES 2026

Dec 23

Google Launches CC, a Gemini-Powered Email Assistant

Google Launches CC, a Gemini-Powered Email Assistant

Dec 16

Gemini Lands in Chrome on iOS, Expanding AI's Mobile Reach

Gemini Lands in Chrome on iOS, Expanding AI's Mobile Reach

Dec 11