Crypto Flexs
  • DIRECTORY
  • CRYPTO
    • ETHEREUM
    • BITCOIN
    • ALTCOIN
  • BLOCKCHAIN
  • EXCHANGE
  • ADOPTION
  • TRADING
  • HACKING
  • SLOT
  • TRADE
Crypto Flexs
  • DIRECTORY
  • CRYPTO
    • ETHEREUM
    • BITCOIN
    • ALTCOIN
  • BLOCKCHAIN
  • EXCHANGE
  • ADOPTION
  • TRADING
  • HACKING
  • SLOT
  • TRADE
Crypto Flexs
Home»ADOPTION NEWS»NVIDIA surpasses 1,000 TPS/users with llama 4 Maverick and Blackwell GPUS.
ADOPTION NEWS

NVIDIA surpasses 1,000 TPS/users with llama 4 Maverick and Blackwell GPUS.

By Crypto FlexsMay 23, 20253 Mins Read
Facebook Twitter Pinterest LinkedIn Tumblr Email
NVIDIA surpasses 1,000 TPS/users with llama 4 Maverick and Blackwell GPUS.
Share
Facebook Twitter LinkedIn Pinterest Email

Lawrence Zenga
May 23, 2025 02:10

NVIDIA uses the BLACKWELL GPUS and LLAMA 4 Maverick to achieve the world’s record reasoning speed of 1,000 TPS/users to set new standards for AI model performance.





NVIDIA has set up a new benchmark with AI performance, breaking LLAMA 4 Maverick Model and Blackwell GPU to break 1,000 tokens (TPS) per user barrier. This achievement has been independently verified by artificial analysis of AI benchmarking service, and significant milestones in the speed of LLM (Lange Language Model) reasoning.

Technology development

This breakthrough has been achieved in a single NVIDIA DGX B200 node equipped with eight NVIDIA BLACKWELL GPUs that can handle more than 1,000 tp per user in LLAMA 4 MAVERICK, an 800 million parameter model. Due to this performance, Blackwell is an optimal hardware for deploying LLAMA 4 to maximize throughput or minimize atmospheric time.

Optimization

NVIDIA has completely utilized the Blackwell GPU by using TensOrt-Llm to implement extensive software optimization. The company also trained a speculative decoding draft model using the EAGLE-3 technology, resulting in a four-fold increase compared to the previous baseline. This improvement maintains response accuracy while improving performance and uses the FP8 data type for gemms and professional mixing to ensure the accuracy that can be compared with BF16 metrics.

The importance of low standby time

In the generated AI application, throughput balance and waiting time are important. In the case of important applications that require quick decision -making, NVIDIA’s BLACKWELL GPU is excellent by minimizing the delay time as shown in the TPS/user record. The function of hardware that handles high throughput and low standby time is ideal for various AI tasks.

CUDA kernel and speculation decoding

NVIDIA optimized the CUDA kernel for the work of Gemms, MoE and stocks to maximize performance by using spatial partitioning and efficient memory data rods. Dumping decoding was used to accelerate the speed of LLM reasoning using a smaller and faster draft model proven by smaller Target LLM. This approach increases significant speed, especially when the prediction of the draft model is correct.

Programming method dependency launch

To further improve performance, NVIDIA has reduced GPU idle time between continuous CUDA kernels using PDL (Programmatic Dependent Lunch). This technique allows you to run the kernel to improve the GPU usage rate and remove the performance interval.

The performance of NVIDIA emphasizes leadership in the field of AI infrastructure and data center technology, setting a new standard for the speed and efficiency of the AI ​​model deployment. Innovation of the Blackwell architecture and software optimization continues to react with possible boundaries of AI performance and guarantee real -time user experience and powerful AI applications.

For more information, visit the NVIDIA official blog.

Image Source: Shutter Stock


Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

Related Posts

Algorand (Algo) Get momentum in the launch and technical growth.

July 14, 2025

It flashes again in July

July 6, 2025

Stablecoin startups surpass 2021 venture capital peaks as institutional money spills.

June 28, 2025
Add A Comment

Comments are closed.

Recent Posts

Aster Launches 24/7 Stock Perpetual Contracts Trading With Exposure To U.S. Equities

July 16, 2025

The New Standard For User-centric Token Ecosystems

July 16, 2025

Hete is released as the largest Web3 REWARDS ecosystem in HEDERA.

July 16, 2025

PBK Miner Launches A New Mining Method To Earn Passive Income From XRP, Easily Earning $18,000 A Day

July 15, 2025

Encryption Inheritance: Industrial Round Up -January 20125

July 15, 2025

$TAC Token Debuts In TVL As TAC Mainnet Goes Live With Leading DeFi Protocols

July 15, 2025

MultiBank Group Announces 7 Million $MBG Tokens Sold Out In Under One Hour During Initial Pre-Sale

July 15, 2025

Allnodes Among First To Launch Bare Metal Servers Powered By AMD Threadripper 9000 Series

July 15, 2025

Global Cryptocurrency Investors Flock To DNSBTC After Bitcoin Surges

July 15, 2025

The BTC price is withdrawn at almost $ 123K height. XRP approaches the highest resistance ever at $ 3.00.

July 15, 2025

Easily Invest In DL Mining Cloud Mining And Earn $6,000 In Passive Income Every Day

July 15, 2025

Crypto Flexs is a Professional Cryptocurrency News Platform. Here we will provide you only interesting content, which you will like very much. We’re dedicated to providing you the best of Cryptocurrency. We hope you enjoy our Cryptocurrency News as much as we enjoy offering them to you.

Contact Us : Partner(@)Cryptoflexs.com

Top Insights

Aster Launches 24/7 Stock Perpetual Contracts Trading With Exposure To U.S. Equities

July 16, 2025

The New Standard For User-centric Token Ecosystems

July 16, 2025

Hete is released as the largest Web3 REWARDS ecosystem in HEDERA.

July 16, 2025
Most Popular

Bitfarms Expands Board of Directors and Appoints Andrew J. Chang as Independent Director

November 24, 2024

Franklin Templeton Renews Support for Solana, Praises Technology and Adoption

July 25, 2024

‘Final Fantasy’ creator Square Enix moves to Arbitrum for Ethereum game NFT

May 30, 2024
  • Home
  • About Us
  • Contact Us
  • Disclaimer
  • Privacy Policy
  • Terms and Conditions
© 2025 Crypto Flexs

Type above and press Enter to search. Press Esc to cancel.