Crypto Flexs
  • DIRECTORY
  • CRYPTO
    • ETHEREUM
    • BITCOIN
    • ALTCOIN
  • BLOCKCHAIN
  • EXCHANGE
  • ADOPTION
  • TRADING
  • HACKING
  • SLOT
Crypto Flexs
  • DIRECTORY
  • CRYPTO
    • ETHEREUM
    • BITCOIN
    • ALTCOIN
  • BLOCKCHAIN
  • EXCHANGE
  • ADOPTION
  • TRADING
  • HACKING
  • SLOT
Crypto Flexs
Home»ADOPTION NEWS»NVIDIA Unveils BigVGAN v2: Pioneering Zero-Shot Waveform Audio Generation
ADOPTION NEWS

NVIDIA Unveils BigVGAN v2: Pioneering Zero-Shot Waveform Audio Generation

By Crypto FlexsSeptember 11, 20243 Mins Read
Facebook Twitter Pinterest LinkedIn Tumblr Email
NVIDIA Unveils BigVGAN v2: Pioneering Zero-Shot Waveform Audio Generation
Share
Facebook Twitter LinkedIn Pinterest Email

Jack Anderson
September 6, 2024 11:03

NVIDIA’s BigVGAN v2 sets a new standard for zero-shot waveform audio generation, delivering state-of-the-art quality at up to 3x faster synthesis speeds.





NVIDIA has announced the release of BigVGAN v2, a groundbreaking generative AI model for zero-shot waveform audio generation, according to the NVIDIA Technical Blog. The new model represents a significant improvement in speed and quality, establishing it as the state-of-the-art solution in audio generation AI.

BigVGAN: A universal neural vocoder

BigVGAN is a general-purpose neural vocoder designed to synthesize audio waveforms from Mel spectrograms. The model uses a fully synthetic architecture with multiple upsampling blocks and residual augmented synthesis layers. The main feature is an anti-aliasing multi-periodical composition (AMP) module that is optimized to generate high-frequency and periodic sound waves, reducing artifacts in the process.

Improvements in BigVGAN v2

BigVGAN v2 introduces several improvements over its predecessors.

  • Cutting edge audio quality Across a variety of measurement criteria and audio types.
  • Up to 3x faster synthesis speed Through optimized CUDA kernels.
  • Pre-trained checkpoints For a variety of audio configurations.
  • Supports sampling rates up to 44kHzContains the highest frequencies that humans can hear.

Generate all the sounds in the world

Waveform audio generation is essential to virtual worlds and has been a major focus of research. BigVGAN v2 overcomes previous limitations by providing high-quality audio with improved fine details. Trained using NVIDIA A100 Tensor Core GPUs and a dataset 100x larger than its predecessor, BigVGAN v2 can generate high-quality sound waves in a variety of domains, including speech, environmental sounds, and music.

Reaching the highest frequency sound that the human ear can detect

Previous models were limited to sampling rates between 22kHz and 24kHz. BigVGAN v2 extends this range to 44kHz, capturing the entire human auditory spectrum. This allows the model to reproduce a wide range of soundscapes, from powerful drums to the crisp cymbals of music.

Faster synthesis using custom CUDA kernels

BigVGAN v2 also provides accelerated synthesis speeds, achieving up to 3x faster inference than the original BigVGAN using custom CUDA kernels. These kernels enable audio waveform generation up to 240x faster than real-time on a single NVIDIA A100 GPU.

Audio quality results

BigVGAN v2 demonstrates superior audio quality for speech and general audio compared to previous models, and achieves similar results to Descript Audio Codec at 44kHz sampling rate. This demonstrates the model’s ability to generate high-quality waveforms for a wide range of audio types.

conclusion

NVIDIA’s BigVGAN v2 sets a new standard for audio synthesis, achieving state-of-the-art quality across all audio types and covering the full range of human hearing. The model’s synthesis speed is now up to 3x faster, making it highly efficient for a wide range of audio configurations.

For more details, see the BigVGAN v2 model card on GitHub.

Image source: Shutterstock


Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

Related Posts

Arthur Breitman talks about the strategic evolution of Tezos in the Coinshares interview.

May 9, 2025

COREWEAVE completes the AI ​​developer platform weight and bias acquisition.

May 9, 2025

HKMA reports stable credit conditions for SMEs in the first quarter of 2025.

May 9, 2025
Add A Comment

Comments are closed.

Recent Posts

Arthur Breitman talks about the strategic evolution of Tezos in the Coinshares interview.

May 9, 2025

COREWEAVE completes the AI ​​developer platform weight and bias acquisition.

May 9, 2025

Ether Lee’s Staying Surges: Is PECTRA attracting more than retail investors?

May 9, 2025

The new blockchain T-Rex raises $ 17 million in Web3 to convert the Layer Layer.

May 9, 2025

HKMA reports stable credit conditions for SMEs in the first quarter of 2025.

May 9, 2025

SEC’s CRENSHAW Slams Ripple Settlement, ‘Regulatory Vacuum’ Warning

May 9, 2025

Tether launches USD ES in KAIA blockchain to promote Web3 adoption in Asia.

May 9, 2025

Easy to get Daily Crypto -Bow Miner’s AI Cloud Mining can benefit while sleeping!

May 9, 2025

Bitcoin hit $ 101K to reclaim six pictures as Trump confirmed us. British trade transaction

May 9, 2025

Bitcoin’s APRIL SURGE sets a promising summer stage.

May 8, 2025

Bitcoin Options BTC’s potential to emphasize the new all -time high

May 8, 2025

Crypto Flexs is a Professional Cryptocurrency News Platform. Here we will provide you only interesting content, which you will like very much. We’re dedicated to providing you the best of Cryptocurrency. We hope you enjoy our Cryptocurrency News as much as we enjoy offering them to you.

Contact Us : Partner(@)Cryptoflexs.com

Top Insights

Arthur Breitman talks about the strategic evolution of Tezos in the Coinshares interview.

May 9, 2025

COREWEAVE completes the AI ​​developer platform weight and bias acquisition.

May 9, 2025

Ether Lee’s Staying Surges: Is PECTRA attracting more than retail investors?

May 9, 2025
Most Popular

Crypto whale lost $36 million to major phishing scam triggering DETH depeg.

October 11, 2024

Tim Draper Leads $3.5 Million Raising for Bitcoin Liquidity Protocol Zest

May 13, 2024

Cryptocurrency Storage in Hong Kong: Compliance Changes Explained

February 21, 2024
  • Home
  • About Us
  • Contact Us
  • Disclaimer
  • Privacy Policy
  • Terms and Conditions
© 2025 Crypto Flexs

Type above and press Enter to search. Press Esc to cancel.