Nvidia’s new AI model Fugatto can generate music and audio

Nvidia on Monday showed a new artificial intelligence model for generating music and audio that can modify voices and generate novel sounds — technology aimed at the producers of music, films and video games.

Nvidia, the world’s biggest supplier of chips and software used to create AI systems, said it does not have immediate plans to publicly release the technology, which it calls Fugatto, short for Foundational Generative Audio Transformer Opus 1.

It joins other technologies shown by startups such as Runway and larger players such as Meta Platforms that can generate audio or video from a text prompt.

Santa Clara, California-based Nvidia’s version generates sound effects and music from a text description, including novel sounds such as making a trumpet bark like a dog.

What makes it different from other AI technologies is its ability to take in and modify existing audio, for example by taking a line played on a piano and transforming it into a line sung by a human voice, or by taking a spoken word recording and changing the accent used and the mood expressed.

“If we think about synthetic audio over the past 50 years, music sounds different now because of computers, because of synthesizers,” said Bryan Catanzaro, vice president of applied deep learning research at Nvidia. “I think that generative AI is going to bring new capabilities to music, to video games and to ordinary folks that want to create things.”

While companies such as OpenAI are negotiating with Hollywood studios over whether and how the AI could be used in the entertainment industry, the relationship between tech and Hollywood has become tense, particularly after Hollywood star Scarlett Johansson accused OpenAI of imitating her voice.

Nvidia’s new model was trained on open-source data, and the company said it is still debating whether and how to release it publicly.

“Any generative technology always carries some risks, because people might use that to generate things that we would prefer they don’t,” Catanzaro said. “We need to be careful about that, which is why we don’t have immediate plans to release this.”

Creators of generative AI models have yet to determine how to prevent abuse of the technology such as a user generating misinformation or infringing on copyrights by generating copyrighted characters.

OpenAI and Meta similarly have not said when they plan to release to the public their models that generate audio or video.

—Stephen Nellis, Reuters

https://www.fastcompany.com/91235722/nvidia-ai-model-generates-music-audio-sounds?partner=rss&utm_source=rss&utm_medium=feed&utm_campaign=rss+fastcompany&utm_content=rss

Creado 9mo | 25 nov 2024, 18:10:04


Inicia sesión para agregar comentarios

Otros mensajes en este grupo.

Netflix is doubling down on full-season drops with season two of Meghan’s show

Meghan, Duchess of Sussex’ latest season of her reality show, With Love, Meghan, drops today on Netflix. In line with the stream

26 ago 2025, 14:40:16 | Fast company - tech
Listen to the 10 most memorable sound effects in the history of tech

For understandable reasons, most technology coverage tends to focus more on the physical or visual

26 ago 2025, 14:40:15 | Fast company - tech
Where solar investments pack the biggest climate punch

The United States’ hourly demand for electricity broke two records last month, reaching its highest-ever level—759,190 megawatts

26 ago 2025, 14:40:14 | Fast company - tech
Doctors love this AI app because it gives them hours of their lives back

A typical physician’s job is much more than just seeing patients. In fact, most doctors spend hours every week outside of clinic hours catching up on typing notes and getting visits and trea

26 ago 2025, 14:40:12 | Fast company - tech
Agentic AI has companies excited and security experts freaked out

Agentic AI is being heralded as the future of the generative AI revolu

26 ago 2025, 12:30:04 | Fast company - tech
This man keeps buying and returning 110-pound anvils on Amazon

An Illinois man keeps buying and returning 110-pound anvils on Amazon—until “someone does something about it,” he says.

The creator, who goes by Johnbo Stockwell on

26 ago 2025, 5:30:09 | Fast company - tech
3 quick and easy ways to clear up storage space in Windows 11

Digital hoarders, unite! I have a game on my PC that I haven’t played in months, and it’s taking up more than 100 GB of disk space. There, I said it.

This is a scenario most of us find o

26 ago 2025, 5:30:07 | Fast company - tech