By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
World of SoftwareWorld of Software
  • News
  • Software
  • Mobile
  • Gadget
  • Gaming
  • Videos
Search
  • About
  • Privacy
  • Terms
  • Contact
Copyright © All Rights Reserved. World of Software.
Reading: Google’s new AI turns text into music
Share
Sign In
Notification Show More
Latest News
COOL GADGETS FOR SELF-DEFENSE
March 24, 2023
Nomad’s New MagSafe Charging Stand Is Worth the Extra Cost
March 24, 2023
Verizon Launches Free Videoconferencing Subscription
March 24, 2023
Twitter Blue subscriptions roll out globally, despite missing many promised features
March 24, 2023
COOL SURVIVAL GADGETS OF THE NEW GENERATION
March 24, 2023
Aa
World of SoftwareWorld of Software
Aa
  • News
  • Gadget
  • Software
  • Mobile
  • Computing
  • Gaming
  • Videos
Search
  • Categories
    • News
    • Gadget
    • Software
    • Mobile
    • Computing
    • Gaming
    • Videos
  • Bookmarks
    • Customize Interests
    • My Bookmarks
Have an existing account? Sign In
Follow US
  • About
  • Privacy
  • Terms
  • Contact
Copyright © All Rights Reserved. World of Software.
World of Software > Blog > News > Google’s new AI turns text into music
News

Google’s new AI turns text into music

Press Room
Press Room January 28, 2023
Updated 2023/01/28 at 3:21 PM
Share
SHARE

Google researchers have made an AI that can generate minutes-long musical pieces from text prompts, and can even transform a whistled or hummed melody into other instruments, similar to how systems like DALL-E generate images from written prompts (via TechCrunch). The model is called MusicLM, and while you can’t play around with it for yourself, the company has uploaded a bunch of samples that it produced using the model.

The examples are impressive. There are 30-second snippets of what sound like actual songs created from paragraph-long descriptions that prescribe a genre, vibe, and even specific instruments, as well as five-minute-long pieces generated from one or two words like “melodic techno.” Perhaps my favorite is a demo of “story mode,” where the model is basically given a script to morph between prompts. For example, this prompt:

electronic song played in a videogame (0:00-0:15)

meditation song played next to a river (0:15-0:30)

fire (0:30-0:45)

fireworks (0:45-0:60)

It may not be for everyone, but I could totally see this being composed by a human (I also listened to it on loop dozens of times while writing this article). Also featured on the demo site are examples of what the model produces when asked to generate 10-second clips of instruments like the cello or maracas (the later example is one where the system does a relatively poor job), eight-second clips of a certain genre, music that would fit a prison escape, and even what a beginner piano player would sound like versus an advanced one. It also includes interpretations of phrases like “futuristic club” and “accordion death metal.”

MusicLM can even simulate human vocals, and while it seems to get the tone and overall sound of voices right, there’s a quality to them that’s definitely off. The best way I can describe it is that they sound grainy or staticky. That quality isn’t as clear in the example above, but I think this one illustrates it pretty well:

That, by the way, is the result of asking it to make music that would play at a gym. You may also have noticed that the lyrics are nonsense, but in a way that you may not necessarily catch if you’re not paying attention — kind of like if you were listening to someone singing in Simlish or that one song that’s meant to sound like English but isn’t.

I won’t pretend to know how Google achieved these results, but it’s released a research paper explaining it in detail if you’re the type of person who would understand this figure:

A figure explaining the “hierarchical sequence- to-sequence modeling task” that the researchers use along with AudioLM, another Google project.
Chart: Google

AI-generated music has a long history dating back decades; there are systems that have been credited with composing pop songs, copying Bach better than a human could in the 90s, and accompanying live performances. One recent version uses AI image generation engine StableDiffusion to turn text prompts into spectrograms that are then turned into music. The paper says that MusicLM can outperform other systems in terms of its “quality and adherence to the caption,” as well as the fact that it can take in audio and copy the melody.

That last part is perhaps one of the coolest demos the researchers put out. The site lets you play the input audio, where someone hums or whistles a tune, then lets you hear how the model reproduces it as an electronic synth lead, string quartet, guitar solo, etc. From the examples I listened to, it manages the task very well.

Like with other forays into this type of AI, Google is being significantly more cautious with MusicLM than some of its peers may be with similar tech. “We have no plans to release models at this point,” concludes the paper, citing risks of “potential misappropriation of creative content” (read: plagiarism) and potential cultural appropriation or misrepresentation.

It’s always possible the tech could show up in one of Google’s fun musical experiments at some point, but for now, the only people who will be able to make use of the research are other people building musical AI systems. Google says it’s publicly releasing a dataset with around 5,500 music-text pairs, which could help when training and evaluating other musical AIs.

You Might Also Like

Twitter Blue subscriptions roll out globally, despite missing many promised features

Congress seems more determined to ban TikTok than ever

PayPal’s bringing its passkey logins to Android

Twitter claims ‘legacy’ blue checkmarks will start to disappear on April Fools’ Day

Is Jack Dorsey going to get blown up by Hindenburg Research?

Sign Up For Daily Newsletter

Be keep up! Get the latest breaking news delivered straight to your inbox.
By signing up, you agree to our Terms of Use and acknowledge the data practices in our Privacy Policy. You may unsubscribe at any time.
Press Room January 28, 2023
Share this Article
Facebook TwitterEmail Print
Share
What do you think?
Love0
Sad0
Happy0
Sleepy0
Angry0
Dead0
Wink0
Previous Article Get a Lifetime Subscription to Dollar Flight Club Starting at $40
Next Article The Apple Watch Series 7 with LTE is down to just $329 today
Leave a comment

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Stay Connected

248.1k Like
69.1k Follow
134k Pin
54.3k Follow
banner banner
Join our Newsletter
Get the latest tech news into your inbox
Join Now

Latest News

COOL GADGETS FOR SELF-DEFENSE
Videos
Nomad’s New MagSafe Charging Stand Is Worth the Extra Cost
Mobile
Verizon Launches Free Videoconferencing Subscription
Software
Twitter Blue subscriptions roll out globally, despite missing many promised features
News

You Might also Like

News

Twitter Blue subscriptions roll out globally, despite missing many promised features

3 Min Read
News

Congress seems more determined to ban TikTok than ever

6 Min Read
News

PayPal’s bringing its passkey logins to Android

2 Min Read
News

Twitter claims ‘legacy’ blue checkmarks will start to disappear on April Fools’ Day

4 Min Read
//

We influence 20 million users and is the number one business and technology news network on the planet

Quick Link

  • About
  • Privacy
  • Terms
  • Contact

Support

  • Customize Interests
  • My Bookmarks

Sign Up for Our Newsletter

Subscribe to our newsletter to get our newest articles instantly!

World of SoftwareWorld of Software
Follow US

Copyright © All Rights Reserved. World of Software.

Join Us!

Subscribe to our newsletter and never miss our latest news, podcasts etc..

Zero spam, Unsubscribe at any time.

Removed from reading list

Undo
Welcome Back!

Sign in to your account

Lost your password?