discoDSP has released two free apps: WavePrompt, a local text-to-music generator for macOS and Windows, and Roboto, a singing robot voice synthesizer for macOS, Windows, Linux and iOS.
WavePrompt
WavePrompt is a free app that generates loops, vocal phrases and full songs from a text prompt. It runs the YuE2 music model on your own computer, on the Apple Silicon GPU on a Mac or an NVIDIA or Intel Arc GPU on Windows, with no server, no account and no upload.
Describe a sound and WavePrompt renders it locally. Loops are cut to exact bars at your tempo, vocal phrases sing your own words in nine languages, and songs run from intro to outro. A piano roll lets you draw or generate the melody and chords the result follows, and the latest update 1.0.6 adds Import Memo, which turns a hummed or sung recording such as an iPhone Voice Memo into a melody of up to 16 bars with matching tempo and key.
Features
- Bar-accurate loops from 1 to 16 bars at 60 to 180 BPM, in any major or minor key and in 4/4, 3/4, 6/8, 5/4 or 7/8.
- Sample-accurate loop starts with nudging, so loops drop straight into a DAW grid.
- Sung vocal phrases from your lyrics in English, Chinese, Japanese, Korean, Spanish, French, German, Italian or Portuguese.
- Eleven voice types including breathy, powerful, deep, raspy, falsetto, whisper, choir and rap, with optional vocal isolation via Demucs.
- Full songs from 15 seconds to 4 minutes, with vocals or instrumental, ending on a planned outro. Lyrics are checked with Whisper.
- Piano roll to draw a melody and pick a chord for each bar, or generate them in the chosen key with density, register and chord options.
- Import Memo turns a hummed or sung recording into a piano roll melody.
- 34 genres, 14 moods and 11 instrument sets, plus a style strength setting.
- Fixed seeds to repeat or vary a result, and a library that keeps every render with its recipe.
- Every result comes with a MIDI file of its score, with melody, chord, tempo, key and section marker tracks.
- Export to WAV (48 kHz, 24-bit), AAC or MIDI, with tempo matching by the Rubber Band Library.
- Works offline after the first model download.
System requirements
- Mac: Apple Silicon (M1 or later), macOS 26 Tahoe or later, 16 GB of unified memory (32 GB recommended for full songs).
- Windows: Windows 10 (version 2004) or Windows 11, 64-bit, NVIDIA GeForce RTX 30 series or newer with 12 GB of video memory or an Intel Arc GPU, 16 GB of RAM (32 GB recommended for full songs).
- About 10 GB of free disk space for the models and the engine.
Music is generated with YuE2 by Multimodal Art Projection. Model weights are licensed under CC BY-NC 4.0 with individual creator permission, and generated audio is yours to use under that permission.
More information: discoDSP / WavePrompt
Roboto
Roboto is a free singing voice synthesizer. Write a melody in the piano roll, type a word on each note, and Roboto sings it. It runs as a plugin or standalone app on macOS, Windows, Linux, iPhone and iPad.
The voice is synthesized with no samples, from a glottal pulse source, five formant resonators and shaped noise for consonants. Roboto sings in English, German, French, Spanish, Japanese, Russian, Finnish and Swedish, and the latest update 1.0.2 adds video export of the animated robot, waveform and lyrics with the song.
Features
- Piano roll where every note carries a word, with snap, zoom and a marker to sing from.
- Text tab that holds the whole song as text, so songs can be typed or pasted.
- Musicalize button that gives plain lyric lines a pentatonic melody, one note per syllable.
- Automation curves along the song for volume, pitch bend, vibrato depth, breath, formant shift and chorus.
- Words are turned into phonemes by built-in rules for each language. Phonemes can also be typed directly, for example S-T-R-IY-T.
- Ten voice presets: Default, Robot, Soprano, Bass, Child, Whisper, Choir, Opera, Giant and Alien.
- Transpose, tempo with host sync, vibrato, jitter, breath, consonant level and length, formant shift and bandwidth, reverb and double track.
- Animated robot face with LED eyes and a mouth that follow the voice, plus a live waveform and lyrics that light up in time.
- MIDI keys from C1 sing from the matching bar while held, and MIDI controllers drive volume, vibrato, breath, formant shift and chorus.
- Import MIDI, karaoke MIDI (.kar), UltraStar, LRC lyrics, SRT subtitles and song text, from the menu or by dropping the file on the window.
- Export to WAV, MIDI, karaoke MIDI, UltraStar, LRC and SRT, with lyric times that match the WAV.
- Video export (MP4, 1280×720 at 60 fps) on macOS, Windows and iOS.
- Built-in songs in every language.
- All voice parameters can be automated by the host.
Formats and system requirements
- macOS 12 or later (Apple Silicon and Intel): VST3, AU and standalone.
- Windows and Linux (64-bit): VST3 and standalone.
- iOS 15 or later (iPhone and iPad): AUv3 and standalone.
More information: discoDSP / Roboto
Read More