Create music, interact with AI, and generate realistic images with their powerful tools.
YuE

About YuE
YuE is an open-source foundation model series focused on music generation, specifically converting lyrics into complete songs. It produces full-length tracks that include both vocal and instrumental components, enabling users to create music from text prompts. The model supports multiple languages, genres, and vocal styles, making it adaptable for diverse musical needs. YuE also offers an inference mode that allows users to generate music in a style similar to a reference song, providing creative flexibility. It includes specialized upsampling models to enhance audio quality, ensuring professional-grade output. Users can explore demos showcasing its vocal performance capabilities and stay informed about ongoing improvements and updates. The tool is designed for AI-powered songwriting and is particularly useful for musicians, producers, and content creators seeking automated music generation solutions.
Key features
- Converts lyrics into full-length songs with vocals and instruments
- Supports multiple languages and genres
- Style matching via reference song inference
- Specialized upsampling models for improved audio quality
- Open-source foundation model series
- Demos available for vocal performance evaluation
- Ongoing model enhancements and updates
- GitHub repository for community access and contributions
Use cases
- Generating complete songs from provided lyrics
- Creating music in specific styles by referencing existing tracks
- Experimenting with multilingual vocal compositions
Pros
- Generates full-length songs (up to five minutes) with coherent musical structure and lyrical alignment
- Supports multiple languages, genres, and vocal styles, including niche styles like Beijing Opera
- Enables style transfer and bidirectional generation through redesigned in-context learning techniques
- Includes specialized upsampling models to enhance audio quality for professional-grade output
- Demonstrates strong performance on music understanding tasks, matching or exceeding state-of-the-art methods
Cons
- Requires technical expertise to fine-tune for additional controls or tail languages
- Long-form generation may demand significant computational resources
- Current examples rely on GPT-generated lyrics, which may limit creative originality
Frequently asked questions about YuE
What is YuE and what does it do?
YuE is an open-source foundation model series designed for full-song generation, particularly converting lyrics into complete tracks with vocals and instruments. It supports multiple languages and genres, enabling users to create music from text prompts or reference styles.
Who is YuE designed for?
YuE is designed for musicians, producers, and content creators seeking automated music generation solutions. It is also suitable for researchers and developers interested in exploring music generation and understanding tasks.
Does YuE support style transfer or reference-based generation?
Yes, YuE includes redesigned in-context learning techniques that enable versatile style transfer, such as converting a song from one genre or language to another while preserving the original accompaniment.
Can YuE generate music in languages other than English?
Yes, YuE supports multiple languages, including Chinese, Japanese, and Korean, and fine-tuning can enhance support for additional tail languages.
What are the technical requirements for using YuE?
YuE requires significant computational resources for long-form music generation and fine-tuning. Users need familiarity with model deployment and may need to adjust settings for optimal performance.
How can I get started with YuE?
Users can explore YuE through its model checkpoints, examples, and documentation available on GitHub and Hugging Face. The tool provides links to different language-specific models and an upsampler for enhancing audio quality.