YuE2: frontier music generation with symbolic planning, zero-shot covers, and agentic music editing.
-
Updated
Sep 29, 2026 - Python
YuE2: frontier music generation with symbolic planning, zero-shot covers, and agentic music editing.
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.
Lab Materials for MIT 6.S191: Introduction to Deep Learning
🎵 The Ultimate Open Source Suno Alternative - Professional UI for ACE-Step 1.5 AI Music Generation. Free, local, unlimited. Stop paying for Suno!
An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.
An AI for Music Generation
SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.
AI Audio Datasets (AI-ADS) 🎵, including Speech, Music, and Sound Effects, which can provide training data for Generative AI, AIGC, AI model training, intelligent audio tool development, and audio applications.
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
MIDI / symbolic music tokenizers for Deep Learning models 🎶
🧠+🎧 Build your music algorithms and AI models with the next-gen DAW 🔥
A unified multimodal language model based on discrete sequence modeling
The most advanced, fully offline client-side AI suite on Android today.
Open‑WebUI Tools is a modular toolkit designed to extend and enrich your Open WebUI instance, turning it into a powerful AI workstation. With a suite of over 15 specialized tools, function pipelines, and filters, this project supports academic research, agentic autonomy, multimodal creativity, workflows, and more
a list of demo websites for automatic music generation research
Apply diffusion models using the new Hugging Face diffusers package to synthesize music instead of images.
Resources on Music Generation with Deep Learning
An all-in-one, 100% local AI video, image, and music studio. Director mode plans full music videos and short films from a single prompt. Built on the WanGP pipeline. Install via Pinokio.
Generate music from the entropy of Linux 🐧🎵
OpenMusic: SOTA Text-to-music (TTM) Generation
To associate your repository with the music-generation topic, visit your repo's landing page and select "manage topics."