YuE2,港科大、纽约大学、斯坦福等几家联合出的开源音乐模型,给它歌词和风格描述,先写出一份旋律和和弦的谱,再按谱生成带人声伴奏的整首歌。
This is a AI post classified by Jev as Open & local models (a tool drop), kept by the AI Radar because it carries real work, not commentary.
YuE2,港科大、纽约大学、斯坦福等几家联合出的开源音乐模型,给它歌词和风格描述,先写出一份旋律和和弦的谱,再按谱生成带人声伴奏的整首歌。 中间那份谱是能看能改的,改一段和声、换个速度、动几个音,再交回去生成新的一版,整首歌怎么走自己说了算。 GitHub:http://github.com/multimodal-art-projection/YuE 翻唱也能做,把一段现成录音转成旋律谱,配新歌词或者换个风格,同一个模型直接出一版新的演绎。 官方演示里一首《The Last Train》通过对话改了 9 步 14 个版本,从中文流行一路改成英文爵士还加了段萨克斯独奏,每一版的谱和对话都能翻。 自家评测里跟 Suno v5、v6 打得有来有回,README 也承认头几名差距很小分不出高下。 配套给了一个 Agent Skill,让 Claude Code 这类工具直接调它写歌改歌。 需要 Linux 加 24 GB 显存的 N 卡,出的是 48 kHz 立体声。个人和音乐人自己用是免费的,生成的作品拿去变现也不用交授权费。
Posted by GitHubDaily (84.6k followers) 1 days ago · 63 likes · 5.3k views · view the original post on X. Kept by the AI Radar as Open & local models. Tools mentioned: yue.
More AI work like this
- dear rtx 3060 owners, and every 12gb card behind it. bonsai 2 27b dense went from 26 to… — @sudoingX
- Nimble is able to process images! — @madiator
- Woah...it claims to be open Jev. And it's 100% opensource. — @Saboo_Shubham_
- your rtx 3060 was running bonsai 2 at 26 tok/s this morning and now does 40 tok/s, i… — @sudoingX
- Meet DeepSeek-V4.1-Flash: a blazing fast image-text-to-text model. It takes both images… — @HuggingModels
- Meet Edge0-35B-A3B-preview: a 35B parameter MoE model built for edge inference. It runs… — @HuggingModels
- Putting Jeff-1 out into the world, this is an attempt at a Jev like model, built on qwen… — @DJLougen
- Train a 9M parameter language model from scratch in five minutes. — @tom_doerr
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 27.2k posts from 4.8k X accounts over the last 14 days, 1.2k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:45 UTC. Full method.