Enable SheetSage2 audio remix in YuE2 UI

This commit is contained in:
Mikei386
2026-09-10 22:56:24 +02:00
parent af25425aee
commit 5c34afa7fa
3 changed files with 121 additions and 1 deletions
+23
View File
@@ -44,6 +44,29 @@ and stable-diffusion.cpp are not part of the Athena setup. Manual composition,
generation, result playback, score editing and the take library work without
those optional components.
### Uploaded-audio remix (SheetSage2)
The **Remix a take** drawer also accepts WAV, FLAC, MP3, M4A, OGG, Opus and
AAC uploads. SheetSage2 transcribes the recording into an editable ABC melody
and chord plan, then YuE2 renders that structure in a newly selected style.
It does not preserve the original samples, singer or production verbatim.
SheetSage2 is kept in `/opt/yue2/.venv-sheetsage2`, while its persistent model
files live below `/data/models/yue2`. The environment intentionally reuses the
image's PyTorch 2.10/CUDA 12.8 runtime: the upstream cu126 recipe is not
Blackwell-capable. Required persistent directories are:
```text
/data/models/yue2/
├── SheetSage2/
└── MERT-v2-FullSong/
```
The local `SheetSage2/config.json` must point `base_model_name_or_path` to
`/opt/yue2/models/MERT-v2-FullSong`, allowing the complete transcription path
to run offline. Analysis and generation share the RTX 5080 and therefore run
sequentially; the UI parks YuE2 before starting SheetSage2.
The small integration patch under `community-webui-patches/` fixes the
community release's missing `refreshArt()` function on Linux and permits the
native YuE2 empty-lyrics request for true instrumentals. It deliberately does