Enable SheetSage2 audio remix in YuE2 UI
This commit is contained in:
@@ -44,6 +44,29 @@ and stable-diffusion.cpp are not part of the Athena setup. Manual composition,
|
||||
generation, result playback, score editing and the take library work without
|
||||
those optional components.
|
||||
|
||||
### Uploaded-audio remix (SheetSage2)
|
||||
|
||||
The **Remix a take** drawer also accepts WAV, FLAC, MP3, M4A, OGG, Opus and
|
||||
AAC uploads. SheetSage2 transcribes the recording into an editable ABC melody
|
||||
and chord plan, then YuE2 renders that structure in a newly selected style.
|
||||
It does not preserve the original samples, singer or production verbatim.
|
||||
|
||||
SheetSage2 is kept in `/opt/yue2/.venv-sheetsage2`, while its persistent model
|
||||
files live below `/data/models/yue2`. The environment intentionally reuses the
|
||||
image's PyTorch 2.10/CUDA 12.8 runtime: the upstream cu126 recipe is not
|
||||
Blackwell-capable. Required persistent directories are:
|
||||
|
||||
```text
|
||||
/data/models/yue2/
|
||||
├── SheetSage2/
|
||||
└── MERT-v2-FullSong/
|
||||
```
|
||||
|
||||
The local `SheetSage2/config.json` must point `base_model_name_or_path` to
|
||||
`/opt/yue2/models/MERT-v2-FullSong`, allowing the complete transcription path
|
||||
to run offline. Analysis and generation share the RTX 5080 and therefore run
|
||||
sequentially; the UI parks YuE2 before starting SheetSage2.
|
||||
|
||||
The small integration patch under `community-webui-patches/` fixes the
|
||||
community release's missing `refreshArt()` function on Linux and permits the
|
||||
native YuE2 empty-lyrics request for true instrumentals. It deliberately does
|
||||
|
||||
Reference in New Issue
Block a user