Files
AI-Profile-Router/experiments/yue2-3b/README.md
T

53 lines
1.8 KiB
Markdown

# YuE2 3B isolated quality test
Prepared, non-starting evaluation of `m-a-p/YuE2-3B` with the standard
`m-a-p/YuE2-Vae` listening decoder. The source is pinned to the official
`yue2-v0.1.6` commit `9c6c4b349be978b06a9d0d958471a07a6cdeff4d`.
Preparation on Athena is complete. The model and VAE files were checked
against their published `weights_manifest.json` SHA-256 values. The Docker
image is built, but no YuE2 container has been created or started.
## Safety and isolation
- This experiment is not part of the profile controller or dashboard.
- The Compose service uses the `manual` profile, has no restart policy and
cannot start through an ordinary `docker compose up`.
- Only the RTX 5080 is exposed to the container.
- Building and downloading do not load the model or use a GPU.
- Do not start it while another Athena GPU job is active.
## Persistent files
```text
/data/models/yue2/
├── YuE2-3B/
└── YuE2-Vae/
/data/music/yue2/
```
The initial control request is a true empty-lyrics instrumental request. No
invented `[Instrumental]` lyrics marker is used.
## Manual test (only after GPU availability was checked)
From `/opt/mike-ai/yue2-3b` on Athena:
```sh
docker compose --profile manual run --rm yue2-test generate \
--offline \
--device cuda:0 \
--budget 16 \
--request /workspace/requests/instrumental-synthwave.json \
--output /workspace/runs
```
Start with the official unquantized BF16 path. If and only if this fails from
VRAM pressure, repeat with `--quantization fp8 --offload-ar`; keep the outputs
separate because that is a different inference configuration.
YuE2 is newly released and officially specifies a 24-GB BF16 GPU. Readiness of
this image and the downloaded weights is not evidence that the 16-GB RTX 5080
run will fit or that its audio quality is acceptable.