model card: long-mix numbers
Browse files
README.md
CHANGED
|
@@ -33,9 +33,11 @@ inverse and the chunked inference with linear fades run in the host —
|
|
| 33 |
|
| 34 |
- The exported network is asserted equal to `model(chunk)` before export.
|
| 35 |
- Spectrogram layout and masked inverse: 151–155 dB PSNR against torch.
|
| 36 |
-
- Each
|
| 37 |
-
end to end 97–
|
| 38 |
-
|
|
|
|
|
|
|
| 39 |
- The host runs every chunk twice and settles a mismatch with a third run, because Core AI's GPU
|
| 40 |
was measured to return a slightly wrong result now and then on other models.
|
| 41 |
|
|
|
|
| 33 |
|
| 34 |
- The exported network is asserted equal to `model(chunk)` before export.
|
| 35 |
- Spectrogram layout and masked inverse: 151–155 dB PSNR against torch.
|
| 36 |
+
- Each of 72 chunks (a 10-second clip and a 6-minute mix) through Core AI on the GPU: 74–110 dB
|
| 37 |
+
PSNR against upstream's output; whole clips end to end 97–114 dB on every stem. Lower than a
|
| 38 |
+
convolutional model's 140 dB because eight transformer layers of fp32 attention accumulate
|
| 39 |
+
GPU-versus-CPU rounding — 1e-4 relative on the worst chunk, 1e-5 typical, repeatable run to
|
| 40 |
+
run, far below anything audible.
|
| 41 |
- The host runs every chunk twice and settles a mismatch with a third run, because Core AI's GPU
|
| 42 |
was measured to return a slightly wrong result now and then on other models.
|
| 43 |
|