Add pipeline tag, link to paper and code

#1
by nielsr HF Staff - opened
Files changed (1) hide show
  1. README.md +10 -5
README.md CHANGED
@@ -1,13 +1,18 @@
1
  ---
2
  tags:
3
- - audio
4
- - music-generation
5
- - midi-to-audio
6
- - audio-super-resolution
 
7
  ---
8
 
9
  # CPR: COMBINING GLOBAL COMPOSING, LOCAL PERFORMING AND FULL-SEQUENCE REFINING IN PIANO RENDERING WITH CONTINUOUS AUTOREGRESSIVE MODELLING
10
 
 
 
 
 
11
  1. **Composer–Performer (CP)** takes prompt audio, prompt MIDI, and target MIDI,
12
  and renders **24 kHz mono audio**. The Composer is an autoregressive Qwen3
13
  Transformer; the Performer renders local Mel spectrograms with flow matching.
@@ -73,4 +78,4 @@ use any existing 24 kHz WAV or the output of Composer-Performer:
73
 
74
  ```bash
75
  python infer_refiner.py --input /path/to/audio_24k.wav --output outputs/refined.wav
76
- ```
 
1
  ---
2
  tags:
3
+ - audio
4
+ - music-generation
5
+ - midi-to-audio
6
+ - audio-super-resolution
7
+ pipeline_tag: audio-to-audio
8
  ---
9
 
10
  # CPR: COMBINING GLOBAL COMPOSING, LOCAL PERFORMING AND FULL-SEQUENCE REFINING IN PIANO RENDERING WITH CONTINUOUS AUTOREGRESSIVE MODELLING
11
 
12
+ This model is presented in the paper [CPR: Combining global composing, local performing and full-sequence refining in piano rendering with continuous autoregressive modelling](https://huggingface.co/papers/2609.18216).
13
+
14
+ Code is available at [github.com/FEAfeatherTHER/CPR_official](https://github.com/FEAfeatherTHER/CPR_official).
15
+
16
  1. **Composer–Performer (CP)** takes prompt audio, prompt MIDI, and target MIDI,
17
  and renders **24 kHz mono audio**. The Composer is an autoregressive Qwen3
18
  Transformer; the Performer renders local Mel spectrograms with flow matching.
 
78
 
79
  ```bash
80
  python infer_refiner.py --input /path/to/audio_24k.wav --output outputs/refined.wav
81
+ ```