Qwen3-4B-GiantMIDI-REMI

Qwen3-4B adapted to symbolic piano generation on GiantMIDI-Piano, with the LoRA adapters and the 316 new REMI music-token embeddings merged into the full weights. This is the "out-of-domain pretraining" arm of a CSE 253 (UCSD) study on whether pretraining domain or scale matters more for symbolic music.

What was trained

The text base never saw music. We added the 316 REMI tokens (<MIDI_TOK_0..315>) as special tokens, froze the base in 4-bit, and trained LoRA (r=16) plus only the 316 new embedding rows. The full bf16 weights here are the merged result.

Usage

The tokenizer includes the 316 <MIDI_TOK_i> tokens, each mapping 1:1 to a MidiTok REMI id (316-token REMI vocabulary). Generation is constrained to those tokens, then decoded to MIDI with MidiTok.

Results (test set, REMI vocabulary)

Perplexity 13.55, pitch KL 0.49, interval KL 0.51, repetition 0.022, 0% invalid.

Base model: Qwen/Qwen3-4B (Apache-2.0). Coursework artifact.

Downloads last month
10
Safetensors
Model size
4B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for sullivanUCSD/Qwen3-4B-GiantMIDI-REMI

Finetuned
Qwen/Qwen3-4B
Finetuned
(1153)
this model