Model Aria is a pretrained autoregressive generative model for symbolic music based on the LLaMA 3.2 (1B) architecture. It was trained on ~60k hours of MIDI transcriptions of expressive solo piano recordings. It has been finetuned to produce realistic continuations of solo piano compositions as well as to produce general purpose contrastive MIDI embeddings. This HuggingFace page contains weights and usage instructions for the embedding model. For the pretrained base model, see aria medium base, and for the generative model, see aria medium gen. 📖 Read our release blog post and paper 🚀 Check out the real time demo in the official GitHub repository 📊 Get access to our training dataset Aria MIDI to train your own models Usage Guidelines Our embedding model was trained to capture composition and performance level attributes by learning to embed different random slices of transcriptions of solo piano performances into similar regions of latent space. As the model was trained to produce global embeddings with data augmentation (e.g., pitch, tempo, etc.), it might not be appropriate for every use case. For more information, see our paper. Quickstart All of our models were trained using…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy