Published September 3, 2026 | Version v1

Modeling Relations Between Musical Events in Continuous Time with Transformer Models for Live Co-Improvisational Interactions

  • 1. ROR icon University of Oslo
  • 2. University of oslo

Description

This paper presents a live semi-autonomous system that co-improvises with a live musician using a model learned from the relationship observed in paired recordings of an improvising duo. This responsive system is based on a deep learning approach that models the relations between sequences of events produced by co-playing musicians in continuous time, using transformer models. We present the architecture for simultaneous sequences of sonic events and provide a quantitative evaluation on several representative tasks. Results are compared against a canonical transformer and a multichannel Factor Oracle, the latter being a widely used model for live sequence-based symbolic music generation. We detail a live implementation employing concatenative synthesis and introduce a customization procedure in which the generative model is iteratively retrained on curated, satisfactory sections from sound recordings of actual human-system co-improvisational interactions. This iterative fine-tuning enables the model’s stylistic output to diverge from the original training corpus. Two musical use cases demonstrate the application of this technique.

Files

Modeling Relations Between Musical Events in Continuous Time with Transformer Models for Live Co-Improvisational Interactions.pdf

Files (213.4 MB)

Name Size Download all
md5:f028d358a6fc14747f7464a4e5023ee7
52.3 MB Preview Download
md5:5d5fa6d4cbf26329c3ed9fcea6ea425c
52.3 MB Preview Download
md5:931efed8800eb32952f3e87c401e7bbf
52.3 MB Preview Download
md5:da38cbc5203040350a556f923afc6f7a
52.3 MB Preview Download
md5:51d88619fd0047de70315c657bf5e0ce
4.0 MB Preview Download