~/matteospanio

> whoami

Matteo Spanio

AI Research Engineer · University of Padua

Matteo Spanio

I’m an AI research engineer working on audio, music and multimodal generative models, and a PhD candidate in Brain, Mind and Computer Science at the University of Padua, based at the Centro di Sonologia Computazionale. In 2025 I spent six months as a visiting PhD student at the Music Technology Group of Universitat Pompeu Fabra in Barcelona.

Three threads run through the work: recovering audio cultural heritage from degraded magnetic tape, as part of the IEEE 3302-2022 standard I help develop at MPAI; teaching generative models the correspondences between taste and sound; and asking what language models actually understand about music notation.

Alongside it I maintain TorchFX, an audio DSP library that runs on the GPU, and I play clarinet with the Orchestra di Padova e del Veneto. The long version is in the CV.

research map

23 papers, 6 projects, 5 posts and 8 news, placed by how similar their text is.
how this is made
Titles and abstracts are embedded with intfloat/multilingual-e5-base over 108 text chunks and projected with UMAP; the lines join the closest pairs. Colour comes from my declared research areas, not from clustering the embedding.
Read the geometry loosely. The areas separate only weakly here — silhouette 0.30 — because nearly everything is audio, music and machine learning: within-area similarity 0.86, between-area 0.82. And the picture flatters them, since the same score in the full embedding is only 0.15. The lines are computed in that embedding rather than in the picture, so who sits next to whom means more than how far apart things look.

AI for audio cultural heritage

Multimodal and crossmodal AI

Symbolic music and LilyPond

Audio DSP tooling

Musicology and performance

Software and research engineering

selected work

all 23 publications →

news

  • Ital-IA 2026
  • Lilybert is available on Hugging Face
  • DAFx 2025
  • New DL model released

all news →

recent writing

all posts →