I’m Michele Mancusi, a Senior Applied Scientist at Music.AI (Moises), where I develop and optimize autoencoders and generative models for speech, music, and general audio enhancement.
Previously, I was a Research Scientist at Sony, where I conducted research on deep-learning-based generative models for speech, audio, and music, including Large Language Model (LLM) and diffusion-based approaches.
Earlier in my career, I interned at Microsoft and Musixmatch. At Microsoft, I worked on deep learning for unsupervised speech separation, while at Musixmatch I focused on singing voice detection.
I earned my Ph.D. from Sapienza University of Rome under the supervision of Prof. Emanuele RodolĂ as a member of the Gladia research group. My doctoral research centered on music generation, source separation, and Natural Language Processing (NLP), contributing to advancements in the field of generative AI.
PhD in Computer Science, 2024
Sapienza University of Rome
M.S. in Physics, 2019
Sapienza University of Rome
B.S. in Physics, 2016
Sapienza University of Rome