Publications

Export 46 results:
Author Title Type [ Year(Asc)]
Filters: First Letter Of Last Name is M  [Clear All Filters]
2024
Melechovsky J., Mehrish A., Sisman B., Herremans D..  2024.  Accent Conversion in Text-To-Speech Using Multi-Level VAE and Adversarial Training. Proc. of IEEE Tencon, Singapore.
Melechovsky J., Mehrish A., Sisman B., Herremans D..  2024.  Accent Conversion in Text-To-Speech Using Multi-Level VAE and Adversarial Training. Proc. of IEEE Tencon, Singapore.
Melechovsky J., Mehrish A., Sisman B., Herremans D..  2024.  Accented Text-to-Speech Synthesis with a Conditional Variational Autoencoder. Proc. of IEEE Tencon, Singapore.
Melechovsky J., Mehrish A., Sisman B., Herremans D..  2024.  Accented Text-to-Speech Synthesis with a Conditional Variational Autoencoder. Proc. of IEEE Tencon, Singapore.
Melechovsky J., Mehrish A., Sisman B., Herremans D..  2024.  DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech. Audio Imagination: NeurIPS 2024 Workshop.
Melechovsky J., Mehrish A., Sisman B., Herremans D..  2024.  DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech. Audio Imagination: NeurIPS 2024 Workshop.
Melechovsky J., Roy A., Herremans D..  2024.  MidiCaps — A large-scale MIDI dataset with text captions. ISMIR. PDF icon 2406.02255v1.pdf (699.83 KB)
Melechovsky J, Guo Z, Ghosal D, Majumder N, Herremans D, Poria S.  2024.  Mustango: Toward Controllable Text-to-Music Generation. Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers). pages 8293–8316. PDF icon 2311.08355 (1).pdf (11.38 MB)
Melechovsky J, Guo Z, Ghosal D, Majumder N, Herremans D, Poria S.  2024.  Mustango: Toward Controllable Text-to-Music Generation. Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers). pages 8293–8316. PDF icon 2311.08355 (1).pdf (11.38 MB)
2022
Makris D., Guo Z, Kaliakatsos-Papakostas N., Herremans D..  2022.  Conditional Drums Generation using Compound Word Representations. EvoMUSART (EVO*) - Lecture Notes in Computer Science. PDF icon 2202.04464.pdf (525.36 KB)
BT B, Hee H.I., Ming C., Lin Y., Priyadarshinee P., Clarke C.J., Herremans D., Chen J.M..  2022.  A Gaussian mixture classifier model to differentiate respiratory symptoms using phonated /ɑː/ sounds. The 18th Australasian International Conference on Speech Science and Technology (SST). PDF icon ahsounds.pdf (1018.01 KB)
Turian J, Shier J, Khan HRaj, Raj B, Schuller BW, Steinmetz CJ, Malloy C, Tzanetakis G, Velarde G, McNally K et al..  2022.  HEAR 2021: Holistic Evaluation of Audio Representations. Proceedings of Machine Learning Research (PMLR): NeurIPS 2021 Competition Track. PDF icon 2203.03022.pdf (406.58 KB)
Turian J, Shier J, Khan HRaj, Raj B, Schuller BW, Steinmetz CJ, Malloy C, Tzanetakis G, Velarde G, McNally K et al..  2022.  HEAR 2021: Holistic Evaluation of Audio Representations. Proceedings of Machine Learning Research (PMLR): NeurIPS 2021 Competition Track. PDF icon 2203.03022.pdf (406.58 KB)
Turian J, Shier J, Khan HRaj, Raj B, Schuller BW, Steinmetz CJ, Malloy C, Tzanetakis G, Velarde G, McNally K et al..  2022.  HEAR 2021: Holistic Evaluation of Audio Representations. Proceedings of Machine Learning Research (PMLR): NeurIPS 2021 Competition Track. PDF icon 2203.03022.pdf (406.58 KB)
Kaliakatsos-Papakostas N., Bastas G., Makris D., Herremans D., Katsouros V., Maragos P..  2022.  A Machine Learning Approach for MIDI to Guitar Tablature Conversion. Sound and Music Computing Conference (SMC). PDF icon 25.pdf (528.42 KB)
Kaliakatsos-Papakostas N., Bastas G., Makris D., Herremans D., Katsouros V., Maragos P..  2022.  A Machine Learning Approach for MIDI to Guitar Tablature Conversion. Sound and Music Computing Conference (SMC). PDF icon 25.pdf (528.42 KB)
Guo R, Simpton I., Kiefer C., Magnusson T, Herremans D..  2022.  MusIAC: An extensible generative framework for Music Infilling Application with multi-level Control. EvoMUSART. PDF icon 2202.05528.pdf (893.23 KB)
Chua P., Makris D., Agres K., Roig G., Herremans D..  2022.  Predicting emotion from music videos: exploring the relative contribution of visual and auditory information to affective responses. Arxiv preprint.
2021
Makris D., Agres K., Herremans D..  2021.  Generating Lead Sheets with Affect: A Novel Conditional seq2seq Framework. Proceedings of the International Joint Conference on Neural Networks (IJCNN). PDF icon 2104.13056.pdf (857.78 KB)
Guo Z, Makris D., Herremans D..  2021.  Hierarchical Recurrent Neural Networks for Conditional Melody Generation with Long-term Structure. Proceedings of the International Joint Conference on Neural Networks (IJCNN). PDF icon 2102.09794.pdf (1015.73 KB)
Agres K., Schaefer R, Volk A, Van Hooren S, Holzapfel A, Bella SDalla, Müller M, de Witte M, Herremans D., Melendez RRamirez et al..  2021.  Music, Computing, and Health: A roadmap for the current and future roles of music technology for healthcare and well-being. Music & Science. PDF icon Preprint for OSF_Agres, Schaefer, Volk, et al. (2021)_Music & Science_watermark.pdf (4.07 MB)
Agres K., Schaefer R, Volk A, Van Hooren S, Holzapfel A, Bella SDalla, Müller M, de Witte M, Herremans D., Melendez RRamirez et al..  2021.  Music, Computing, and Health: A roadmap for the current and future roles of music technology for healthcare and well-being. Music & Science. PDF icon Preprint for OSF_Agres, Schaefer, Volk, et al. (2021)_Music & Science_watermark.pdf (4.07 MB)
Agres K., Schaefer R, Volk A, Van Hooren S, Holzapfel A, Bella SDalla, Müller M, de Witte M, Herremans D., Melendez RRamirez et al..  2021.  Music, Computing, and Health: A roadmap for the current and future roles of music technology for healthcare and well-being. Music & Science. PDF icon Preprint for OSF_Agres, Schaefer, Volk, et al. (2021)_Music & Science_watermark.pdf (4.07 MB)
Agres K., Schaefer R, Volk A, Van Hooren S, Holzapfel A, Bella SDalla, Müller M, de Witte M, Herremans D., Melendez RRamirez et al..  2021.  Music, Computing, and Health: A roadmap for the current and future roles of music technology for healthcare and well-being. Music & Science. PDF icon Preprint for OSF_Agres, Schaefer, Volk, et al. (2021)_Music & Science_watermark.pdf (4.07 MB)
2013
Herremans D., Martens D, Sörensen K..  2013.  Dance Hit Song Science. International Workshop on Music and Machine Learning. PDF icon abstract_preprint_MML2013_DH.pdf (194.82 KB)