Publications
Export 46 results:
Author Title Type [ Year
] Filters: First Letter Of Last Name is M [Clear All Filters]
.
2026. Development of Interpretable Deep Learning-based Segmentation Algorithm for Automated Assessment of Oral Diadochokinesis in Progressive Neurological Diseases. Journal of Speech, Language, and Hearing Research.
.
2026. JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment. Empirical Methods in Natural Language Processing (EMNLP).
.
2026. MelodySim: Measuring Melody-aware Music Similarity for Plagiarism Detection. IEEE Tencon.
.
2026. SonicMaster: Towards Controllable All-in-One Music Restoration and Mastering. Proceedings of ICML.
2508.03448v2.pdf (3.31 MB)
.
2026. SonicMaster: Towards Controllable All-in-One Music Restoration and Mastering. Proceedings of ICML.
2508.03448v2.pdf (3.31 MB)
.
2025. Analysis and Synthesis of Audio with AI: from Neurological Disease to Accented Speech and Music.
thesis_Jan.pdf (26.4 MB)
.
2025. End-to-End Text-to-SQL with Dataset Selection: Leveraging LLMs for Adaptive Query Generation. Proceedings of IJCNN, Rome, Italy.
.
2024. Accent Conversion in Text-To-Speech Using Multi-Level VAE and Adversarial Training. Proc. of IEEE Tencon, Singapore.
.
2024. Accent Conversion in Text-To-Speech Using Multi-Level VAE and Adversarial Training. Proc. of IEEE Tencon, Singapore.
.
2024. Accented Text-to-Speech Synthesis with a Conditional Variational Autoencoder. Proc. of IEEE Tencon, Singapore.
.
2024. Accented Text-to-Speech Synthesis with a Conditional Variational Autoencoder. Proc. of IEEE Tencon, Singapore.
.
2024. DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech. Audio Imagination: NeurIPS 2024 Workshop.
.
2024. DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech. Audio Imagination: NeurIPS 2024 Workshop.
.
2024. MidiCaps — A large-scale MIDI dataset with text captions. ISMIR.
2406.02255v1.pdf (699.83 KB)
.
2024. Mustango: Toward Controllable Text-to-Music Generation. Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers). pages 8293–8316.
2311.08355 (1).pdf (11.38 MB)
.
2024. Mustango: Toward Controllable Text-to-Music Generation. Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers). pages 8293–8316.
2311.08355 (1).pdf (11.38 MB)
.
2023. DiffRoll: Diffusion-based Generative Music Transcription with Unsupervised Pretraining Capability. ICASSP.
diffroll.pdf (2.2 MB)
.
2023. DiffRoll: Diffusion-based Generative Music Transcription with Unsupervised Pretraining Capability. ICASSP.
diffroll.pdf (2.2 MB)
.
2023. Learning accent representation with multi-level VAE towards controllable speech synthesis. IEEE Spoken Language Technology (SLT) Workshop.
.
2023. Learning accent representation with multi-level VAE towards controllable speech synthesis. IEEE Spoken Language Technology (SLT) Workshop.
.
2022. Conditional Drums Generation using Compound Word Representations. EvoMUSART (EVO*) - Lecture Notes in Computer Science.
2202.04464.pdf (525.36 KB)
.
2022. A Gaussian mixture classifier model to differentiate respiratory symptoms using phonated /ɑː/ sounds. The 18th Australasian International Conference on Speech Science and Technology (SST).
ahsounds.pdf (1018.01 KB)
.
2022. HEAR 2021: Holistic Evaluation of Audio Representations. Proceedings of Machine Learning Research (PMLR): NeurIPS 2021 Competition Track.
2203.03022.pdf (406.58 KB)
.
2022. HEAR 2021: Holistic Evaluation of Audio Representations. Proceedings of Machine Learning Research (PMLR): NeurIPS 2021 Competition Track.
2203.03022.pdf (406.58 KB)
.
2022. HEAR 2021: Holistic Evaluation of Audio Representations. Proceedings of Machine Learning Research (PMLR): NeurIPS 2021 Competition Track.
2203.03022.pdf (406.58 KB)
.
2022. A Machine Learning Approach for MIDI to Guitar Tablature Conversion. Sound and Music Computing Conference (SMC).
25.pdf (528.42 KB)
.
2022. A Machine Learning Approach for MIDI to Guitar Tablature Conversion. Sound and Music Computing Conference (SMC).
25.pdf (528.42 KB)
.
2022. MusIAC: An extensible generative framework for Music Infilling Application with multi-level Control. EvoMUSART.
2202.05528.pdf (893.23 KB)
.
2021. Generating Lead Sheets with Affect: A Novel Conditional seq2seq Framework. Proceedings of the International Joint Conference on Neural Networks (IJCNN).
2104.13056.pdf (857.78 KB)
.
2021. Hierarchical Recurrent Neural Networks for Conditional Melody Generation with Long-term Structure. Proceedings of the International Joint Conference on Neural Networks (IJCNN).
2102.09794.pdf (1015.73 KB)
.
2021. Music, Computing, and Health: A roadmap for the current and future roles of music technology for healthcare and well-being. Music & Science.
Preprint for OSF_Agres, Schaefer, Volk, et al. (2021)_Music & Science_watermark.pdf (4.07 MB)
.
2021. Music, Computing, and Health: A roadmap for the current and future roles of music technology for healthcare and well-being. Music & Science.
Preprint for OSF_Agres, Schaefer, Volk, et al. (2021)_Music & Science_watermark.pdf (4.07 MB)
.
2020. A variational autoencoder for music generation controlled by tonal tension. Joint Conference on AI Music Creativity (CSMC + MuMe).
2010.06230.pdf (622.82 KB)
.
2019. Machine Learning Research that Matters for Music Creation: A Case Study. Journal of New Music Research. 48(1):36-55.
concert_paper_preprint.pdf (1.6 MB)
.
2019. Midi Miner – A Python library for tonal tension and track classification. ISMIR - Late Breaking Demo.
midi_miner.pdf (83.7 KB)
.
2015. Classification and generation of composer-specific music using global feature models and variable neighborhood search. Computer Music Journal. 39(3):91.
papercmj-dh_preprint.pdf (637.63 KB)
.
2015. Composer Classification Models for Music-Theory Building. Computational Music Analysis.
Chapter_HerremansEtAl_preprint.pdf (475.26 KB)
.
2015. Composer Classification Models for Music-Theory Building. Computational Music Analysis.
Chapter_HerremansEtAl_preprint.pdf (475.26 KB)
.
2015. Generating Fingerings for Polyphonic Piano Music with a Tabu Search Algorithm. Mathematics and Computation in Music. 9110:149-160.
paper_mcm_preprint.pdf (405.73 KB)
.
2014. Dance hit song prediction. Journal of New music Research. 43:302.
wp_hit.pdf (689.07 KB)
.
2013. Dance Hit Song Science. International Workshop on Music and Machine Learning.
abstract_preprint_MML2013_DH.pdf (194.82 KB)