Publications
Export 79 results:
Author Title Type [ Year
] Filters: First Letter Of Last Name is C [Clear All Filters]
.
2026. Emerging AI Technologies for Music: Towards Controllable, Collaborative, and Creative Systems. Proceedings of Machine Learning Research, PMLR 303:1-5, 2026.
bhandari26a.pdf (161.47 KB)
.
2026. Text2Score: Generating Sheet Music From Textual Prompts. arXiv:2605.13431.
2605.13431v1.pdf (395.67 KB)
.
2026. Text2Score: Generating Sheet Music From Textual Prompts. arXiv:2605.13431.
2605.13431v1.pdf (395.67 KB)
.
2025. ImprovNet: Generating Controllable Musical Improvisations with Iterative Corruption Refinement. Proceedings of IJCNN.
.
2025. ImprovNet: Generating Controllable Musical Improvisations with Iterative Corruption Refinement. Proceedings of IJCNN.
.
2025. PRESENT: Zero-Shot Text-to-Prosody Control. IEEE Signal Processing Letters.
2408.06827v1.pdf (367.55 KB)
.
2025. SonicVerse: Multi-Task Learning for Music Feature-Informed Captioning. Proceedings of the 6th Conference on AI Music Creativity (AIMC 2025), Brussels, Belgium, September 10th - 12th, 2025.
.
2025. Text2midi: Generating Symbolic Music from Captions. Proceedings of AAAI, Philadelphia.
2412.16526v2.pdf (569.51 KB)
.
2025. Towards the future of education: cyber-physical learning. Discover Education. 4:1–16.
.
2024. Gamification and skills tree. Trends and Foresight Report on Cyber-Physical Learning.
.
2024. MIRFLEX: Music Information Retrieval Feature Library for Extraction. ISMIR, Late Breaking Demos.
2411.00469v1.pdf (89.86 KB)
.
2024. SNIPER Training: Variable Sparsity Rate Training For Text-To-Speech. Proc. of IEEE Tencon, Singapore.
2211.07283.pdf (435.22 KB)
.
2023. DiffRoll: Diffusion-based Generative Music Transcription with Unsupervised Pretraining Capability. ICASSP.
diffroll.pdf (2.2 MB)
.
2023. MERP: A Music Dataset with Emotion Ratings and Raters’ Profile Information. Sensors - Intelligent Sensors. 23(1)
sensors-23-00382 (2).pdf (1.21 MB)
.
2022. Computationally Efficient Physics Approximating Neural Networks for Highly Nonlinear Maps. 2022 International Conference on Research in Adaptive and Convergent Systems.
.
2022. Computationally Efficient Physics Approximating Neural Networks for Highly Nonlinear Maps. 2022 International Conference on Research in Adaptive and Convergent Systems.
.
2022. Computationally Efficient Physics Approximating Neural Networks for Highly Nonlinear Maps. 2022 International Conference on Research in Adaptive and Convergent Systems.
.
2022. A Gaussian mixture classifier model to differentiate respiratory symptoms using phonated /ɑː/ sounds. The 18th Australasian International Conference on Speech Science and Technology (SST).
ahsounds.pdf (1018.01 KB)
.
2022. A Gaussian mixture classifier model to differentiate respiratory symptoms using phonated /ɑː/ sounds. The 18th Australasian International Conference on Speech Science and Technology (SST).
ahsounds.pdf (1018.01 KB)
.
2022. HEAR 2021: Holistic Evaluation of Audio Representations. Proceedings of Machine Learning Research (PMLR): NeurIPS 2021 Competition Track.
2203.03022.pdf (406.58 KB)
.
2022. Jointist: Joint Learning for Multi-instrument Transcription and Its Applications.
2206.10805.pdf (427.51 KB)
.
2022. Jointist: Joint Learning for Multi-instrument Transcription and Its Applications.
2206.10805.pdf (427.51 KB)
.
2022. Predicting emotion from music videos: exploring the relative contribution of visual and auditory information to affective responses. Arxiv preprint.
.
2022. Single Image Video Prediction with Auto-Regressive GANs. Sensors. 22:3533.
.
2022. Understanding Audio Features via Trainable Basis Functions. Arxiv preprint.
2204.11437.pdf (7.36 MB)
.
2022. A white paper on cyberphysical learning. White paper, Singapore University of Technology and Design.
LSL_WhitePaper_Cyber-physical-Campus-Higher-Education.pdf (6.98 MB)
.
2021. Deep Neural Network Based Respiratory Pathology Classification Using Cough Sounds. Sensors. 21(16):5555.
2106.12174.pdf (6.52 MB)
.
2021. The Effect of Spectrogram Reconstructions on Automatic Music Transcription:An Alternative Approach to Improve Transcription Accuracy. Proceedings of the International Conference on Pattern Recognition (ICPR2020).
2010.09969.pdf (3.46 MB)
.
2021. Evaluating the Effectiveness of an Augmented Reality Game Promoting Environmental Action. Sustainability. 13(24):13912.
sustainability-13-13912.pdf (16.23 MB)
.
2021. ReconVAT: A Semi-Supervised Automatic Music Transcription Framework for Low-Resource Real-World Data. ACM Multimedia.
.
2021. Revisiting the Onsets and Frames Model with Additive Attention. Proceedings of the International Joint Conference on Neural Networks (IJCNN).
2104.06607.pdf (1.52 MB)
.
2020. Acoustic prediction of flowrate: varying liquid jet stream onto a free surface. IEEE International Conference on Signal Processing and Communications (SPCOM).
preprint flow.pdf (1.01 MB)
.
2020. Acoustic prediction of flowrate: varying liquid jet stream onto a free surface. IEEE International Conference on Signal Processing and Communications (SPCOM).
preprint flow.pdf (1.01 MB)
.
2020. Asthmatic versus healthy child classification based on cough and vocalised /a:/ sounds. The Journal of the Acoustical Society of America (JASA). 148, EL253
.
2020. The impact of Audio input representations on neural network based music transcription. Proceedings of the International Joint Conference on Neural Networks (IJCNN).
2001.09989.pdf (1.87 MB)
.
2020. nnAudio: An on-the-fly GPU Audio to Spectrogram Conversion Toolbox Using 1D Convolution Neural Networks. IEEE Access.
nnAudio.pdf (10.2 MB)
.
2020. Regression-based music emotion prediction using triplet neural networks. Proceedings of the International Joint Conference on Neural Networks (IJCNN).
2001.09988.pdf (777.31 KB)
.
2020. Unsupervised disentanglement of pitch and timbre for isolated musical instrument sounds. Proceedings of the International Society of Music Information Retrieval (ISMIR).
.
2019. Development of Machine Learning for asthmatic and healthy voluntary cough - a proof of concept study. Applied Sciences. 9(14)
applsci-09-02833.pdf (2.06 MB)
.
2019. The emergence of deep learning: new opportunities for music and audio technologies. Neural Computing and Applications.
main_preprint.pdf (102.16 KB)
.
2019. Latent space representation for multi-target speaker detection and identification with a sparse dataset using Triplet neural networks. IEEE Automatic Speech Recognition and Understanding Workshop (ASRU 2019).
1910.01463.pdf (934.76 KB)
.
2019. Machine Learning Research that Matters for Music Creation: A Case Study. Journal of New Music Research. 48(1):36-55.
concert_paper_preprint.pdf (1.6 MB)
.
2019. Machine Learning Research that Matters for Music Creation: A Case Study. Journal of New Music Research. 48(1):36-55.
concert_paper_preprint.pdf (1.6 MB)
.
2019. nnAudio: A PyTorch Audio Processing Tool Using 1D Convolution neural networks. ISMIR - Late Breaking Demo.
nnAudio.pdf (399.08 KB)
.
2019. Towards emotion based music generation: A tonal tension model based on the spiral array. Proceedings of Cognitive Science (CogSci).
CogSci_tension (1).pdf (610.91 KB)
.
2019. Towards robust audio spoofing detection: a detailed comparison of traditional and learned features. IEEE Access. 7:84229-84241.
ieee_access_herremans.pdf (14.31 MB)
.
2018. Blacklisted speaker identification using triplet neural networks. MCE2018 competition.
SUTD_description.pdf (133.08 KB)
.
2018. From Context to Concept: Exploring Semantic Relationships in Music with Word2Vec. Neural Computing and Applications.
paper.pdf (1.64 MB)
.
2018. Minimally Simple Binaural Room Modelling Using a Single Feedback Delay Network. Journal of the Audio Engineering Society. 66(10):791-807.
angus_jaes_preprint.pdf (6.39 MB)
.
2018. Modeling temporal tonal relations in polyphonic music through deep networks with a novel image-based representation. The Thirty-Second AAAI Conference on Artificial Intelligence.
preprint_lstm.pdf (741.28 KB)
.
2018. A Novel Interface for the Graphical Analysis of Music Practice Behaviours. Frontiers in Psychology - Human-Media Interaction. 9
practice_browser.pdf (4.9 MB)
.
2018. O.R. and music generation. OR/MS Today. 45(1)
O.R. and music generation - INFORMS.pdf (825.66 KB)
.
2018. Perceptual evaluation of measures of spectral variance. Journal of the Acoustical Society of America. 143(6):3300–3311.
jasa_an_dh_preprint.pdf (2.46 MB)
.
2017. A Functional Taxonomy of Music Generation Systems. ACM Computing Surveys. 50(5):30.
music_generation_survey_dh_preprint.pdf (349.15 KB)
.
2017. A Functional Taxonomy of Music Generation Systems. ACM Computing Surveys. 50(5):30.
music_generation_survey_dh_preprint.pdf (349.15 KB)
.
2017. Generating guitar solos by integer programming. Journal of the Operational Research Society. :971-985.
preprint_guitar_solo_generation_dh.pdf (772.59 KB)
.
2017. Harmonic Structure Predicts the Enjoyment of Uplifting Trance Music. Frontiers in Psychology, Cognitive Science. 7(1999)
agres16ut.pdf (1.15 MB)
.
2017. IMMA-Emo: A Multimodal Interface for Visualising Score- and Audio-synchronised Emotion Annotations. Audio Mostly.
IMMA-emo_preprint.pdf (1.4 MB)
.
2017. IMMA-Emo: A Multimodal Interface for Visualising Score- and Audio-synchronised Emotion Annotations. Audio Mostly.
IMMA-emo_preprint.pdf (1.4 MB)
.
2017. Modeling Musical Context with Word2vec. First International Workshop On Deep Learning and Music. 1:11-18.
herremans2017work2vec.pdf (745.8 KB)
.
2017. MorpheuS: generating structured music with constrained patterns and tension. IEEE Transactions on Affective Computing. PP (In Press)(99)
herremans2017morpheusFullIEEE.pdf (5.71 MB)
.
2017. A multi-modal platform for semantic music analysis: visualizing audio- and score-based tension. 11th International Conference on Semantic Computing IEEE ICSC 2017.
paper_preprint.pdf (1.63 MB)
.
2017. A variable neighborhood search algorithm to generate piano fingerings for polyphonic sheet music. International Transactions in Operational Research, Special Issue on Variable Neighbourhood Search. 24(3):509–535.
ITOR_VNS_APF_preprint.pdf (840.28 KB)
.
2016. The Effect of Repetitive Structure on Enjoyment in Uplifting Trance Music. 14th International Conference for Music Perception and Cognition (ICMPC). :280-282.
preprint_trance.pdf (139.27 KB)
.
2016. MorpheuS: Automatic music generation with recurrent pattern constraints and tension profiles. IEEE TENCON.
paper_morpheus_dh_ieee.pdf (550.61 KB)
.
2016. MorpheuS: constraining structure in automatic music generation. Dagstuhl seminar on Computational Music Structure Analysis.
abstract_dagstuhl_dh.pdf (88.49 KB)
.
2016. Music generation with structural constraints: an operations research approach. 30th Annual Conference of the Belgian Operational Research (OR) Society (ORBEL30). :37-39.
orbel30_dh.pdf (117.78 KB)
.
2016. Tension ribbons: Quantifying and visualising tonal tension. Second International Conference on Technologies for Music Notation and Representation (TENOR). 2:8-18.
paper_tenor_dh_preprint_small.pdf (1.67 MB)
.
2016. Uma abordagem baseada em programação linear inteira para a geração de solos de guitarra. XLVIII Simpósio Brasileiro de Pesquisa Operacional (SBPO).
sbpo_dh.pdf (346.61 KB)
.
2015. The effect of repetitive structure on enjoyment and altered states in uplifting trance music. 2nd International Conference on Music and Consciousness (MUSCON 2), Brighton.
AgresEtAl_muscon.pdf (12.47 KB)
.
2015. Generating Fingerings for Polyphonic Piano Music with a Tabu Search Algorithm. Mathematics and Computation in Music. 9110:149-160.
paper_mcm_preprint.pdf (405.73 KB)
.
2015. Generating Fingerings for Polyphonic Piano Music with a Tabu Search Algorithm. Mathematics and Computation in Music. 9110:149-160.
paper_mcm_preprint.pdf (405.73 KB)
.
2015. Generating music with an optimization algorithm using a Markov based objective function. ORBEL29, Belgian Conference on Operations Research.
orbel29abs.pdf (138.67 KB)
.
2015. Generating structured music for bagana using quality metrics based on Markov models. Expert Systems With Applications. 42 (21)(21):424–7435.
paper-bagana.pdf (1.73 MB)
.
2014. First species counterpoint generation with VNS and vertical viewpoints. Annual Conference of the Belgian Operation Research Society (ORBEL28).
orbel28_dh.pdf (216.63 KB)
.
2014. Generating structured music using quality metrics based on Markov models.
wp_bagana.pdf (1.7 MB)
.
2014. Markov Based Quality Metrics For Generating Structured Music With Optimization Techniques. Digital Music Research Network (DMNR+9).
dmrn9_dh.pdf (133.29 KB)
.
2014. Sampling the extrema from statistical models of music with variable neighbourhood search. ICMC/SMC.
icmc_dh.pdf (1.07 MB)
.
2013. First species counterpoint generation with VNS and vertical viewpoints. Digital Music Research Network (DMNR+8).
dnmr8_dh_dc.pdf (147.73 KB)