Publications

Export 28 results:
Author Title Type [ Year(Asc)]
Filters: First Letter Of Last Name is R  [Clear All Filters]
2026
Herremans D., Roy A..  2026.  Aligning Generative Music AI with Human Preferences: Methods and Challenges. Proceedings of AAAI, senior member track. PDF icon 2511.15038v1.pdf (417.24 KB)
Bhandari K., Roy A., Colton S., Herremans D..  2026.  Emerging AI Technologies for Music: Towards Controllable, Collaborative, and Creative Systems. Proceedings of Machine Learning Research, PMLR 303:1-5, 2026. PDF icon bhandari26a.pdf (161.47 KB)
A. Putri M, Saide S., D. Riau K, Herremans D..  2026.  Generative AI in Education for SDG 4: Insights from Indonesia and Kazakhstan. Proceedings of the Pacific Asia Conference on Information Systems (PACIS)..
Ghosh A., Roy A., Herremans D..  2026.  KARMA-MV: A Benchmark for Causal Question Answering on Music Videos. arXiv:2605.08175. PDF icon 2605.08175v1.pdf (3.32 MB)
Liu R., Roy A., Herremans D..  2026.  Leveraging LLM Embeddings for Cross Dataset Label Alignment and Zero Shot Music Emotion Prediction. Conference on AI Music Creativity (AIMC).
Lu T., Geist C-M, Melechovsky J., Roy A., Herremans D..  2026.  MelodySim: Measuring Melody-aware Music Similarity for Plagiarism Detection. IEEE Tencon.
Roy A, Liang J, Herremans D.  2026.  nnAudio 2: Overcoming Dynamic Compilation Barriers and Transform Inconsistencies. Conference on AI Music Creativity (AIMC).
Melechovsky J., Mehrish A., Roy A., Herremans D..  2026.  SonicMaster: Towards Controllable All-in-One Music Restoration and Mastering. Proceedings of ICML. PDF icon 2508.03448v2.pdf (3.31 MB)
Roy A., Puri G., Herremans D..  2026.  Text2midi-InferAlign: Improving Symbolic Music Generation with Inference-Time Alignment. ICASSP. PDF icon 2505.12669v1.pdf (360.69 KB)
Bhandari K., Chang S., Roy A., Ronchini F., Benetos E., Herremans D., Colton S..  2026.  Text2Score: Generating Sheet Music From Textual Prompts. arXiv:2605.13431. PDF icon 2605.13431v1.pdf (395.67 KB)
Bhandari K., Chang S., Roy A., Ronchini F., Benetos E., Herremans D., Colton S..  2026.  Text2Score: Generating Sheet Music From Textual Prompts. arXiv:2605.13431. PDF icon 2605.13431v1.pdf (395.67 KB)
2025
Roy A., Liu R., Lu T., Herremans D..  2025.  JamendoMaxCaps: A Large Scale Music-caption Dataset with Imputed Metadata. Proceedings of IJCNN, Rome, Italy.
Chopra A., Roy A., Herremans D..  2025.  SonicVerse: Multi-Task Learning for Music Feature-Informed Captioning. Proceedings of the 6th Conference on AI Music Creativity (AIMC 2025), Brussels, Belgium, September 10th - 12th, 2025.
Bhandari K., Roy A., Wang K., Puri G., Colton S., Herremans D..  2025.  Text2midi: Generating Symbolic Music from Captions. Proceedings of AAAI, Philadelphia. PDF icon 2412.16526v2.pdf (569.51 KB)
2020
Cheuk K.W., Luo Y.J., BT B, Roig G., Herremans D..  2020.  Regression-based music emotion prediction using triplet neural networks. Proceedings of the International Joint Conference on Neural Networks (IJCNN). PDF icon 2001.09988.pdf (777.31 KB)
2019
Cheuk K.W., BT B, Roig G., Herremans D..  2019.  Latent space representation for multi-target speaker detection and identification with a sparse dataset using Triplet neural networks. IEEE Automatic Speech Recognition and Understanding Workshop (ASRU 2019). PDF icon 1910.01463.pdf (934.76 KB)
T. Phuong HThi, Herremans D., Roig G..  2019.  Multimodal Deep Models for Predicting Affective Responses Evoked by Movies. The 2nd International Workshop on Computer Vision for Physiological Measurement as part of ICCV. Seoul, South Korea. 2019. PDF icon 1909.06957.pdf (836.3 KB)