Publications

Export 27 results:
Author Title Type [ Year(Desc)]
Filters: First Letter Of Last Name is P  [Clear All Filters]
2019
Sturm B., Ben-Tal O., Monaghan U., Collins N., Herremans D., Chew E., Hadjeres G., Deruty E., Pachet F..  2019.  Machine Learning Research that Matters for Music Creation: A Case Study. Journal of New Music Research. 48(1):36-55.PDF icon concert_paper_preprint.pdf (1.6 MB)
T. Phuong HThi, Herremans D., Roig G..  2019.  Multimodal Deep Models for Predicting Affective Responses Evoked by Movies. The 2nd International Workshop on Computer Vision for Physiological Measurement as part of ICCV. Seoul, South Korea. 2019. PDF icon 1909.06957.pdf (836.3 KB)
2021
T. Phuong HThi, BT B, Roig G., Herremans D..  2021.  AttendAffectNet – Emotion Prediction of Movie Viewers Using Multimodal Fusion with Self-attention. Sensors. Special issue on Intelligent Sensors: Sensor Based Multi-Modal Emotion Recognition. PDF icon sensors-21-08356.pdf (1.03 MB)
T. Phuong HThi, BT B, Herremans D., Roig G..  2021.  AttendAffectNet: Self-Attention based Networks for Predicting Affective Responses from Movies. Proceedings of the International Conference on Pattern Recognition (ICPR2020). PDF icon 2010.11188.pdf (7.07 MB)
2022
Clarke C.J., Chowdhury J., BT B, Priyadarshinee P., Lim C.M.Ying, I. Tan FXing, Herremans D., Chen J.M..  2022.  Computationally Efficient Physics Approximating Neural Networks for Highly Nonlinear Maps. 2022 International Conference on Research in Adaptive and Convergent Systems.
Pham Q-H, Herremans D., Roig G..  2022.  EmoMV: Affective Music-Video Correspondence Learning Datasets for Classification and Retrieval. Information Fusion. PDF icon SSRN-id4189323.pdf (2.01 MB)
BT B, Hee H.I., Ming C., Lin Y., Priyadarshinee P., Clarke C.J., Herremans D., Chen J.M..  2022.  A Gaussian mixture classifier model to differentiate respiratory symptoms using phonated /ɑː/ sounds. The 18th Australasian International Conference on Speech Science and Technology (SST). PDF icon ahsounds.pdf (1018.01 KB)
Turian J, Shier J, Khan HRaj, Raj B, Schuller BW, Steinmetz CJ, Malloy C, Tzanetakis G, Velarde G, McNally K et al..  2022.  HEAR 2021: Holistic Evaluation of Audio Representations. Proceedings of Machine Learning Research (PMLR): NeurIPS 2021 Competition Track. PDF icon 2203.03022.pdf (406.58 KB)
Sockalingam N., Lo K., n KO., Herremans D., Raghunath N., Cancion H.GC, Kejun H., Leong H., Tan J., Nizharzharudin K. et al..  2022.  A white paper on cyberphysical learning. White paper, Singapore University of Technology and Design. PDF icon LSL_WhitePaper_Cyber-physical-Campus-Higher-Education.pdf (6.98 MB)
2024
Lanzendörfer L.A., Lu T., Perraudin N., Herremans D., Wattenhofer R..  2024.  Coarse-to-Fine Text-to-Music Latent Diffusion. Audio Imagination: NeurIPS 2024 Workshop.
Melechovsky J, Guo Z, Ghosal D, Majumder N, Herremans D, Poria S.  2024.  Mustango: Toward Controllable Text-to-Music Generation. Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers). pages 8293–8316. PDF icon 2311.08355 (1).pdf (11.38 MB)
Kang J, Poria S, Herremans D..  2024.  Video2Music: Suitable Music Generation from Videos using an Affective Multimodal Transformer model. Expert Systems with Applications. PDF icon 2311.00968.pdf (5.51 MB)
2026
Puri G., Socklingam N., Herremans D..  2026.  Digital Lifelong Learning in the Age of AI: Trends and Insights.
A. Putri M, Saide S., D. Riau K, Herremans D..  2026.  Generative AI in Education for SDG 4: Insights from Indonesia and Kazakhstan. Proceedings of the Pacific Asia Conference on Information Systems (PACIS)..
Liu R., Hung C.Y., Majumder N., Herremans D., Poria S..  2026.  JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment. Empirical Methods in Natural Language Processing (EMNLP).
Song M., Pala T.D, Jin W., Zadeh A., Li C., Herremans D., Poria S..  2026.  Measuring and Mitigating Rapport Bias of Large Language Models under Multi-Agent Social Interactions. Proceedings of ICLR.
Song M., Pala T.D, Jin W., Zadeh A., Li C., Herremans D., Poria S..  2026.  Measuring and Mitigating Rapport Bias of Large Language Models under Multi-Agent Social Interactions. Proceedings of ICLR.
Jin W., Song M., Pala T.D, Ken C.Y., Herremans D., Poria S..  2026.  PromptDistill: Query-based Selective Token Retention in Intermediate Layers for Efficient Large Language Model Inference. 19th International Natural Language Generation Conference (INLG).
Jin W., Song M., Pala T.D, Ken C.Y., Herremans D., Poria S..  2026.  PromptDistill: Query-based Selective Token Retention in Intermediate Layers for Efficient Large Language Model Inference. 19th International Natural Language Generation Conference (INLG).
Jiang Z., Yeo S., Herremans D., Perrault S..  2026.  Scaffolded Vulnerability: Chatbot-Mediated Reciprocal Self-Disclosure and Need-Supportive Interaction in Couples. Proceedings of CHI.
Roy A., Puri G., Herremans D..  2026.  Text2midi-InferAlign: Improving Symbolic Music Generation with Inference-Time Alignment. ICASSP. PDF icon 2505.12669v1.pdf (360.69 KB)