Publications

Export 164 results:
Author [ Title(Desc)] Type Year
A B C D E F G H I J K L M N O P Q R S T U V W X Y Z 
A
Melechovsky J., Mehrish A., Sisman B., Herremans D..  2024.  Accent Conversion in Text-To-Speech Using Multi-Level VAE and Adversarial Training. Proc. of IEEE Tencon, Singapore.
Melechovsky J., Mehrish A., Sisman B., Herremans D..  2024.  Accented Text-to-Speech Synthesis with a Conditional Variational Autoencoder. Proc. of IEEE Tencon, Singapore.
BT B, Aslim E.J, Ng YShu Lynn, Kuo TLi Chuen, Chen JShihang, Herremans D., Ng LGuat, Chen J.M..  2020.  Acoustic prediction of flowrate: varying liquid jet stream onto a free surface. IEEE International Conference on Signal Processing and Communications (SPCOM). PDF icon preprint flow.pdf (1.01 MB)
Herremans D.  2021.  aiSTROM - A roadmap for developing a successful AI strategy. IEEE Access.
Herremans D., Roy A..  2026.  Aligning Generative Music AI with Human Preferences: Methods and Challenges. Proceedings of AAAI, senior member track. PDF icon 2511.15038v1.pdf (417.24 KB)
Melechovsky J..  2025.  Analysis and Synthesis of Audio with AI: from Neurological Disease to Accented Speech and Music. PDF icon thesis_Jan.pdf (26.4 MB)
Husain J.A., Herremans D..  2026.  APEX: Large-scale Multi-task Aesthetic-Informed Popularity Prediction for AI-Generated Music. arXiv:2605.03395. PDF icon 2605.03395v1 (1).pdf (292.45 KB)
Kang J., Herremans D..  2025.  Are we there yet? A brief survey of Music Emotion Prediction Datasets, Models and Outstanding Challenges IEEE Transactions on Affective Computing. PDF icon 2406.08809v1.pdf (156.19 KB)
BT B, Hee H.I., Teoh O.H., Lee K.P., Kapoor S., Herremans D., Chen J.M..  2020.  Asthmatic versus healthy child classification based on cough and vocalised /a:/ sounds. The Journal of the Acoustical Society of America (JASA). 148, EL253
T. Phuong HThi, BT B, Roig G., Herremans D..  2021.  AttendAffectNet – Emotion Prediction of Movie Viewers Using Multimodal Fusion with Self-attention. Sensors. Special issue on Intelligent Sensors: Sensor Based Multi-Modal Emotion Recognition. PDF icon sensors-21-08356.pdf (1.03 MB)
T. Phuong HThi, BT B, Herremans D., Roig G..  2021.  AttendAffectNet: Self-Attention based Networks for Predicting Affective Responses from Movies. Proceedings of the International Conference on Pattern Recognition (ICPR2020). PDF icon 2010.11188.pdf (7.07 MB)
C
Herremans D., Sörensen K., Martens D.  2015.  Classification and generation of composer-specific music using global feature models and variable neighborhood search. Computer Music Journal. 39(3):91.PDF icon papercmj-dh_preprint.pdf (637.63 KB)
Lanzendörfer L.A., Lu T., Perraudin N., Herremans D., Wattenhofer R..  2025.  Coarse-to-Fine Text-to-Music Latent Diffusion. Proceedings of ICASSP.
Lanzendörfer L.A., Lu T., Perraudin N., Herremans D., Wattenhofer R..  2024.  Coarse-to-Fine Text-to-Music Latent Diffusion. Audio Imagination: NeurIPS 2024 Workshop.
Herremans D.  2015.  Compose ≡ compute. 4OR. 13:335–336.
Herremans D..  2014.  Compose=Compute - Computer Generation And Classification Of Music Through Operations Research Methods. PhD Thesis, University of Antwerp. :250.
Herremans D., Martens D, Sörensen K., Meredith D..  2015.  Composer Classification Models for Music-Theory Building. Computational Music Analysis. PDF icon Chapter_HerremansEtAl_preprint.pdf (475.26 KB)
Herremans D., Sörensen K..  2012.  Composing counterpoint musical scores with variable neighborhood search. Annual Conference of the Belgian Operation Research Society (ORBEL26). PDF icon orbel26abs_vnsforcp.pdf (116.85 KB)
Herremans D., Sörensen K..  2013.  Composing Fifth Species Counterpoint Music With A Variable Neighborhood Search Algorithm. Expert Systems with Applications. 40PDF icon paper_preprint_cp5.pdf (405.75 KB)
Herremans D., Sörensen K..  2012.  Composing Fifth Species Counterpoint Music With Variable Neighborhood Search. PDF icon wp_cp5.pdf (508 KB)
Herremans D., Sörensen K..  2012.  Composing first species counterpoint musical scores with a variable neighbourhood search algorithm. Journal of Mathematics and the Arts. 6:169-189.
Clarke C.J., Chowdhury J., BT B, Priyadarshinee P., Lim C.M.Ying, I. Tan FXing, Herremans D., Chen J.M..  2022.  Computationally Efficient Physics Approximating Neural Networks for Highly Nonlinear Maps. 2022 International Conference on Research in Adaptive and Convergent Systems.
Makris D., Guo Z, Kaliakatsos-Papakostas N., Herremans D..  2022.  Conditional Drums Generation using Compound Word Representations. EvoMUSART (EVO*) - Lecture Notes in Computer Science. PDF icon 2202.04464.pdf (525.36 KB)
Ong J., Herremans D..  2023.  Constructing Time-Series Momentum Portfolios with Deep Multi-Task Learning. Expert Systems with Applications. 230(120587)PDF icon 2306.13661.pdf (707.95 KB)
D
Herremans D., Martens D, Sörensen K..  2014.  Dance hit song prediction. Journal of New music Research. 43:302.PDF icon wp_hit.pdf (689.07 KB)
Herremans D., Martens D, Sörensen K..  2013.  Dance Hit Song Science. International Workshop on Music and Machine Learning. PDF icon abstract_preprint_MML2013_DH.pdf (194.82 KB)
Melechovsky J., Mehrish A., Sisman B., Herremans D..  2024.  DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech. Audio Imagination: NeurIPS 2024 Workshop.
Pham Q-H.  2020.  Data-driven 3D Scene Understanding. PhD
Nahar F., Agres K., BT B, Herremans D..  2020.  A dataset and classification model for Malay, Hindi, Tamil and Chinese music. 13th Workshop on music and machine learning (MML) as part of ECML/PKDD. PDF icon 2009.04459.pdf (234.8 KB)
T BB, Hee HIng, Kapoor S, Teoh OHoe, Teng SShin, Lee KPin, Herremans D, Chen JMing.  2021.  Deep Neural Network Based Respiratory Pathology Classification Using Cough Sounds. Sensors. 21(16):5555.PDF icon 2106.12174.pdf (6.52 MB)
Ong J., Herremans D..  2024.  DeepUnifiedMom: Unified Time-series Momentum Portfolio Construction via Multi-Task Learning with Multi-Gate Mixture of Experts. arXiv:2406.08742. PDF icon 2406.08742v1.pdf (1.06 MB)
Song M., Liu R., Wang X, Jiang Y, Xie P, Huang F, Zhou J, Herremans D., Poria S..  2025.  Demystifying deep search: a holistic evaluation with hint-free multi-hop questions and factorised metrics. arXiv:2510.05137.
Melechovsky J., Novotny M., Tykalova T., Klempir J., Herremans D., Rusz J..  2026.  Development of Interpretable Deep Learning-based Segmentation Algorithm for Automated Assessment of Oral Diadochokinesis in Progressive Neurological Diseases. Journal of Speech, Language, and Hearing Research.
Hee H.I., BT B, Karunakaran A., Herremans D., Teoh O.H., Lee K.P., Teng S.S., Lui S., Chen J.M..  2019.  Development of Machine Learning for asthmatic and healthy voluntary cough - a proof of concept study. Applied Sciences. 9(14)PDF icon applsci-09-02833.pdf (2.06 MB)
Cheuk K.W., Sawata R, Uesaka T, Murata N, Takahashi N, Takahashi S, Herremans D., Mitsufuji Y.  2023.  DiffRoll: Diffusion-based Generative Music Transcription with Unsupervised Pretraining Capability. ICASSP. PDF icon diffroll.pdf (2.2 MB)
Puri G., Socklingam N., Herremans D..  2026.  Digital Lifelong Learning in the Age of AI: Trends and Insights.
Wang K., Herremans D..  2024.  DisfluencySpeech -- Single-Speaker Conversational Speech Dataset with Paralanguage. Proc. of IEEE Tencon, Singapore.
Guo Z, Kang J., Herremans D..  2023.  A Domain-Knowledge-Inspired Music Embedding Space and a Novel Attention Mechanism for Symbolic Music Modeling. Proceedings of the 37th AAAI Conference on Artificial Intelligence. PDF icon 2212.00973.pdf (1.74 MB)
Lee-Leon A., Yuen C., Herremans D..  2019.  Doppler Invariant Demodulation for Shallow Water Acoustic Communications Using Deep Belief Networks. 16th IEEE Asia Pacific Wireless Communications Symposium (APWCS). PDF icon 1909.02850.pdf (790.54 KB)
Levers O.D, Herremans D., Dipankar A., Blessing L..  2022.  Downscaling using Deep Convolutional Autoencoders, a case study for South East Asia. Egusphere preprint. PDF icon egusphere-2022-234.pdf (8.99 MB)
Herremans D..  2010.  Drupal 6: Ultimate Community Site Guide.
E
Agres K, Bigo L, Herremans D, Conklin D.  2015.  The effect of repetitive structure on enjoyment and altered states in uplifting trance music. 2nd International Conference on Music and Consciousness (MUSCON 2), Brighton. PDF icon AgresEtAl_muscon.pdf (12.47 KB)
Agres K., Bigo L., Herremans D., Conklin D..  2016.  The Effect of Repetitive Structure on Enjoyment in Uplifting Trance Music. 14th International Conference for Music Perception and Cognition (ICMPC). :280-282.PDF icon preprint_trance.pdf (139.27 KB)
Cheuk K.W., Luo Y.J., Benetos E., Herremans D..  2021.  The Effect of Spectrogram Reconstructions on Automatic Music Transcription:An Alternative Approach to Improve Transcription Accuracy. Proceedings of the International Conference on Pattern Recognition (ICPR2020). PDF icon 2010.09969.pdf (3.46 MB)
Herremans D., Chuan C.-H..  2019.  The emergence of deep learning: new opportunities for music and audio technologies. Neural Computing and Applications. PDF icon main_preprint.pdf (102.16 KB)
Bhandari K., Roy A., Colton S., Herremans D..  2026.  Emerging AI Technologies for Music: Towards Controllable, Collaborative, and Creative Systems. Proceedings of Machine Learning Research, PMLR 303:1-5, 2026. PDF icon bhandari26a.pdf (161.47 KB)
Pham Q-H, Herremans D., Roig G..  2022.  EmoMV: Affective Music-Video Correspondence Learning Datasets for Classification and Retrieval. Information Fusion. PDF icon SSRN-id4189323.pdf (2.01 MB)
Tripathi A., Patle V., Jain A., Pundir A., Menon S., A. Singh K, Herremans D..  2025.  End-to-End Text-to-SQL with Dataset Selection: Leveraging LLMs for Adaptive Query Generation. Proceedings of IJCNN, Rome, Italy.
Wang K., Tekler Z., Cheah L., Herremans D., Blessing L..  2021.  Evaluating the Effectiveness of an Augmented Reality Game Promoting Environmental Action. Sustainability. 13(24):13912.PDF icon sustainability-13-13912.pdf (16.23 MB)
Guo R, Herremans D..  2025.  An exploration of controllability in symbolic music infilling. IEEE Access.
F
Herremans D., Sörensen K., Conklin D..  2013.  First species counterpoint generation with VNS and vertical viewpoints. Digital Music Research Network (DMNR+8). PDF icon dnmr8_dh_dc.pdf (147.73 KB)
Herremans D., Sörensen K., Conklin D..  2014.  First species counterpoint generation with VNS and vertical viewpoints. Annual Conference of the Belgian Operation Research Society (ORBEL28). PDF icon orbel28_dh.pdf (216.63 KB)
Herremans D., Low K.W..  2025.  Forecasting Bitcoin Volatility Spikes from Whale Transactions and Cryptoquant Data Using Synthesizer Transformer Models. IEEE Access. 13:117788-117807.PDF icon SSRN-id4247684.pdf (5.05 MB)
Zhong C., Wu S., Yu J., Wu W., Herremans D., Zhang K..  2026.  Formalizing Semi-Structured Interviews for Design Requirement Discovery: A Multi-Agent Framework for Clarification and Empathic Interaction. Available at SSRN 7199726.
Chuan C.-H., Agres K., Herremans D..  2018.  From Context to Concept: Exploring Semantic Relationships in Music with Word2Vec. Neural Computing and Applications. PDF icon paper.pdf (1.64 MB)
Kadir N., Herremans D..  2026.  A Functional Taxonomy of Intelligent Systems in Education.
Herremans D., Chuan C.-H., Chew E..  2017.  A Functional Taxonomy of Music Generation Systems. ACM Computing Surveys. 50(5):30.PDF icon music_generation_survey_dh_preprint.pdf (349.15 KB)
Herremans D., Sörensen K..  2013.  FuX, an Android app that generates counterpoint. IEEE Symposium on Computational Intelligence for Creativity and Affective Computing (CICAC). :48-55.PDF icon wp_fux.pdf (486.27 KB)
G
Chow D., Herremans D..  2024.  Gamification and skills tree. Trends and Foresight Report on Cyber-Physical Learning.
BT B, Hee H.I., Ming C., Lin Y., Priyadarshinee P., Clarke C.J., Herremans D., Chen J.M..  2022.  A Gaussian mixture classifier model to differentiate respiratory symptoms using phonated /ɑː/ sounds. The 18th Australasian International Conference on Speech Science and Technology (SST). PDF icon ahsounds.pdf (1018.01 KB)
Balliauw M., Herremans D., D. Cuervo P, Sörensen K..  2015.  Generating Fingerings for Polyphonic Piano Music with a Tabu Search Algorithm. Mathematics and Computation in Music. 9110:149-160.PDF icon paper_mcm_preprint.pdf (405.73 KB)
Cunha N., A. S, Herremans D.  2017.  Generating guitar solos by integer programming. Journal of the Operational Research Society. :971-985.PDF icon preprint_guitar_solo_generation_dh.pdf (772.59 KB)
Makris D., Agres K., Herremans D..  2021.  Generating Lead Sheets with Affect: A Novel Conditional seq2seq Framework. Proceedings of the International Joint Conference on Neural Networks (IJCNN). PDF icon 2104.13056.pdf (857.78 KB)
Herremans D., Weisser S., Sörensen K., Conklin D..  2015.  Generating music with an optimization algorithm using a Markov based objective function. ORBEL29, Belgian Conference on Operations Research. PDF icon orbel29abs.pdf (138.67 KB)
Herremans D., Weisser S., Sörensen K., Conklin D..  2015.  Generating structured music for bagana using quality metrics based on Markov models. Expert Systems With Applications. 42 (21)(21):424–7435.PDF icon paper-bagana.pdf (1.73 MB)
Herremans D., Weisser S., Sörensen K., Conklin D..  2014.  Generating structured music using quality metrics based on Markov models. PDF icon wp_bagana.pdf (1.7 MB)
A. Putri M, Saide S., D. Riau K, Herremans D..  2026.  Generative AI in Education for SDG 4: Insights from Indonesia and Kazakhstan. Proceedings of the Pacific Asia Conference on Information Systems (PACIS)..
Tan HHao, Luo Y.J., Herremans D..  2020.  Generative Modelling for Controllable Audio Synthesis of Expressive Piano Performance. Workshop on Machine Learning for Music Discover (ML4MD) as part of ICML. PDF icon 2006.09833.pdf (2.81 MB)
H
Agres K., Herremans D., Bigo L., Conklin D..  2017.  Harmonic Structure Predicts the Enjoyment of Uplifting Trance Music. Frontiers in Psychology, Cognitive Science. 7(1999)PDF icon agres16ut.pdf (1.15 MB)
Turian J, Shier J, Khan HRaj, Raj B, Schuller BW, Steinmetz CJ, Malloy C, Tzanetakis G, Velarde G, McNally K et al..  2022.  HEAR 2021: Holistic Evaluation of Audio Representations. Proceedings of Machine Learning Research (PMLR): NeurIPS 2021 Competition Track. PDF icon 2203.03022.pdf (406.58 KB)
Tripathi A., A. Singh K, Surya R., Gupta A., Veikho S.L., Herremans D., Bisane S..  2025.  HHNAS-AM: Hierarchical Hybrid Neural Architecture Search using Adaptive Mutation Policies. arXiv:2508.14946.
Guo Z, Makris D., Herremans D..  2021.  Hierarchical Recurrent Neural Networks for Conditional Melody Generation with Long-term Structure. Proceedings of the International Joint Conference on Neural Networks (IJCNN). PDF icon 2102.09794.pdf (1015.73 KB)
Herremans D., Bergmans T..  2017.  Hit Song Prediction Based on Early Adopter Data and Audio Features. The 18th International Society for Music Information Retrieval Conference (ISMIR) - Late Breaking Demo. PDF icon paper_preprint_hit.pdf (221.73 KB)
Lee-Leon A., Yuen C., Herremans D..  2019.  A Hybrid Fuzzy Logic-Neural Network Approach For Multi-path Separation Of Underwater Acoustic Signals. 89th IEEE Vehicular Technology Conference. PDF icon fuzzy logic.pdf (1.66 MB)
M
Kaliakatsos-Papakostas N., Bastas G., Makris D., Herremans D., Katsouros V., Maragos P..  2022.  A Machine Learning Approach for MIDI to Guitar Tablature Conversion. Sound and Music Computing Conference (SMC). PDF icon 25.pdf (528.42 KB)
Sturm B., Ben-Tal O., Monaghan U., Collins N., Herremans D., Chew E., Hadjeres G., Deruty E., Pachet F..  2019.  Machine Learning Research that Matters for Music Creation: A Case Study. Journal of New Music Research. 48(1):36-55.PDF icon concert_paper_preprint.pdf (1.6 MB)
Herremans D., Weisser S., Sörensen K., Conklin D..  2014.  Markov Based Quality Metrics For Generating Structured Music With Optimization Techniques. Digital Music Research Network (DMNR+9). PDF icon dmrn9_dh.pdf (133.29 KB)
Song M., Pala T.D, Jin W., Zadeh A., Li C., Herremans D., Poria S..  2026.  Measuring and Mitigating Rapport Bias of Large Language Models under Multi-Agent Social Interactions. Proceedings of ICLR.
Lu T., Geist C-M, Melechovsky J., Roy A., Herremans D..  2026.  MelodySim: Measuring Melody-aware Music Similarity for Plagiarism Detection. IEEE Tencon.
Koh E., Cheuk K.W., Heung K.Y., Agres K., Herremans D..  2023.  MERP: A Music Dataset with Emotion Ratings and Raters’ Profile Information. Sensors - Intelligent Sensors. 23(1)PDF icon sensors-23-00382 (2).pdf (1.21 MB)
Guo R, Herremans D, Magnusson T.  2019.  Midi Miner – A Python library for tonal tension and track classification. ISMIR - Late Breaking Demo. PDF icon midi_miner.pdf (83.7 KB)
Melechovsky J., Roy A., Herremans D..  2024.  MidiCaps — A large-scale MIDI dataset with text captions. ISMIR. PDF icon 2406.02255v1.pdf (699.83 KB)
Agus N., Anderson H., Chen J.M., Lui S., Herremans D..  2018.  Minimally Simple Binaural Room Modelling Using a Single Feedback Delay Network. Journal of the Audio Engineering Society. 66(10):791-807.PDF icon angus_jaes_preprint.pdf (6.39 MB)
Chopra A., Roy A., Herremans D..  2024.  MIRFLEX: Music Information Retrieval Feature Library for Extraction. ISMIR, Late Breaking Demos. PDF icon 2411.00469v1.pdf (89.86 KB)
Herremans D., Chuan C.-H..  2017.  Modeling Musical Context with Word2vec. First International Workshop On Deep Learning and Music. 1:11-18.PDF icon herremans2017work2vec.pdf (745.8 KB)

Pages