PUBLICATIONS

Members of CASPR have been involved in the research documented in the following scientific publications:

2027

Journal Papers

  1. Ioannis Stylianou, Achintya Kumar Sarkar, Nauman Dawalatabad, James Glass, and Zheng-Hua Tan. “LibriVAD: a scalable open dataset with deep learning benchmarks for robust voice activity detection”. In: Computer Speech & Language 102 (2027), p. 102043. (Accepted).

2026

Journal Papers

  1. Eleftheria Lydaki, Zheng-Hua Tan, Jesper Jensen, and Meng Guo. “Deep feedback cancellation in hearing aids”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 34 (2026), pp. 3159–3173.
  2. Filippo Villani, Wai-Yip Geoffrey Chan, Zheng-Hua Tan, Jan Østergaard, and Jesper Jensen. “Near-end listening enhancement via combined noise-dependent spectro-temporal energy reallocation”. In: IEEE Access 14 (2026), pp. 43795–43812.
  3. Gustav Wagner Zakarias, Lars Kai Hansen, and Zheng-Hua Tan. “BiSSL: enhancing the alignment between self-supervised pretraining and downstream fine-tuning via bilevel optimization”. In: Transactions on Machine Learning Research 2 (2026). This publication was awarded the J2C certification and presented at the 43rd International Conference on Machine Learning (ICML) in Seoul, Korea.
  4. Haolan Wang, Wai-Yip Geoffrey Chan, and Jesper Jensen. “TSIP-Net: no-reference speech intelligibility prediction in the presence of competing speech”. In: Speech Communication 179 (2026), p. 103377.
  5. Ioannis Stylianou, Jon Francombe, Pablo Martinez-Nuevo, Sven Ewan Shepstone, and Zheng-Hua Tan. “One prompt, many sounds: modeling listener variability in LLM-based equalization”. In: IEEE Journal of Selected Topics in Signal Processing (2026).
  6. James Brooks-Park, Søren Bech, Jan Østergaard, and Steven van de Par. “Room compensation for loudspeaker reproduction using a supporting source”. In: The Journal of the Acoustical Society of America 159.4 (2026), pp. 3006–3017.
  7. Jesper Brunnström, Martin Bo Møller, Jan Østergaard, Shoichi Koyama, Toon van Waterschoot, and Marc Moonen. “Sound field estimation with moving microphones using kernel ridge regression”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 34 (2026), pp. 4084–4097.
  8. Jesper Brunnström, Martin Bo Møller, Jan Østergaard, Shoichi Koyama, Toon van Waterschoot, and Marc Moonen. “Time-domain sound field estimation using kernel ridge regression”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 34 (2026), pp. 1243–1258.
  9. José Miguel Cadavid Tobón, Martin Bo Møller, Toon van Waterschoot, Søren Bech, and Jan Østergaard. “Reducing computational complexity in adaptive sound zones with online room impulse response estimation”. In: The Journal of the Acoustical Society of America 159.6 (2026), pp. 5729–5744.
  10. Nikolai Lund Kühne, Jesper Jensen, Jan Østergaard, and Zheng-Hua Tan. “MambAttention: mamba with multi-head attention for generalizable single-channel speech enhancement”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 34 (2026), pp. 820–833.
  11. Payam Shahsavari Baboukani, Emina Alickovic, Jan Østergaard, and Kasper Eskelund. “Parietal alpha-band connectivity tracks listening effort in hearing-aid users under competing speech and noise”. In: European Journal of Neuroscience 63.6 (2026), p. e70466.
  12. Pia Nancy Porysek Moreta, Søren Bech, Jon Francombe, Jan Østergaard, and Steven van de Par. “Continuous and overall attribute evaluation of spatial audio reproduction systems with spatially dynamic content”. In: Journal of the Audio Engineering Society 74.4 (2026), pp. 199–211.

Conference Papers

  1. Kevin Wilkinghoff, Gordon Wichern, Jonathan Le Roux, and Zheng-Hua Tan. “Mind the gap: detecting cluster exits for robust local density-based score normalization in anomalous sound detection”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Sydney, Australia, September 27–October 1, 2026.
  2. Niels Overby, Svend Feldt, Zheng-Hua Tan, and Jan Østergaard. “Speech separation: from simulated spaces to realistic situations”. In: Proceedings of the 19th International Workshop on Acoustic Signal Enhancement (IWAENC). Cremona, Italy, September 7–10, 2026.
  3. Panos Apostolidis, Svend Feldt, Zheng-Hua Tan, Jan Østergaard, and Jesper Jensen. “Listen first: output-based multi-microphone speech enhancement”. In: Proceedings of the 19th International Workshop on Acoustic Signal Enhancement (IWAENC). Cremona, Italy, September 7–10, 2026.
  4. Asjid Tanveer, Jesper Jensen, Emina Alickovic, Zheng-Hua Tan, and Jan Østergaard. “Subject-independent auditory attention decoding from mobile EEG via contrastive learning”. In: Proceedings of the 34th European Signal Processing Conference (EUSIPCO). Bruges, Belgium, August 31–September 4, 2026.
  5. Kevin Wilkinghoff, Keisuke Imoto, and Zheng-Hua Tan. “How much does machine identity matter in anomalous sound detection at test time?” In: Proceedings of the 34th European Signal Processing Conference (EUSIPCO). Bruges, Belgium, August 31–September 4, 2026.
  6. Sif Bjerre Lindby, Jesper Jensen, Zheng-Hua Tan, and Jan Østergaard. “Electroencephalography phase synchrony differentiates linguistic violations in noisy speech”. In: Proceedings of the 48th Annual International Conference of the IEEE Engineering in Medicine and Biology Society (EMBC). Toronto, Canada, July 26–30, 2026, pp. 727–733.
  7. Sif Bjerre Lindby, Jesper Jensen, Zheng-Hua Tan, and Jan Østergaard. “Investigating the effect of sentence-level syntactic structure on information loss in the human auditory system”. In: Proceedings of the 51st Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2026, pp. 16577–16581.
  8. Nikolai Lund Kühne, Jesper Jensen, Jan Østergaard, and Zheng-Hua Tan. “Exploring resolution-wise shared attention in hybrid Mamba-U-Nets for improved cross-corpus speech enhancement”. In: Proceedings of the 51st Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2026, pp. 16327–16331.
  9. Mojtaba Farmani, Svend Feldt, and Jesper Jensen. “Beamforming using virtual microphones for hearing aid applications”. In: Proceedings of the 51st Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2026, pp. 15552–15556.
  10. Peter Asbjørn Leer Bysted, Svend Feldt, Zheng-Hua Tan, Jan Østergaard, and Jesper Jensen. “Ranking the impact of contextual specialization in neural speech enhancement”. In: Proceedings of the 51st Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2026, pp. 17432–17436.
  11. Kevin Wilkinghoff and Zheng-Hua Tan. “DSPAST: disentangled representations for spatial audio reasoning with large language models”. In: Proceedings of the 51st Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2026, pp. 14747–14751.
  12. Kevin Wilkinghoff, Alessia Cornaggia-Urrigshardt, and Zheng-Hua Tan. “Quantization-based score calibration for few-shot keyword spotting with dynamic time warping in noisy environments”. In: Proceedings of the 51st Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2026, pp. 18787–18791.
  13. Lasse Borgholt, Jakob Havtorn, Christian Igel, Lars Maaløe, and Zheng-Hua Tan. “A text-to-text alignment algorithm for better evaluation of modern speech recognition systems”. In: Proceedings of the 51st Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2026, pp. 3476–3480.

2025

Journal Papers

  1. Vasudha Sathyapriyan, Michael Syskind Pedersen, Mike Brookes, Jan Østergaard, Patrick A. Naylor, and Jesper Jensen. “Head-steered channel selection for hearing aid applications using remote microphones”. In: IEEE Access 13 (2025), pp. 201478–201491.
  2. Vasudha Sathyapriyan, Michael Syskind Pedersen, Mike Brookes, Jan Østergaard, Patrick A. Naylor, and Jesper Jensen. “Binary estimator selection methods for hearing aids with a remote microphone”. In: IEEE Access 13 (2025), pp. 159610–159627.
  3. Kaspar Müller, Markus Buck, Simon Doclo, Jan Østergaard, and Tobias Wolff. “A steered response power method for sound source localization with generic acoustic models”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 33 (2025), pp. 4004–4019.
  4. Yuying Xie and Zheng-Hua Tan. “A survey of deep learning for complex speech spectrograms”. In: Speech Communication 175 (2025), p. 103319.
  5. Mohamad Al Ahdab, Zheng-Hua Tan, and John-Josef Leth. “Distributions and direct parametrization for stable stochastic state-space models”. In: IEEE Control Systems Letters 9 (2025), pp. 444–449.
  6. Asjid Tanveer, Jesper Jensen, Zheng-Hua Tan, and Jan Østergaard. “Single-microphone deep envelope separation based auditory attention decoding for competing speech and music”. In: Journal of Neural Engineering 22.3 (2025), p. 036006.
  7. José Miguel Cadavid Tobón, Martin Bo Møller, Toon van Waterschoot, Søren Bech, and Jan Østergaard. “Effect of reduced information in the performance of low-frequency sound zones”. In: Journal of the Audio Engineering Society 73.7-8 (2025), pp. 429–445.
  8. Iván López-Espejo, Eros Roselló, Amin Edraki, and Jesper Jensen. “Noise-robust hearing aid voice control”. In: IEEE Signal Processing Letters 32 (2025), pp. 241–245.
  9. Yiming Zhang, Xuenan Xu, Ruoyi Du, Haohe Liu, Yuan Dong, Zheng-Hua Tan, Wenwu Wang, and Zhanyu Ma. “Zero-shot audio captioning using soft and hard prompts”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 33 (2025), pp. 2045–2058.
  10. Peter Asbjørn Leer Bysted, Jesper Jensen, Laurel Carney, Zheng-Hua Tan, Jan Østergaard, and Lars Bramsløw. “Hearing-loss compensation using deep neural networks: a framework and results from a listening test”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 33 (2025), pp. 828–841.

Conference Papers

  1. Kevin Wilkinghoff, Takuya Fujimura, Keisuke Imoto, Jonathan Le Roux, Zheng-Hua Tan, and Tomoki Toda. “Handling domain shifts for anomalous sound detection: a review of DCASE-related work”. In: Proceedings of the Detection and Classification of Acoustic Scenes and Events Workshop (DCASE). Barcelona, Spain, October 29–31, 2025.
  2. Asjid Tanveer, Jesper Jensen, Zheng-Hua Tan, and Jan Østergaard. “Multivariate phase synchrony and auditory attention decoding for speech and music”. In: Proceedings of the International Conference on Biomedical Engineering and Bioinformatics (ICBEB). Prague, Czech Republic, September 19–21, 2025.
  3. Jesper Brunnström, Martin Bo Møller, Jan Østergaard, Toon van Waterschoot, Marc Moonen, and Filip Elvander. “Spatial covariance estimation for sound field reproduction using kernel ridge regression”. In: Proceedings of the 33rd European Signal Processing Conference (EUSIPCO). Palermo, Italy, September 8–12, 2025.
  4. Filippo Villani, Wai-Yip Geoffrey Chan, Zheng-Hua Tan, Jan Østergaard, and Jesper Jensen. “Analysis and extension of a near-end listening enhancement method based on long-term fractile noise statistics”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Rotterdam, The Netherlands, August 17–21, 2025, pp. 783–787.
  5. Aymen Bashir, Haolan Wang, Amin Edraki, Wai-Yip Geoffrey Chan, and Jesper Jensen. “Intelligibility prediction for time-modified speech signals using spectro-temporal modulation features”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Rotterdam, The Netherlands, August 17–21, 2025, pp. 5478–5482.
  6. Nikolai Lund Kühne, Jan Østergaard, Jesper Jensen, and Zheng-Hua Tan. “xLSTM-SENet: xLSTM for single-channel speech enhancement”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Rotterdam, The Netherlands, August 17–21, 2025, pp. 5148–5152.
  7. Eleftheria Lydaki, Zheng-Hua Tan, Jesper Jensen, and Meng Guo. “Deep feedback cancellation for hearing aids with improved system stability and sound quality”. In: Proceedings of the 50th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Hyderabad, India, April 6–11, 2025.
  8. Nikolai Lund Kühne, Astrid Holm Filtenborg Kitchen, Marie Saugstrup Jensen, Mikkel Sebastian Lundsgaard Brøndt, Martin Gonzalez, Christophe Biscio, and Zheng-Hua Tan. “Detecting and defending against adversarial attacks on automatic speech recognition via diffusion models”. In: Proceedings of the 50th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Hyderabad, India, April 6–11, 2025.
  9. Mohamad Al Ahdab, Zheng-Hua Tan, and John-Josef Leth. “Optimal sensor scheduling for continuous-discrete Kalman filtering with auxiliary dynamics”. In: Proceedings of the 42nd International Conference on Machine Learning (ICML). Vancouver, Canada, July 13–19, 2025.
  10. Sarthak Yadav, Sergios Theodoridis, and Zheng-Hua Tan. “AxLSTMs: learning self-supervised audio representations with xLSTMs”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Rotterdam, The Netherlands, August 17–21, 2025.
  11. Sarthak Yadav, Sergios Theodoridis, and Zheng-Hua Tan. “AudioMAE++: learning better masked audio representations with SwiGLU FFNs”. In: Proceedings of the IEEE International Workshop on Machine Learning for Signal Processing (MLSP). Istanbul, Turkey, August 31–September 3, 2025.
  12. Holger Severin Bovbjerg, Jan Østergaard, Jesper Jensen, Shinji Watanabe, and Zheng-Hua Tan. “Learning robust spatial representations from binaural audio through feature distillation”. In: Proceedings of the IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA). Granlibakken Tahoe, Tahoe City, CA, USA, October 12–15, 2025.
  13. Jan Østergaard, Sangeeth Geetha Jayaprakash, and Rodrigo Ordoñez. “Rate-distortion under neural tracking of speech: a directed redundancy approach”. In: Proceedings of the IEEE Data Compression Conference (DCC). Snowbird, UT, USA, March 18–21, 2025.

2024

Journal Papers

  1. Mohammad Bokaei, Jesper Jensen, Simon Doclo, and Jan Østergaard. “Low-latency deep analog speech transmission using joint source channel coding”. In: IEEE Journal of Selected Topics in Signal Processing 18.8 (2024), pp. 1401–1413.
  2. Pia Nancy Porysek Moreta, Søren Bech, Jon Francombe, Jan Østergaard, and Steven van de Par. “Identifying principal attributes for evaluating audio quality of reproduction systems with spatially dynamic program material”. In: Journal of the Audio Engineering Society 72.9 (2024), pp. 552–570.
  3. Adèle Simon, Søren Bech, Gérard Loquet, and Jan Østergaard. “Cortical linear encoding and decoding of sounds: similarities and differences between naturalistic speech and music listening”. In: European Journal of Neuroscience 59.8 (2024), pp. 2059–2074.
  4. Yousef Mohammadi, Carina Graversen, José Biurrun Manresa, Jan Østergaard, and Ole Kæseler Andersen. “Effects of background noise and linguistic violations on frontal theta oscillations during effortful listening”. In: Ear and Hearing 45.3 (2024), pp. 721–729.
  5. Andreas Jonas Fuglsig, Zheng-Hua Tan, Lars Søndergaard Bertelsen, Jesper Jensen, Jens Christian Lindof, and Jan Østergaard. “Joint far- and near-end speech and listening enhancement with minimum processing”. In: IEEE Access 12 (2024), pp. 119983–120004.
  6. José Miguel Cadavid Tobón, Martin Bo Møller, Christian Sejer Pedersen, Søren Bech, Toon van Waterschoot, and Jan Østergaard. “Performance of low-frequency sound zones with very fast room impulse response measurements”. In: The Journal of the Acoustical Society of America 155.1 (2024), pp. 757–768.
  7. Yiming Zhang, Ruoyi Du, Zheng-Hua Tan, Wenwu Wang, and Zhanyu Ma. “Generating accurate and diverse audio captions through variational autoencoder framework”. In: IEEE Signal Processing Letters 31 (2024), pp. 2520–2524.
  8. Peter Asbjørn Leer Bysted, Jesper Jensen, Zheng-Hua Tan, Jan Østergaard, and Lars Bramsløw. “How to train your ears: auditory-model emulation for large-dynamic-range inputs and mild-to-severe hearing losses”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 32 (2024), pp. 2006–2020.
  9. Martin Bo Møller, Jorge Martinez, and Jan Østergaard. “Reduced complexity for sound zones with subband block adaptive filters and a loudspeaker line array”. In: The Journal of the Acoustical Society of America 155.4 (2024), pp. 2314–2326.
  10. Philippe Gonzalez, Zheng-Hua Tan, Jan Østergaard, Jesper Jensen, Tommy Sonne Alstrøm, and Tobias May. “Investigating the design space of diffusion models for speech enhancement”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 32 (2024), pp. 4486–4500.
  11. Philippe Gonzalez, Zheng-Hua Tan, Jan Østergaard, Jesper Jensen, Tommy Sonne Alstrøm, and Tobias May. “The effect of training dataset size on discriminative and diffusion-based speech enhancement systems”. In: IEEE Signal Processing Letters 31 (2024), pp. 2225–2229.
  12. Mathias Bach Pedersen, Zheng-Hua Tan, Søren Holdt Jensen, and Jesper Rindom Jensen. “Data-driven non-intrusive speech intelligibility prediction using speech presence probability”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 32 (2024), pp. 55–67.

Conference Papers

  1. Jesper Brunnström, Martin Bo Møller, Jan Østergaard, and Marc Moonen. “Bayesian sound field estimation using uncertain data”. In: Proceedings of the 18th International Workshop on Acoustic Signal Enhancement (IWAENC). Aalborg, Denmark, September 9–12, 2024.
  2. Haolan Wang, Jesper Jensen, Iván López-Espejo, and Wai-Yip Geoffrey Chan. “No-reference speech intelligibility prediction leveraging a noisy-speech ASR pre-trained model”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Kos, Greece, September 1–5, 2024.
  3. Mohammad Bokaei, Jesper Jensen, Simon Doclo, and Jan Østergaard. “Channel-configurable deep wireless speech transmission”. In: Proceedings of the IEEE Wireless Communications and Networking Conference (WCNC). Dubai, United Arab Emirates, April 21–24, 2024.
  4. Mohammad Bokaei, Jesper Jensen, Simon Doclo, and Jan Østergaard. “Deep low-latency joint speech transmission and enhancement over a Gaussian channel”. In: Proceedings of the 49th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Seoul, Korea, April 14–19, 2024.
  5. Jan Østergaard. “Directed redundancy in time series”. In: Proceedings of the IEEE International Symposium on Information Theory (ISIT). Athens, Greece, July 7–12, 2024.
  6. José Miguel Cadavid Tobón, Martin Bo Møller, Toon van Waterschoot, Søren Bech, and Jan Østergaard. “Spatial sampling versus acquisition time of room impulse responses for low-frequency sound zones”. In: Proceedings of the 156th Audio Engineering Society (AES) Convention. Madrid, Spain, June 15–17, 2024.
  7. Jan Østergaard and Payam Shahsavari Baboukani. “Synergy and redundancy dominated effects in time series via transfer entropy decompositions”. In: Proceedings of the IEEE International Symposium on Information Theory Workshop on Neural Information Theory (NeurIT). Athens, Greece, July 7, 2024.
  8. Filippo Villani, Wai-Yip Geoffrey Chan, Zheng-Hua Tan, Jan Østergaard, and Jesper Jensen. “Near-end listening enhancement using a noise-robust linear time-invariant filter”. In: Proceedings of the 18th International Workshop on Acoustic Signal Enhancement (IWAENC). Aalborg, Denmark, September 9–12, 2024.
  9. Yuying Xie, Michael Kuhlmann, Frederik Rautenberg, Zheng-Hua Tan, and Reinhold Haeb-Umbach. “Speaker and style disentanglement of speech based on contrastive predictive coding supported factorized variational autoencoder”. In: Proceedings of the 32nd European Signal Processing Conference (EUSIPCO). Lyon, France, August 26–30, 2024.
  10. Asjid Tanveer, Jesper Jensen, Zheng-Hua Tan, and Jan Østergaard. “Envelope based deep source separation and EEG auditory attention decoding for speech and music”. In: Proceedings of the 32nd European Signal Processing Conference (EUSIPCO). Lyon, France, August 26–30, 2024.
  11. Deividas Eringis, John-Josef Leth, Zheng-Hua Tan, Rafal Wisniewski, and Mihály Petreczky. “PAC-Bayesian error bound, via Rényi divergence, for a class of linear time-invariant state-space models”. In: Proceedings of the 41st International Conference on Machine Learning (ICML). Vienna, Austria, July 21–27, 2024.
  12. Sarthak Yadav, Sergios Theodoridis, Lars Kai Hansen, and Zheng-Hua Tan. “Masked autoencoders with multi-window local-global attention are better audio learners”. In: Proceedings of the 12th International Conference on Learning Representations (ICLR). Vienna, Austria, May 7–11, 2024.
  13. Yuying Xie, Thomas Arildsen, and Zheng-Hua Tan. “Complex recurrent variational autoencoder for speech resynthesis and enhancement”. In: Proceedings of the IEEE World Congress on Computational Intelligence (WCCI). Yokohama, Japan, June 30–July 5, 2024.
  14. Holger Severin Bovbjerg, Jesper Jensen, Jan Østergaard, and Zheng-Hua Tan. “Self-supervised pre-training for robust personalized voice activity detection in adverse conditions”. In: Proceedings of the 49th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Seoul, Korea, April 14–19, 2024.
  15. Philippe Gonzalez, Zheng-Hua Tan, Jan Østergaard, Jesper Jensen, Tommy Sonne Alstrøm, and Tobias May. “Diffusion-based speech enhancement in matched and mismatched conditions using a Heun-based sampler”. In: Proceedings of the 49th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Seoul, Korea, April 14–19, 2024.
  16. Amin Edraki, Wai-Yip Geoffrey Chan, Jesper Jensen, and Daniel Fogerty. “Speaker adaptation for enhancement of bone-conducted speech”. In: Proceedings of the 49th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Seoul, Korea, April 14–19, 2024.
  17. Vasudha Sathyapriyan, Michael Syskind Pedersen, Mike Brookes, Jan Østergaard, Patrick A. Naylor, and Jesper Jensen. “Speech enhancement in hearing aids using target speech presence estimation based on a delayed remote microphone signal”. In: Proceedings of the 49th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Seoul, Korea, April 14–19, 2024.
  18. Vikas Tokala, Eric Grinstein, Mike Brookes, Simon Doclo, Jesper Jensen, and Patrick A. Naylor. “Binaural speech enhancement using deep complex convolutional transformer networks”. In: Proceedings of the 49th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Seoul, Korea, April 14–19, 2024.
  19. Deividas Eringis, John-Josef Leth, Zheng-Hua Tan, Rafal Wisniewski, and Mihály Petreczky. “PAC-Bayes generalisation bounds for dynamical systems including stable RNNs”. In: Proceedings of the 38th AAAI Conference on Artificial Intelligence (AAAI). Vancouver, Canada, February 20–27, 2024.
  20. Sarthak Yadav and Zheng-Hua Tan. “Audio Mamba: selective state spaces for self-supervised audio representations”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Kos, Greece, September 1–5, 2024.
  21. Mohammad Bokaei, Jesper Jensen, Simon Doclo, and Jan Østergaard. “Deep digital joint source-channel based wireless speech transmission”. In: Proceedings of the 32nd European Signal Processing Conference (EUSIPCO). Lyon, France, August 26–30, 2024.
  22. Aditya Joglekar, Iván López-Espejo, and John H. L. Hansen. “Fearless Steps APOLLO: identifying conversational mission-critical topics in NASA Apollo missions audio based on keyword spotting”. In: Proceedings of the NASA Human Research Program Investigators’ Workshop (IWS). Galveston, TX, USA, February 13–16, 2024.
  23. Aditya Joglekar, Iván López-Espejo, and John H. L. Hansen. “Fearless Steps APOLLO: challenges in keyword spotting and topic detection for naturalistic audio streams”. In: Proceedings of the 184th Meeting of the Acoustical Society of America. Chicago, IL, USA, May 13–17, 2024.

2023

Journal Papers

  1. Kateřina Žmolíková, Michael Syskind Pedersen, and Jesper Jensen. “Masked spectrogram prediction for unsupervised domain adaptation in speech enhancement”. In: IEEE Open Journal of Signal Processing 5 (2023), pp. 274–283.
  2. Yousef Mohammadi, Jan Østergaard, Carina Graversen, Ole Kæseler Andersen, and José Biurrun Manresa. “Validity and reliability of self-reported and neural measures of listening effort”. In: European Journal of Neuroscience 58.11 (2023), pp. 4357–4370.
  3. Yousef Mohammadi, Carina Graversen, Jan Østergaard, Ole Kæseler Andersen, and Tobias Reichenbach. “Phase-locking of neural activity to the envelope of speech in the delta frequency band reflects differences between word lists and sentences”. In: Journal of Cognitive Neuroscience 35.8 (2023), pp. 1301–1311.
  4. Yiming Zhang, Hong Yu, Ruoyi Du, Zheng-Hua Tan, Wenwu Wang, Zhanyu Ma, and Yuan Dong. “ACTUAL: audio captioning with caption feature space regularization”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 31 (2023), pp. 2643–2657.
  5. Achintya Kumar Sarkar and Zheng-Hua Tan. “On training targets and activation functions for deep representation learning in text-dependent speaker verification”. In: Acoustics 5.3 (2023), pp. 693–713.
  6. Christian Heider and Zheng-Hua Tan. “Leveraging domain features for detecting adversarial attacks against deep speech recognition in noise”. In: IEEE Open Journal of Signal Processing 4 (2023), pp. 179–187.
  7. Adèle Simon, Gérard Loquet, Jan Østergaard, and Søren Bech. “Cortical auditory attention decoding during music and speech listening”. In: IEEE Transactions on Neural Systems and Rehabilitation Engineering 31 (2023), pp. 2903–2911.
  8. Luca Turchet, Mathieu Lagrange, Cristina Rottondi, György Fazekas, Nils Peters, Jan Østergaard, Frederic Font, Tom Bäckström, and Carlo Fischione. “The Internet of Sounds: convergent trends, insights and future directions”. In: IEEE Internet of Things Journal 10.13 (2023), pp. 11264–11292.
  9. Iván López-Espejo, Amin Edraki, Wai-Yip Geoffrey Chan, Zheng-Hua Tan, and Jesper Jensen. “On the deficiency of intelligibility metrics as proxies for subjective intelligibility”. In: Speech Communication 150 (2023), pp. 9–22.

Conference Papers

  1. Vikas Tokala, Eric Grinstein, Mike Brookes, Simon Doclo, Jesper Jensen, and Patrick A. Naylor. “Binaural speech enhancement using complex convolutional recurrent networks”. In: Proceedings of the 57th Asilomar Conference on Signals, Systems, and Computers (ACSSC). Asilomar Conference Grounds, Pacific Grove, CA, USA, October 19–November 1, 2023.
  2. Mohammad Bokaei, Jesper Jensen, Simon Doclo, and Jan Østergaard. “Deep joint source-channel analog coding for low-latency speech transmission over Gaussian channels”. In: Proceedings of the 31st European Signal Processing Conference (EUSIPCO). Helsinki, Finland, September 4–8, 2023.
  3. Juan F. Montesinos, Daniel Michelsanti, Gloria Haro, Zheng-Hua Tan, and Jesper Jensen. “Speech inpainting: context-based speech synthesis guided by video”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Dublin, Ireland, August 20–24, 2023.
  4. Mo Zhou, Martin Bo Møller, Christian Sejer Pedersen, Niels Evert Marius de Koeijer, and Jan Østergaard. “Robust sound zone filters for synchronization errors”. In: Proceedings of Forum Acusticum 2023. Turin, Italy, September 11–15, 2023.
  5. Christian Sejer Pedersen, Mo Zhou, Martin Bo Møller, Niels Evert Marius de Koeijer, and Jan Østergaard. “Sound quality evaluation of packet loss concealment for wireless low-frequency sound zones”. In: Proceedings of Forum Acusticum 2023. Turin, Italy, September 11–15, 2023.
  6. Kaspar Müller, Bilgesu Çakmak, Paul Didier, Simon Doclo, Jan Østergaard, and Tobias Wolff. “Head orientation estimation with distributed microphones using speech radiation patterns”. In: Proceedings of the 57th Asilomar Conference on Signals, Systems, and Computers (ACSSC). Asilomar Conference Grounds, Pacific Grove, CA, USA, October 19–November 1, 2023.
  7. Matthias Blochberger, Jan Østergaard, Randall Ali, Marc Moonen, Filip Elvander, Jesper Jensen, and Toon van Waterschoot. “Adaptive coding in wireless acoustic sensor networks for distributed blind system identification”. In: Proceedings of the 57th Asilomar Conference on Signals, Systems, and Computers (ACSSC). Asilomar Conference Grounds, Pacific Grove, CA, USA, October 19–November 1, 2023, pp. 1420–1424.
  8. Ahmed Alghamdi, Leonard Moen, Wai-Yip Geoffrey Chan, Daniel Fogerty, and Jesper Jensen. “Correlation based glimpse proportion index”. In: Proceedings of the IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA). New Paltz, NY, USA, October 22–25, 2023, pp. 1–5.
  9. Vasudha Sathyapriyan, Michael Syskind Pedersen, Jan Østergaard, Mike Brookes, Patrick A. Naylor, and Jesper Jensen. “Speech enhancement using binary estimator selection applied to hearing aids with a remote microphone”. In: Proceedings of the IEEE International Conference on Frontiers of Signal Processing (ICFSP). Corfu, Greece, October 23–25, 2023.
  10. Yuying Xie, Thomas Arildsen, and Zheng-Hua Tan. “Improved disentangled speech representations using contrastive learning in factorized hierarchical variational autoencoder”. In: Proceedings of the 31st European Signal Processing Conference (EUSIPCO). Helsinki, Finland, September 4–8, 2023.
  11. Daniel Michelsanti, Zheng-Hua Tan, Sergi Rotger-Griful, and Jesper Jensen. “A vision-assisted hearing aid system based on deep learning”. In: Proceedings of the IEEE ICASSP 2023 Satellite Workshop on Augmented Hearing and Audition Technologies (AMHAT). Rhodes, Greece, June 4–10, 2023.
  12. Holger Severin Bovbjerg and Zheng-Hua Tan. “Improving label-deficient keyword spotting through self-supervised pretraining”. In: Proceedings of the IEEE ICASSP 2023 Satellite Workshop on Self-supervised Learning for Audio, Speech and Beyond (SASB). Rhodes, Greece, June 4–10, 2023.
  13. Iván López-Espejo, Santi Prieto, Alfonso Ortega, and Eduardo Lleida. “Improved vocal effort transfer vector estimation for vocal effort-robust speaker verification”. In: Proceedings of the IEEE International Workshop on Machine Learning for Signal Processing (MLSP). Rome, Italy, September 17–20, 2023.
  14. Christian Sejer Pedersen, Mo Zhou, Martin Bo Møller, Niels Evert Marius de Koeijer, and Jan Østergaard. “AR model for low latency packet loss concealment for wireless sound zones at low frequencies”. In: Proceedings of the 154th Audio Engineering Society (AES) Convention. Helsinki, Finland, May 20–23, 2023, pp. 1–10.
  15. Matthias Blochberger, Filip Elvander, Randall Ali, Jan Østergaard, Jesper Jensen, Marc Moonen, and Toon van Waterschoot. “Distributed adaptive norm estimation for blind system identification in wireless sensor networks”. In: Proceedings of the 48th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Rhodes, Greece, June 4–10, 2023.
  16. Manuel Morante, Jan Østergaard, and Sergios Theodoridis. “Interpretable nonnegative incoherent deep dictionary learning for fMRI data analysis”. In: Proceedings of the 48th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Rhodes, Greece, June 4–10, 2023.
  17. Iván López-Espejo, Ram C. M. C. Shekar, Zheng-Hua Tan, Jesper Jensen, and John H. L. Hansen. “Filterbank learning for noise-robust small-footprint keyword spotting”. In: Proceedings of the 48th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Rhodes, Greece, June 4–10, 2023.
  18. Mo Zhou, Martin Bo Møller, Christian Sejer Pedersen, and Jan Østergaard. “Robust FIR filters for wireless low-frequency sound zones”. In: Proceedings of the 48th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Rhodes, Greece, June 4–10, 2023.
  19. Jan Østergaard, Christian Sejer Pedersen, Mo Zhou, Niels Evert Marius de Koeijer, and Martin Bo Møller. “Multiple description audio coding for wireless low-frequency sound zones”. In: Proceedings of the IEEE Data Compression Conference (DCC). Snowbird, UT, USA, March 21–24, 2023.

2022

Journal Papers

  1. Georg Ø. Rønsch, Iván López-Espejo, Daniel Michelsanti, Yuying Xie, Petar Popovski, and Zheng-Hua Tan. “Utilization of acoustic signals with generative Gaussian and autoencoder modeling for condition-based maintenance of injection moulds”. In: International Journal of Computer Integrated Manufacturing 37.4 (2022), pp. 438–453.
  2. Himavanth Reddy, Asutosh Kar, and Jan Østergaard. “Performance analysis of low complexity fully connected neural networks for monaural speech enhancement”. In: Applied Acoustics 190 (2022), p. 108627.
  3. Srikanth Burra, Sanjana Sankar, Asutosh Kar, and Jan Østergaard. “A family of split kernel adaptive filtering algorithms for nonlinear stereophonic acoustic echo cancellation”. In: Journal of Ambient Intelligence and Humanized Computing 14.4 (2022), pp. 9907–9924.
  4. Jan Østergaard, Uri Erez, and Ram Zamir. “Incremental refinements and multiple descriptions with feedback”. In: IEEE Transactions on Information Theory 68.10 (2022), pp. 6915–6940.
  5. Payam Shahsavari Baboukani, Carina Graversen, Emina Alickovic, and Jan Østergaard. “Speech to noise ratio improvement induces nonlinear parietal phase synchrony in hearing aid users”. In: Frontiers in Neuroscience 16 (2022), p. 932959.
  6. Bjørn Uttrup Dideriksen, Kristoffer Derosche, and Zheng-Hua Tan. “iVAE-GAN: identifiable VAE-GAN models for latent representation learning”. In: IEEE Access 10 (2022), pp. 48405–48418.
  7. Mathias Bach Pedersen, Asger Heidemann Andersen, Søren Holdt Jensen, Zheng-Hua Tan, and Jesper Jensen. “Training data-driven speech intelligibility predictors on heterogeneous listening test data”. In: IEEE Access 10 (2022), pp. 66175–66189.
  8. Poul Hoang, Zheng-Hua Tan, Jan-Mark de Haan, and Jesper Jensen. “The minimum overlap-gap algorithm for speech enhancement”. In: IEEE Access 10 (2022), pp. 14698–14716.
  9. Iván López-Espejo, Zheng-Hua Tan, John H. L. Hansen, and Jesper Jensen. “Deep spoken keyword spotting: an overview”. In: IEEE Access 10 (2022), pp. 4169–4199.
  10. Santi Prieto, Alfonso Ortega, Iván López-Espejo, and Eduardo Lleida. “Shouted and whispered speech compensation for speaker verification systems”. In: Digital Signal Processing 127 (2022), p. 103536.
  11. Poul Hoang, Jan-Mark de Haan, Zheng-Hua Tan, and Jesper Jensen. “Multichannel speech enhancement with own voice-based interfering speech suppression for hearing assistive devices”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 30 (2022), pp. 706–720.

Conference Papers

  1. Iván López-Espejo, Zheng-Hua Tan, and Jesper Jensen. “An experimental study on light speech features for small-footprint keyword spotting”. In: Proceedings of IberSPEECH. Granada, Spain, November 23–25, 2022.
  2. Ángel M. Gómez, Victoria Sánchez, Antonio M. Peinado, Juan Manuel Martín-Doñas, Alejandro Gómez-Alanis, Amelia Villegas-Morcillo, Eros Roselló, Manuel Chica, Celia García, and Iván López-Espejo. “Fusion of classical digital signal processing and deep learning methods (FTCAPPS)”. In: Proceedings of IberSPEECH. Granada, Spain, November 23–25, 2022.
  3. Kaspar Müller, Simon Doclo, Jan Østergaard, and Tobias Wolff. “Model-based estimation of in-car-communication feedback applied to speech zone detection”. In: Proceedings of the International Workshop on Acoustic Signal Enhancement (IWAENC). Bamberg, Germany, September 5–8, 2022.
  4. Matthias Blochberger, Filip Elvander, Randall Ali, Marc Moonen, Jan Østergaard, Jesper Jensen, and Toon van Waterschoot. “Distributed cross-relation-based frequency-domain blind system identification using online-ADMM”. In: Proceedings of the International Workshop on Acoustic Signal Enhancement (IWAENC). Bamberg, Germany, September 5–8, 2022.
  5. Vasudha Sathyapriyan, Michael Syskind Pedersen, Jan Østergaard, Mike Brookes, Patrick A. Naylor, and Jesper Jensen. “A linear MMSE filter using delayed remote microphone signals for speech enhancement in hearing aid applications”. In: Proceedings of the International Workshop on Acoustic Signal Enhancement (IWAENC). Bamberg, Germany, September 5–8, 2022.
  6. José Miguel Cadavid Tobón, Martin Bo Møller, Søren Bech, Toon van Waterschoot, and Jan Østergaard. “Performance of low frequency sound zones based on truncated room impulse responses”. In: Proceedings of Audio Mostly. St. Pölten, Austria, September 6–9, 2022.
  7. Claus Meyer Larsen, Peter Koch, and Zheng-Hua Tan. “Adversarial multi-task deep learning for noise-robust voice activity detection with low algorithmic delay”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Incheon, Korea, September 18–22, 2022.
  8. Peter Asbjørn Leer Bysted, Jesper Jensen, Zheng-Hua Tan, Jan Østergaard, and Lars Bramsløw. “A neural network framework for modelling parameterized auditory models”. In: Proceedings of the Baltic-Nordic Acoustic Meeting (BNAM) Joint Acoustics Conference. Aalborg, Denmark, May 9–11, 2022.
  9. Christian Sejer Pedersen, Martin Bo Møller, and Jan Østergaard. “Effect of wireless transmission errors on sound zone performance at low frequencies”. In: Proceedings of the Baltic-Nordic Acoustic Meeting (BNAM) Joint Acoustics Conference. Aalborg, Denmark, May 9–11, 2022.
  10. Adèle Simon, Søren Bech, Gérard Loquet, and Jan Østergaard. “Electrodes selection for cortical auditory attention decoding during speech and music listening”. In: Proceedings of the International Conference on Information Fusion (FUSION). Linköping, Sweden, July 4–7, 2022.
  11. Adèle Simon, Jan Østergaard, Søren Bech, and Gérard Loquet. “Optimal time lags for linear cortical auditory attention detection: differences between speech and music listening”. In: Proceedings of the International Symposium on Hearing. Lyon, France, June 19–24, 2022.
  12. Peter Koch and Jan Østergaard. “The effect of fixed-point arithmetic on low frequency sound zone control”. In: Proceedings of the Baltic-Nordic Acoustic Meeting (BNAM) Joint Acoustics Conference. Aalborg, Denmark, May 9–11, 2022, pp. 125–134.
  13. Payam Shahsavari Baboukani, Sergios Theodoridis, and Jan Østergaard. “A stimuli-relevant directed dependency index for time series”. In: Proceedings of the 47th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Singapore, Singapore, May 22–27, 2022.
  14. Andreas Jonas Fuglsig, Jan Østergaard, Jesper Jensen, Lars Søndergaard Bertelsen, Peter Mariager, and Zheng-Hua Tan. “Joint far- and near-end speech intelligibility enhancement based on the approximated speech intelligibility index”. In: Proceedings of the 47th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Singapore, Singapore, May 22–27, 2022.

2021

Journal Papers

  1. Adel Zahedi, Michael Syskind Pedersen, Jan Østergaard, Thomas Ulrich Christiansen, Lars Bramsløw, and Jesper Jensen. “Minimum processing beamforming”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 29 (2021), pp. 2710–2724.
  2. Iván López-Espejo, Zheng-Hua Tan, and Jesper Jensen. “A novel loss function and training strategy for noise-robust keyword spotting”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 29 (2021), pp. 2254–2266.
  3. Jiyang Xie, Zhanyu Ma, Jianjun Lei, Guoqiang Zhang, Jing-Hao Xue, Zheng-Hua Tan, and Jun Guo. “Advanced dropout: a model-free methodology for Bayesian dropout optimization”. In: IEEE Transactions on Pattern Analysis and Machine Intelligence 44.9 (2021), pp. 4605–4625.
  4. Achintya Kumar Sarkar and Zheng-Hua Tan. “Self-segmentation of pass-phrase utterances for deep feature learning in text-dependent speaker verification”. In: Computer Speech & Language 70 (2021), p. 101229.
  5. Achintya Kumar Sarkar and Zheng-Hua Tan. “Vocal tract length perturbation for text-dependent speaker verification with autoregressive prediction coding”. In: IEEE Signal Processing Letters 28 (2021), pp. 364–368.
  6. Guttikonda Gowtham, Srikanth Burra, Asutosh Kar, Jan Østergaard, Pitikhate Sooraksa, Vladimir Mladenovic, and Diego B. Haddad. “A family of adaptive Volterra filters based on maximum correntropy criterion for improved active control of impulse noise”. In: Circuits, Systems, and Signal Processing 41.2 (2021), pp. 1019–1037.
  7. Srikanth Burra, Asutosh Kar, and Jan Østergaard. “Multiple sub-filter based proportionate filtering for nonlinear acoustic echo cancellation”. In: Applied Acoustics 182 (2021), p. 108215.
  8. Juan Manuel Martín-Doñas, Jesper Jensen, Zheng-Hua Tan, Ángel M. Gómez, and Antonio M. Peinado. “Online multichannel speech enhancement based on recursive EM and DNN-based speech presence estimation”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 28 (2021), pp. 3080–3094.
  9. Daniel Michelsanti, Zheng-Hua Tan, Shi-Xiong Zhang, Yong Xu, Meng Yu, Dong Yu, and Jesper Jensen. “An overview of deep-learning-based audio-visual speech enhancement and separation”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 29 (2021), pp. 1368–1396.
  10. Amin Edraki, Wai-Yip Geoffrey Chan, Jesper Jensen, and Daniel Fogerty. “Speech intelligibility prediction using spectro-temporal modulation analysis”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 29 (2021), pp. 210–225.
  11. Milan S. Derpich and Jan Østergaard. “Directed data-processing inequalities for systems with feedback”. In: Entropy 23.5 (2021).

Conference Papers

  1. Payam Shahsavari Baboukani, Carina Graversen, Emina Alickovic, and Jan Østergaard. “EEG phase synchrony reflects SNR levels during continuous speech-in-noise tasks”. In: Proceedings of the 43rd Annual International Conference of the IEEE Engineering in Medicine and Biology Society (EMBC). Guadalajara, Mexico, November 1–5, 2021.
  2. Morten Østergaard Nielsen, Jan Østergaard, Jesper Jensen, and Zheng-Hua Tan. “Compression of DNNs using magnitude pruning and nonlinear information bottleneck training”. In: Proceedings of the IEEE 31st International Workshop on Machine Learning for Signal Processing (MLSP). Gold Coast, Australia, October 25–28, 2021.
  3. Yuying Xie, Thomas Arildsen, and Zheng-Hua Tan. “Disentangled speech representation learning based on factorized hierarchical variational autoencoder with self-supervised objective”. In: Proceedings of the IEEE 31st International Workshop on Machine Learning for Signal Processing (MLSP). Gold Coast, Australia, October 25–28, 2021.
  4. Amin Edraki, Wai-Yip Geoffrey Chan, Jesper Jensen, and Daniel Fogerty. “A spectro-temporal glimpsing index (STGI) for speech intelligibility prediction”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Brno, Czech Republic, August 30–September 3, 2021.
  5. Giovanni Morrone, Daniel Michelsanti, Zheng-Hua Tan, and Jesper Jensen. “Audio-visual speech inpainting with deep learning”. In: Proceedings of the 46th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Toronto, Canada, June 6–11, 2021.
  6. Poul Hoang, Zheng-Hua Tan, Jan-Mark de Haan, and Jesper Jensen. “Joint maximum likelihood estimation of power spectral densities and relative acoustic transfer functions for acoustic beamforming”. In: Proceedings of the 46th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Toronto, Canada, June 6–11, 2021.
  7. Adèle Simon, Søren Bech, Gérard Loquet, and Jan Østergaard. “Auditory attention decoding during naturalistic music listening: a pilot study”. In: Proceedings of the 16th International Conference on Music Perception and Cognition (ICMPC). July 28–31, 2021.
  8. Jan Østergaard. “Low delay robust audio coding by noise shaping, fractional sampling, and source prediction”. In: Proceedings of the IEEE Data Compression Conference (DCC). Snowbird, UT, USA, March 23–26, 2021.
  9. Uri Erez, Jan Østergaard, and Ram Zamir. “An orthogonality principle for select-maximum estimation of exponential variables”. In: Proceedings of the IEEE International Symposium on Information Theory (ISIT). Melbourne, Australia, July 11–16, 2021.
  10. Md Sahidullah, Achintya Kumar Sarkar, Ville Vestman, Xuechen Liu, Romain Serizel, Tomi Kinnunen, Zheng-Hua Tan, and Emmanuel Vincent. “UIAI system for Short-Duration Speaker Verification Challenge 2020”. In: Proceedings of the IEEE Spoken Language Technology Workshop (SLT). Shenzhen, China, January 19–22, 2021.
  11. Zeyu Song, Dongliang Chang, Zhanyu Ma, Xiaoxu Li, and Zheng-Hua Tan. “CC-loss: channel correlation loss for image classification”. In: Proceedings of the 25th International Conference on Pattern Recognition (ICPR). Milan, Italy, January 10–15, 2021, pp. 7601–7608.

2020

Journal Papers

  1. Xiaoxu Li, Dongliang Chang, Zhanyu Ma, Zheng-Hua Tan, Jing-Hao Xue, Jie Cao, and Jun Guo. “Deep InterBoost networks for small-sample image classification”. In: Neurocomputing 456 (2020), pp. 492–503.
  2. Zhanyu Ma, Xiaoou Lu, Jiyang Xie, Zhen Yang, Jing-Hao Xue, Zheng-Hua Tan, Bo Xiao, and Jun Guo. “On the comparisons of decorrelation approaches for non-Gaussian neutral vector variables”. In: IEEE Transactions on Neural Networks and Learning Systems 34.4 (2020), pp. 1823–1837.
  3. Iván López-Espejo, Zheng-Hua Tan, and Jesper Jensen. “Improved external speaker-robust keyword spotting for hearing assistive devices”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 28 (2020), pp. 1233–1247.
  4. Xiaoxu Li, Dongliang Chang, Zhanyu Ma, Zheng-Hua Tan, Jing-Hao Xue, Jie Cao, Jingyi Yu, and Jun Guo. “OSLNet: deep small-sample classification with an orthogonal softmax layer”. In: IEEE Transactions on Image Processing 29 (2020), pp. 6482–6495.
  5. Morten Kolbæk, Zheng-Hua Tan, Søren Holdt Jensen, and Jesper Jensen. “On loss functions for supervised monaural time-domain speech enhancement”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 28 (2020), pp. 825–838.
  6. Miklas Strøm Kristoffersen, Sven Ewan Shepstone, and Zheng-Hua Tan. “The importance of context when recommending TV content: dataset and algorithms”. In: IEEE Transactions on Multimedia 22.6 (2020), pp. 1531–1541.
  7. Yonggang Qi and Zheng-Hua Tan. “SketchSegNet+: an end-to-end learning of RNN for multi-class sketch semantic segmentation”. In: IEEE Access 7 (2020), pp. 102717–102726.
  8. Zheng-Hua Tan, Achintya Kumar Sarkar, and Najim Dehak. “rVAD: an unsupervised segment-based robust voice activity detection method”. In: Computer Speech & Language 59 (2020), pp. 1–21.
  9. Martin Bo Møller and Jan Østergaard. “A moving horizon framework for sound zones”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 28 (2020), pp. 256–265.
  10. Payam Shahsavari Baboukani, Carina Graversen, Emina Alickovic, and Jan Østergaard. “Estimating conditional transfer entropy in time series using mutual information and nonlinear prediction”. In: Entropy 22.10 (2020), pp. 1–21.
  11. Jamal Amini, Richard Christian Hendriks, Richard Heusdens, Meng Guo, and Jesper Jensen. “Rate-constrained noise reduction in wireless acoustic sensor networks”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 28.1 (2020), pp. 1–12.

Conference Papers

  1. Juan Manuel Martín-Doñas, Antonio M. Peinado, Iván López-Espejo, and Ángel M. Gómez. “Dual-channel eKF-RTF framework for speech enhancement with DNN-based speech presence estimation”. In: Proceedings of IberSPEECH. Valladolid, Spain, November 25–27, 2020.
  2. Mathias Bach Pedersen, Morten Kolbæk, Asger Heidemann Andersen, Søren Holdt Jensen, and Jesper Jensen. “End-to-end speech intelligibility prediction using time-domain fully convolutional neural networks”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Shanghai, China, October 25–29, 2020.
  3. Santi Prieto, Alfonso Ortega, Iván López-Espejo, and Eduardo Lleida. “Shouted speech compensation for speaker verification robust to vocal effort conditions”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Shanghai, China, October 25–29, 2020.
  4. Daniel Michelsanti, Olga Slizovskaia, Gloria Haro, Emilia Gómez, Zheng-Hua Tan, and Jesper Jensen. “Vocoder-based speech synthesis from silent videos”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Shanghai, China, October 25–29, 2020.
  5. Iván López-Espejo, Zheng-Hua Tan, and Jesper Jensen. “Exploring filterbank learning for keyword spotting”. In: Proceedings of the 28th European Signal Processing Conference (EUSIPCO). Amsterdam, The Netherlands, January 18–22, 2021.
  6. Payam Shahsavari Baboukani, Carina Graversen, and Jan Østergaard. “Estimation of directed dependencies in time series using conditional mutual information and non-linear prediction”. In: Proceedings of the 28th European Signal Processing Conference (EUSIPCO). Amsterdam, The Netherlands, January 18–22, 2021.
  7. Saeid Samizade, Zheng-Hua Tan, Chao Shen, and Xiaohong Guan. “Adversarial example detection by classification for deep speech recognition”. In: Proceedings of the 45th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2020.
  8. Mathias Bach Pedersen, Asger Heidemann Andersen, Søren Holdt Jensen, and Jesper Jensen. “A neural network for monaural intrusive speech intelligibility prediction”. In: Proceedings of the 45th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2020.
  9. Poul Hoang, Zheng-Hua Tan, Thomas Lunner, Jan-Mark de Haan, and Jesper Jensen. “Maximum likelihood estimation of the interference-plus-noise cross power spectral density matrix for own voice retrieval”. In: Proceedings of the 45th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2020.
  10. Adel Zahedi, Michael Syskind Pedersen, Jan Østergaard, Lars Bramsløw, Thomas Ulrich Christiansen, and Jesper Jensen. “A constrained maximum likelihood estimator of speech and noise spectra with application to multi-microphone noise reduction”. In: Proceedings of the 45th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2020.
  11. Andreas Koutrouvelis, Richard Christian Hendriks, Richard Heusdens, and Jesper Jensen. “Robust joint estimation of multimicrophone signal model parameters”. In: Proceedings of the 45th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2020.
  12. Jamal Amini, Richard Christian Hendriks, Richard Heusdens, Meng Guo, and Jesper Jensen. “Rate-constrained noise reduction in wireless acoustic sensor networks”. In: Proceedings of the 45th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Barcelona, Spain, May 4–8, 2020.
  13. Uri Erez, Jan Østergaard, and Ram Zamir. “The exponential distribution in rate distortion theory: the case of compression with independent encodings”. In: Proceedings of the IEEE Data Compression Conference (DCC). Snowbird, UT, USA, March 24–27, 2020.

2019

Journal Papers

  1. Andreas Jonas Fuglsig and Jan Østergaard. “Zero-delay multiple descriptions of stationary scalar Gauss-Markov sources”. In: Entropy 21.12 (2019), p. 1185.
  2. Daniel Michelsanti, Zheng-Hua Tan, and Jesper Jensen. “Deep-learning-based audio-visual speech enhancement in presence of Lombard effect”. In: Speech Communication 115 (2019), pp. 38–50.
  3. Achintya Kumar Sarkar, Zheng-Hua Tan, Hao Tang, Suwon Shon, and James Glass. “Time-contrastive learning based deep bottleneck features for text-dependent speaker verification”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 27.8 (2019), pp. 1267–1279.
  4. Juan Manuel Martín-Doñas, Antonio M. Peinado, Iván López-Espejo, and Ángel M. Gómez. “Dual-channel speech enhancement based on extended Kalman filter relative transfer function estimation”. In: Applied Sciences 9.12 (2019), pp. 1–21.
  5. Morten Kolbæk, Zheng-Hua Tan, and Jesper Jensen. “On the relationship between short-time objective intelligibility and short-time spectral-amplitude mean-square error for speech enhancement”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 27.2 (2019), pp. 283–295.
  6. Asutosh Kar, Ankita Anand, Jan Østergaard, Søren Holdt Jensen, and M. N. Srikanta Swamy. “Sound quality improvement for hearing aids in presence of multiple inputs”. In: Circuits, Systems, and Signal Processing 38.8 (2019), pp. 3591–3615.
  7. Asutosh Kar, Ankita Anand, Jan Østergaard, Søren Holdt Jensen, and M. N. Srikanta Swamy. “Mean square performance evaluation in frequency domain for an improved adaptive feedback cancellation in hearing aids”. In: Signal Processing 157 (2019), pp. 45–61.
  8. Mohsen Zareian Jahromi, Adel Zahedi, Jesper Jensen, and Jan Østergaard. “Information loss in the human auditory system”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 27.3 (2019), pp. 472–481.

Conference Papers

  1. Amin Edraki, Wai-Yip Geoffrey Chan, Jesper Jensen, and Daniel Fogerty. “Improvement and assessment of spectro-temporal modulation analysis for speech intelligibility estimation”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Graz, Austria, September 15–19, 2019, pp. 1378–1382.
  2. Andreas Koutrouvelis, Richard Christian Hendriks, Richard Heusdens, and Jesper Jensen. “Estimation of sensor array signal model parameters using factor analysis”. In: Proceedings of the 27th European Signal Processing Conference (EUSIPCO). A Coruña, Spain, September 2–6, 2019.
  3. Poul Hoang, Zheng-Hua Tan, Jan-Mark de Haan, Thomas Lunner, and Jesper Jensen. “Robust Bayesian and maximum a posteriori beamforming for hearing assistive devices”. In: Proceedings of the 7th IEEE Global Conference on Signal and Information Processing (GlobalSIP). Ottawa, Canada, November 11–14, 2019.
  4. Jiyang Xie, Zhanyu Ma, Guoqiang Zhang, Jing-Hao Xue, Zheng-Hua Tan, and Jun Guo. “Soft dropout and its variational Bayes approximation”. In: Proceedings of the IEEE International Workshop on Machine Learning for Signal Processing (MLSP). Pittsburgh, PA, USA, October 13–16, 2019.
  5. Iván López-Espejo, Zheng-Hua Tan, and Jesper Jensen. “Keyword spotting for hearing assistive devices robust to external speakers”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Graz, Austria, September 15–19, 2019.
  6. Miklas Strøm Kristoffersen, Jacob L. Wieland, Sven Ewan Shepstone, Zheng-Hua Tan, and Vinoba Vinayagamoorthy. “Deep joint embeddings of context and content for recommendation”. In: Proceedings of the CARS 2.0 Workshop on Context-Aware Recommender Systems, held in conjunction with RecSys. Copenhagen, Denmark, September 20, 2019.
  7. Daniel Michelsanti, Zheng-Hua Tan, Sigurdur Sigurdsson, and Jesper Jensen. “Effects of Lombard reflex on the performance of deep-learning-based audio-visual speech enhancement systems”. In: Proceedings of the 44th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Brighton, United Kingdom, May 12–17, 2019.
  8. Daniel Michelsanti, Zheng-Hua Tan, Sigurdur Sigurdsson, and Jesper Jensen. “On training targets and objective functions for deep-learning-based audio-visual speech enhancement”. In: Proceedings of the 44th Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Brighton, United Kingdom, May 12–17, 2019.
  9. Andrea Coifman, Peter Rohoska, Miklas Strøm Kristoffersen, Sven Ewan Shepstone, and Zheng-Hua Tan. “Subjective annotations for vision-based attention level estimation”. In: Proceedings of the 14th International Conference on Computer Vision Theory and Applications (VISAPP). Prague, Czech Republic, February 25–27, 2019.

2018

Journal Papers

  1. Photios A. Stavrou, Jan Østergaard, and Charalambos D. Charalambous. “Zero-delay rate distortion via filtering for vector-valued Gaussian sources”. In: IEEE Journal of Selected Topics in Signal Processing 12.5 (2018), pp. 841–856.
  2. Jamal Amini, Richard Christian Hendriks, Richard Heusdens, Meng Guo, and Jesper Jensen. “Asymmetric coding for rate-constrained noise reduction in binaural hearing aids”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 27.1 (2018), pp. 154–167.
  3. Asger Heidemann Andersen, Jan-Mark de Haan, Zheng-Hua Tan, and Jesper Jensen. “Refinement and validation of the binaural short-time objective intelligibility measure for spatially diverse conditions”. In: Speech Communication 102 (2018), pp. 1–13.
  4. Asger Heidemann Andersen, Jan-Mark de Haan, Zheng-Hua Tan, and Jesper Jensen. “Non-intrusive speech intelligibility prediction using convolutional neural networks”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 26.10 (2018), pp. 1925–1939.
  5. Xiaodong Duan and Zheng-Hua Tan. “A spatial self-similarity based feature learning method for face recognition under varying poses”. In: Pattern Recognition Letters 111 (2018), pp. 109–116.
  6. Mojtaba Farmani, Michael Syskind Pedersen, Zheng-Hua Tan, and Jesper Jensen. “Bias-compensated informed sound source localization using relative transfer functions”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 26.7 (2018), pp. 1275–1289.
  7. Sven Ewan Shepstone, Zheng-Hua Tan, and Miklas Strøm Kristoffersen. “Using closed-set speaker identification score confidence to enhance audio-based collaborative filtering for multiple users”. In: IEEE Transactions on Consumer Electronics 64.1 (2018), pp. 11–18.
  8. Sebastian Braun, Adam Kuklasinski, Ofer Schwartz, Oliver Thiergart, Emanuël A. P. Habets, Sharon Gannot, Simon Doclo, and Jesper Jensen. “Evaluation and comparison of late reverberation power spectral density estimators”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 26.6 (2018), pp. 1056–1071.

Conference Papers

  1. Evgenios Vlachos and Zheng-Hua Tan. “Public perception of Android robots: indications from an analysis of YouTube comments”. In: Proceedings of the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). Madrid, Spain, October 1–5, 2018.
  2. Hong Yu, Tianrui Hu, Zhanyu Ma, Zheng-Hua Tan, and Jun Guo. “Multi-task adversarial network bottleneck features for noise-robust speaker verification”. In: Proceedings of the IEEE International Conference on Network Infrastructure and Digital Content (IC-NIDC). Guiyang, China, August 22–24, 2018.
  3. Gabriele Trovato, Renato Paredes, Javier Balvin, Francisco Fabian Cuéllar, Nicolai Bæk Thomsen, Søren Bech, and Zheng-Hua Tan. “The sound or silence: investigating the influence of robot noise on proxemics”. In: Proceedings of the 27th IEEE International Conference on Robot and Human Interactive Communication (RO-MAN). Nanjing and Tai’an, China, August 27–31, 2018.
  4. Peter Sibbern Frederiksen, Jess Villalba, Shinji Watanabe, Zheng-Hua Tan, and Najim Dehak. “Effectiveness of single-channel BLSTM enhancement for language identification”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Hyderabad, India, September 2–6, 2018.
  5. Mojtaba Farmani, Michael Syskind Pedersen, and Jesper Jensen. “Sound source localization for hearing aid applications using wireless microphones”. In: Proceedings of the IEEE Sensor Array and Multichannel Signal Processing Workshop (SAM). Sheffield, United Kingdom, July 8–11, 2018.
  6. Jamal Amini, Richard Christian Hendriks, Richard Heusdens, Meng Guo, and Jesper Jensen. “Operational rate-constrained noise reduction for generalized binaural hearing aid setups”. In: Proceedings of the Symposium on Information Theory and Signal Processing in the Benelux. 2018.
  7. Andreas Koutrouvelis, Richard Christian Hendriks, Richard Heusdens, Steven van de Par, Jesper Jensen, and Meng Guo. “Evaluation of binaural noise reduction methods in terms of intelligibility and perceived localization”. In: Proceedings of the 26th European Signal Processing Conference (EUSIPCO). Rome, Italy, September 3–7, 2018.
  8. Jamal Amini, Richard Christian Hendriks, Richard Heusdens, Meng Guo, and Jesper Jensen. “Operational rate-constrained beamforming in binaural hearing aids”. In: Proceedings of the 26th European Signal Processing Conference (EUSIPCO). Rome, Italy, September 3–7, 2018.
  9. Photios A. Stavrou, Jan Østergaard, and Mikael Skoglund. “On zero-delay source coding of LTI Gauss-Markov systems with covariance matrix distortion constraints”. In: Proceedings of the European Control Conference (ECC). Limassol, Cyprus, June 12–15, 2018.
  10. Morten Kolbæk, Zheng-Hua Tan, and Jesper Jensen. “Monaural speech enhancement using deep neural networks by maximizing a short-time objective intelligibility measure”. In: Proceedings of the 43rd Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Calgary, Canada, April 15–20, 2018.
  11. Photios A. Stavrou and Jan Østergaard. “Fixed-rate zero-delay source coding for stationary vector-valued Gauss-Markov sources”. In: Proceedings of the IEEE Data Compression Conference (DCC). Snowbird, UT, USA, March 20–23, 2018.

2017

Journal Papers

  1. Renhua Peng, Zheng-Hua Tan, Xiaodong Li, and Chengshi Zheng. “A perceptually motivated LP residual estimator in noisy and reverberant environments”. In: Speech Communication 96 (2017), pp. 129–141.
  2. Hong Yu, Zheng-Hua Tan, Zhanyu Ma, Rainer Martin, and Jun Guo. “Spoofing detection in automatic speaker verification systems using DNN classifiers and dynamic acoustic features”. In: IEEE Transactions on Neural Networks and Learning Systems 29.10 (2017), pp. 4633–4644.
  3. Md Sahidullah, Dennis Alexander Lehmann Thomsen, Rosa González Hautamäki, Tomi Kinnunen, Zheng-Hua Tan, Robert Parts, and Martti Pitkanen. “Robust voice liveness detection and speaker verification using throat microphones”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 26.1 (2017), pp. 44–56.
  4. Zheng-Hua Tan, Nicolai Bæk Thomsen, Xiaodong Duan, Evgenios Vlachos, Sven Ewan Shepstone, Morten Højfeldt Rasmussen, and Jesper Lisby Højvang. “iSocioBot: a multimodal interactive social robot”. In: International Journal of Social Robotics 10.1 (2017), pp. 5–19.
  5. Achintya Kumar Sarkar and Zheng-Hua Tan. “Incorporating pass-phrase dependent background models for text-dependent speaker verification”. In: Computer Speech & Language 47 (2017), pp. 259–271.
  6. Jen-Tzung Chien, Chao-Hsi Lee, and Zheng-Hua Tan. “Latent Dirichlet mixture model”. In: Neurocomputing 278 (2017), pp. 12–22.
  7. Stefanos Astaras, Aristodemos Pnevmatikakis, and Zheng-Hua Tan. “Visual detection of events of interest from urban activity”. In: Wireless Personal Communications 97.2 (2017), pp. 1877–1888.
  8. Morten Kolbæk, Dong Yu, Zheng-Hua Tan, and Jesper Jensen. “Multi-talker speech separation with utterance-level permutation invariant training of deep recurrent neural networks”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 25.10 (2017), pp. 1901–1913.
  9. Hong Yu, Zheng-Hua Tan, Yiming Zhang, Zhanyu Ma, and Jun Guo. “DNN filter bank cepstral coefficients for spoofing detection”. In: IEEE Access 5 (2017), pp. 4779–4787.
  10. Mojtaba Farmani, Michael Syskind Pedersen, Zheng-Hua Tan, and Jesper Jensen. “Informed sound source localization using relative transfer functions for hearing aid applications”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 25.3 (2017), pp. 611–623.
  11. Zhanyu Ma, Jing-Hao Xue, Arne Leijon, Zheng-Hua Tan, Zhen Yang, and Jun Guo. “Decorrelation of neutral vector variables: theory and applications”. In: IEEE Transactions on Neural Networks and Learning Systems 29.1 (2017), pp. 129–143.
  12. Sven Ewan Shepstone, Zheng-Hua Tan, and Søren Holdt Jensen. “Audio-based granularity-adapted emotion classification”. In: IEEE Transactions on Affective Computing 9.2 (2017), pp. 176–190.
  13. Zhanyu Ma, Hong Yu, Zheng-Hua Tan, and Jun Guo. “Text-independent speaker identification using the histogram transform model”. In: IEEE Access 4 (2017), pp. 9733–9739.
  14. Adam Kuklasiński and Jesper Jensen. “Multi-channel Wiener filters in binaural and bilateral hearing aids: speech intelligibility improvement and robustness to DoA errors”. In: Journal of the Audio Engineering Society 65.1–2 (2017), pp. 8–16.
  15. Andreas Koutrouvelis, Richard Christian Hendriks, Richard Heusdens, and Jesper Jensen. “Relaxed binaural LCMV beamforming”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 25.1 (2017), pp. 133–148.
  16. Morten Kolbæk, Zheng-Hua Tan, and Jesper Jensen. “Speech intelligibility potential of general and specialized deep neural network based speech enhancement systems”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 25.1 (2017), pp. 149–163.

Conference Papers

  1. Achintya Kumar Sarkar and Zheng-Hua Tan. “Time-contrastive learning based DNN bottleneck features for text-dependent speaker verification”. In: Proceedings of the NIPS 2017 Time Series Workshop. Long Beach, CA, USA, December 8, 2017.
  2. Xiaodong Duan, Nicolai Bæk Thomsen, Zheng-Hua Tan, Borge Lindberg, and Søren Holdt Jensen. “Weighted score based fast converging co-training with application to audio-visual person identification”. In: Proceedings of the 29th IEEE International Conference on Tools with Artificial Intelligence (ICTAI). Boston, MA, USA, November 6–8, 2017.
  3. Photios A. Stavrou, Jan Østergaard, Charalambos D. Charalambous, and Milan S. Derpich. “An upper bound to zero-delay rate distortion via Kalman filtering for vector Gaussian sources”. In: Proceedings of the IEEE Information Theory Workshop (ITW). Kaohsiung, Taiwan, November 6–10, 2017.
  4. Morten Kolbæk, Dong Yu, Zheng-Hua Tan, and Jesper Jensen. “Joint separation and denoising of noisy multi-talker speech using recurrent neural networks and permutation invariant training”. In: Proceedings of the IEEE International Workshop on Machine Learning for Signal Processing (MLSP). Tokyo, Japan, September 25–28, 2017. This publication received the Best Student Paper Award.
  5. Photios A. Stavrou and Jan Østergaard. “A lower bound on causal and zero-delay rate distortion for scalar Gaussian autoregressive sources”. In: Proceedings of the Symposium on Information Theory and Signal Processing in the Benelux. Delft, The Netherlands, May 24–25, 2017, pp. 207–214.
  6. Mohsen Zareian Jahromi, Jan Østergaard, and Jesper Jensen. “Humans do not maximize the probability of correct decision when recognizing DANTALE words in noise”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Stockholm, Sweden, August 20–24, 2017.
  7. Asger Heidemann Andersen, Jan-Mark de Haan, Zheng-Hua Tan, and Jesper Jensen. “On the use of band importance weighting in the short-time objective intelligibility measure”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Stockholm, Sweden, August 20–24, 2017.
  8. Hong Yu, Zheng-Hua Tan, Zhanyu Ma, and Jun Guo. “Adversarial network bottleneck features for noise robust speaker verification”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Stockholm, Sweden, August 20–24, 2017.
  9. Daniel Michelsanti and Zheng-Hua Tan. “Conditional generative adversarial networks for speech enhancement and noise-robust speaker verification”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Stockholm, Sweden, August 20–24, 2017.
  10. Achintya Kumar Sarkar, Md Sahidullah, Zheng-Hua Tan, and Tomi Kinnunen. “Improving speaker verification performance in presence of spoofing attacks using out-of-domain spoofed data”. In: Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech). Stockholm, Sweden, August 20–24, 2017.
  11. Dong Yu, Morten Kolbæk, Zheng-Hua Tan, and Jesper Jensen. “Permutation invariant training of deep models for speaker-independent multi-talker speech separation”. In: Proceedings of the 42nd Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). New Orleans, LA, USA, March 5–9, 2017.
  12. Asger Heidemann Andersen, Jan-Mark de Haan, Zheng-Hua Tan, and Jesper Jensen. “A non-intrusive short-time objective intelligibility measure”. In: Proceedings of the 42nd Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). New Orleans, LA, USA, March 5–9, 2017.
  13. Tomi Kinnunen, Md Sahidullah, Mauro Falcone, Luca Costantini, Rosa González Hautamäki, Dennis Alexander Lehmann Thomsen, Achintya Kumar Sarkar, Zheng-Hua Tan, Hector Delgado, Massimiliano Todisco, Nicholas Evans, Ville Hautamaki, and Kong Aik Lee. “RedDots replayed: a new replay spoofing attack corpus for text-dependent speaker verification research”. In: Proceedings of the 42nd Annual IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). New Orleans, LA, USA, March 5–9, 2017.
  14. Jan Østergaard, Yuval Kochman, and Ram Zamir. “An asymmetric difference multiple description Gaussian noise channel”. In: Proceedings of the IEEE Data Compression Conference (DCC). Snowbird, UT, USA, April 4–7, 2017.

2016

Journal Papers

  1. Adel Zahedi, Jan Østergaard, Søren Holdt Jensen, Patrick A. Naylor, and Søren Bech. “Source coding in networks with covariance distortion constraints”. In: IEEE Transactions on Signal Processing 64.22 (2016), pp. 5943–5958.
  2. Jesper Jensen and Cees H. Taal. “An algorithm for predicting the intelligibility of speech masked by modulated noise maskers”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 24.11 (2016), pp. 2009–2022.
  3. Asger Heidemann Andersen, Zheng-Hua Tan, Jan-Mark de Haan, and Jesper Jensen. “Predicting the intelligibility of noisy and nonlinearly processed binaural speech”. In: IEEE/ACM Transactions on Audio, Speech, and Language Processing 24.11 (2016), pp. 1908–1920.

Conference Papers

  1. Mojtaba Farmani, Richard Heusdens, Michael Syskind Pedersen, Zheng-Hua Tan, and Jesper Jensen. “TDOA-based self-calibration of dual-microphone arrays”. In: Proceedings of the 19th International Conference on Information Fusion (FUSION). Heidelberg, Germany, July 5–8, 2016, pp. 1931–1936.
  2. Morten Kolbæk, Zheng-Hua Tan, and Jesper Jensen. “Speech enhancement using long short-term memory based recurrent neural networks for noise robust speaker verification”. In: Proceedings of the IEEE Spoken Language Technology Workshop (SLT). San Diego, CA, USA, December 13–16, 2016.
  3. Hector Delgado, Massimiliano Todisco, Md Sahidullah, Achintya Kumar Sarkar, Nicholas Evans, Tomi Kinnunen, and Zheng-Hua Tan. “Further optimisations of constant Q cepstral processing for integrated utterance and text-dependent speaker verification”. In: Proceedings of the IEEE Spoken Language Technology Workshop (SLT). San Diego, CA, USA, December 13–16, 2016.
  4. Adam Mashiach, Yuval Kochman, Jan Østergaard, and Ram Zamir. “Two asymmetric descriptions from many symmetric descriptions”. In: Proceedings of the International Conference on the Science of Electrical Engineering (ICSEE). Eilat, Israel, November 16–18, 2016.
  5. Mohsen Zareian Jahromi, Jan Østergaard, and Jesper Jensen. “Detection of spoken words in noise: comparison of human performance to maximum likelihood detection”. In: Proceedings of the IEEE Global Conference on Signal and Information Processing (GlobalSIP). Washington, DC, USA, December 7–9, 2016.