Haizhou Li
Researcher Next ID · RN-025750
Researcher · Computer Science
National University of Singapore
Singapore, Singapore
- Works count
- 1,627
- Citation count
- 30,663
- H-index
- 75
- i10-index
- 623
Research interests
Publications
HuatuoGPT, Towards Taming Language Model to Be a Doctor
Journal · 2023 · https://doi.org/10.18653/v1/2023.findings-emnlp.725
Ego4D: Around the World in 3,000 Hours of Egocentric Video
2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) · 2022 · https://doi.org/10.1109/cvpr52688.2022.01842
ADD 2022: the first Audio Deep Synthesis Detection Challenge
ICASSP 2022 - 2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) · 2022 · 10.1109/icassp43922.2022.9746939
ADD 2022: the first Audio Deep Synthesis Detection Challenge
ICASSP 2022 - 2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) · 2022 · 10.1109/icassp43922.2022.9746939
Seen and Unseen Emotional Style Transfer for Voice Conversion with A New Emotional Speech Dataset
Journal · 2021 · https://doi.org/10.1109/icassp39728.2021.9413391
Emotional voice conversion: Theory, databases and ESD
Speech Communication · 2021 · https://doi.org/10.1016/j.specom.2021.11.006
An Overview of Voice Conversion and Its Challenges: From Statistical Modeling to Deep Learning
IEEE/ACM Transactions on Audio Speech and Language Processing · 2020 · https://doi.org/10.1109/taslp.2020.3038524
A Cost-Sensitive Deep Belief Network for Imbalanced Classification
IEEE Transactions on Neural Networks and Learning Systems · 2018 · https://doi.org/10.1109/tnnls.2018.2832648
A learning-based approach to direction of arrival estimation in noisy and reverberant environments
Journal · 2015 · https://doi.org/10.1109/icassp.2015.7178484
Text-dependent speaker verification: Classifiers, databases and RSR2015
Speech Communication · 2014 · https://doi.org/10.1016/j.specom.2014.03.001
Spoofing and countermeasures for speaker verification: A survey
Speech Communication · 2014 · https://doi.org/10.1016/j.specom.2014.10.005
Exemplar-Based Sparse Representation With Residual Compensation for Voice Conversion
IEEE/ACM Transactions on Audio Speech and Language Processing · 2014 · https://doi.org/10.1109/taslp.2014.2333242
Precise-Spike-Driven Synaptic Plasticity: Learning Hetero-Association of Spatiotemporal Spike Patterns
PLoS ONE · 2013 · https://doi.org/10.1371/journal.pone.0078318
Making Social Robots More Attractive: The Effects of Voice Pitch, Humor and Empathy
International Journal of Social Robotics · 2013 · https://doi.org/10.1007/s12369-012-0171-x
Spoken Language Recognition: From Fundamentals to Practice
Proceedings of the IEEE · 2013 · https://doi.org/10.1109/jproc.2012.2237151
A first speech recognition system for Mandarin-English code-switch conversational speech
· 2012 · 10.1109/icassp.2012.6289015
Vulnerability of speaker verification systems against voice conversion spoofing attacks: The case of telephone speech
Journal · 2012 · https://doi.org/10.1109/icassp.2012.6288895
Detecting converted speech and natural speech for anti-spoofing attack in speaker recognition
Journal · 2012 · https://doi.org/10.21437/interspeech.2012-465
IRIS: a Chat-oriented Dialogue System based on the Vector Space Model
Journal · 2012
A first speech recognition system for Mandarin-English code-switch conversational speech
· 2012 · 10.1109/icassp.2012.6289015
Spectrogram Image Feature for Sound Event Classification in Mismatched Conditions
IEEE Signal Processing Letters · 2010 · https://doi.org/10.1109/lsp.2010.2100380
An overview of text-independent speaker recognition: From features to supervectors
Speech Communication · 2009 · https://doi.org/10.1016/j.specom.2009.08.009
Level-set based automatic cup-to-disc ratio determination using retinal fundus images in ARGALI
Journal · 2008 · https://doi.org/10.1109/iembs.2008.4649648
A Vector Space Modeling Approach to Spoken Language Identification
IEEE Transactions on Audio Speech and Language Processing · 2006 · https://doi.org/10.1109/tasl.2006.876860
Efficient and Robust Feature Extraction by Maximum Margin Criterion
IEEE Transactions on Neural Networks · 2006 · https://doi.org/10.1109/tnn.2005.860852
A joint source-channel model for machine transliteration
Journal · 2004 · https://doi.org/10.3115/1218955.1218976
Automated Feature Extraction in Color Retinal Images by a Model Based Approach
IEEE Transactions on Biomedical Engineering · 2004 · https://doi.org/10.1109/tbme.2003.820400
Current projects
No projects listed.