@shinjiw_at_cmu: My Google Scholar h-index just reached 100! My first paper was in 1999, and behind these 100 are students and collabora…

X AI KOLs Following News

Summary

Shinji Watanabe announces his Google Scholar h-index has reached 100, crediting decades of work on speech processing research and open-source tooling to his students and collaborators. The post highlights landmark works like ESPnet, Deep Clustering, and SUPERB that underpin modern speech recognition.

My Google Scholar h-index just reached 100! 🙂 My first paper was in 1999, and behind these 100 are students and collaborators in academia and industry, plus the open-source community and challenges. This number is really yours. Thank you all! 🙏 https://t.co/IVGQM30qe8 https://t.co/PUpOCJBMxJ
Original Article
View Cached Full Text

Cached at: 10/04/26, 01:04 AM

My Google Scholar h-index just reached 100! 🙂

My first paper was in 1999, and behind these 100 are students and collaborators in academia and industry, plus the open-source community and challenges.

This number is really yours. Thank you all! 🙏

https://t.co/IVGQM30qe8 https://t.co/PUpOCJBMxJ


Shinji Watanabe

Source: https://scholar.google.com/citations?user=U5xRA6QAAAAJ ESPnet: End-to-end speech processing toolkitS Watanabe, T Hori, S Karita, T Hayashi, J Nishitoba, Y Unno, NEY Soplin, ...

arXiv preprint arXiv:1804.00015, 2018

21282018Deep clustering: Discriminative embeddings for segmentation and separationJR Hershey, Z Chen, J Le Roux, S Watanabe

2016 IEEE international conference on acoustics, speech and signal …, 2016

19442016Superb: Speech processing universal performance benchmarkS Yang, PH Chi, YS Chuang, CIJ Lai, K Lakhotia, YY Lin, AT Liu, J Shi, ...

arXiv preprint arXiv:2105.01051, 2021

15022021Joint CTC-attention based end-to-end speech recognition using multi-task learningS Kim, T Hori, S Watanabe

2017 IEEE international conference on acoustics, speech and signal …, 2017

12532017Hybrid CTC/attention architecture for end-to-end speech recognitionS Watanabe, T Hori, S Kim, JR Hershey, T Hayashi

IEEE Journal of Selected Topics in Signal Processing 11 (8), 1240-1253, 2017

12012017A comparative study on transformer vs rnn in speech applicationsS Karita, N Chen, T Hayashi, T Hori, H Inaguma, Z Jiang, M Someki, ...

2019 IEEE automatic speech recognition and understanding workshop (ASRU …, 2019

10882019The third ‘CHiME’speech separation and recognition challenge: Dataset, task and baselinesJ Barker, R Marxer, E Vincent, S Watanabe

2015 IEEE workshop on automatic speech recognition and understanding (ASRU …, 2015

9312015Phase-sensitive and recognition-boosted speech separation using deep recurrent neural networksH Erdogan, JR Hershey, S Watanabe, J Le Roux

2015 IEEE International Conference on Acoustics, Speech and Signal …, 2015

8722015Speech enhancement with LSTM recurrent neural networks and its application to noise-robust ASRF Weninger, H Erdogan, S Watanabe, E Vincent, J Le Roux, JR Hershey, ...

International conference on latent variable analysis and signal separation …, 2015

8322015Self-supervised speech representation learning: A reviewA Mohamed, H Lee, L Borgholt, JD Havtorn, J Edin, C Igel, K Kirchhoff, ...

IEEE Journal of Selected Topics in Signal Processing 16 (6), 1179-1210, 2022

7102022A review of speaker diarization: Recent advances with deep learningTJ Park, N Kanda, D Dimitriadis, KJ Han, S Watanabe, S Narayanan

Computer speech & language 72, 101317, 2022

6532022Single-channel multi-speaker separation using deep clusteringY Isik, JL Roux, Z Chen, S Watanabe, JR Hershey

arXiv preprint arXiv:1607.02173, 2016

5692016CHiME-6 challenge: Tackling multispeaker speech recognition for unsegmented recordingsS Watanabe, M Mandel, J Barker, E Vincent, A Arora, X Chang, ...

arXiv preprint arXiv:2004.09249, 2020

5522020The fifth’CHiME’speech separation and recognition challenge: dataset, task and baselinesJ Barker, S Watanabe, E Vincent, J Trmal

arXiv preprint arXiv:1803.10609, 2018

5392018End-to-end speech recognition: A surveyR Prabhavalkar, T Hori, TN Sainath, R Schlüter, S Watanabe

IEEE/ACM Transactions on Audio, Speech, and Language Processing 32, 325-351, 2023

5142023Gigaspeech: An evolving, multi-domain asr corpus with 10,000 hours of transcribed audioG Chen, S Chai, G Wang, J Du, WQ Zhang, C Weng, D Su, D Povey, ...

arXiv preprint arXiv:2106.06909, 2021

5132021Audiogpt: Understanding and generating speech, music, sound, and talking headR Huang, M Li, D Yang, J Shi, X Chang, Z Ye, Y Wu, Z Hong, J Huang, ...

Proceedings of the AAAI conference on artificial intelligence 38 (21), 23802 …, 2024

4952024An analysis of environment, microphone and data simulation mismatches in robust speech recognitionE Vincent, S Watanabe, AA Nugraha, J Barker, R Marxer

Computer Speech & Language 46, 535-557, 2017

4882017Improved MVDR beamforming using single-channel mask prediction networksH Erdogan, JR Hershey, S Watanabe, MI Mandel, JL Roux

Proc. Interspeech 2016, 1981-1985, 2016

4432016End-to-end neural speaker diarization with permutation-free objectivesY Fujita, N Kanda, S Horiguchi, K Nagamatsu, S Watanabe

arXiv preprint arXiv:1909.05952, 2019

4262019

Similar Articles