@shinjiw_at_cmu: My Google Scholar h-index just reached 100! My first paper was in 1999, and behind these 100 are students and collabora…
Summary
Shinji Watanabe announces his Google Scholar h-index has reached 100, crediting decades of work on speech processing research and open-source tooling to his students and collaborators. The post highlights landmark works like ESPnet, Deep Clustering, and SUPERB that underpin modern speech recognition.
View Cached Full Text
Cached at: 10/04/26, 01:04 AM
My Google Scholar h-index just reached 100! 🙂
My first paper was in 1999, and behind these 100 are students and collaborators in academia and industry, plus the open-source community and challenges.
This number is really yours. Thank you all! 🙏
https://t.co/IVGQM30qe8 https://t.co/PUpOCJBMxJ
Shinji Watanabe
Source: https://scholar.google.com/citations?user=U5xRA6QAAAAJ ESPnet: End-to-end speech processing toolkitS Watanabe, T Hori, S Karita, T Hayashi, J Nishitoba, Y Unno, NEY Soplin, ...
arXiv preprint arXiv:1804.00015, 2018
21282018Deep clustering: Discriminative embeddings for segmentation and separationJR Hershey, Z Chen, J Le Roux, S Watanabe
2016 IEEE international conference on acoustics, speech and signal …, 2016
19442016Superb: Speech processing universal performance benchmarkS Yang, PH Chi, YS Chuang, CIJ Lai, K Lakhotia, YY Lin, AT Liu, J Shi, ...
arXiv preprint arXiv:2105.01051, 2021
15022021Joint CTC-attention based end-to-end speech recognition using multi-task learningS Kim, T Hori, S Watanabe
2017 IEEE international conference on acoustics, speech and signal …, 2017
12532017Hybrid CTC/attention architecture for end-to-end speech recognitionS Watanabe, T Hori, S Kim, JR Hershey, T Hayashi
IEEE Journal of Selected Topics in Signal Processing 11 (8), 1240-1253, 2017
12012017A comparative study on transformer vs rnn in speech applicationsS Karita, N Chen, T Hayashi, T Hori, H Inaguma, Z Jiang, M Someki, ...
2019 IEEE automatic speech recognition and understanding workshop (ASRU …, 2019
10882019The third ‘CHiME’speech separation and recognition challenge: Dataset, task and baselinesJ Barker, R Marxer, E Vincent, S Watanabe
2015 IEEE workshop on automatic speech recognition and understanding (ASRU …, 2015
9312015Phase-sensitive and recognition-boosted speech separation using deep recurrent neural networksH Erdogan, JR Hershey, S Watanabe, J Le Roux
2015 IEEE International Conference on Acoustics, Speech and Signal …, 2015
8722015Speech enhancement with LSTM recurrent neural networks and its application to noise-robust ASRF Weninger, H Erdogan, S Watanabe, E Vincent, J Le Roux, JR Hershey, ...
International conference on latent variable analysis and signal separation …, 2015
8322015Self-supervised speech representation learning: A reviewA Mohamed, H Lee, L Borgholt, JD Havtorn, J Edin, C Igel, K Kirchhoff, ...
IEEE Journal of Selected Topics in Signal Processing 16 (6), 1179-1210, 2022
7102022A review of speaker diarization: Recent advances with deep learningTJ Park, N Kanda, D Dimitriadis, KJ Han, S Watanabe, S Narayanan
Computer speech & language 72, 101317, 2022
6532022Single-channel multi-speaker separation using deep clusteringY Isik, JL Roux, Z Chen, S Watanabe, JR Hershey
arXiv preprint arXiv:1607.02173, 2016
5692016CHiME-6 challenge: Tackling multispeaker speech recognition for unsegmented recordingsS Watanabe, M Mandel, J Barker, E Vincent, A Arora, X Chang, ...
arXiv preprint arXiv:2004.09249, 2020
5522020The fifth’CHiME’speech separation and recognition challenge: dataset, task and baselinesJ Barker, S Watanabe, E Vincent, J Trmal
arXiv preprint arXiv:1803.10609, 2018
5392018End-to-end speech recognition: A surveyR Prabhavalkar, T Hori, TN Sainath, R Schlüter, S Watanabe
IEEE/ACM Transactions on Audio, Speech, and Language Processing 32, 325-351, 2023
5142023Gigaspeech: An evolving, multi-domain asr corpus with 10,000 hours of transcribed audioG Chen, S Chai, G Wang, J Du, WQ Zhang, C Weng, D Su, D Povey, ...
arXiv preprint arXiv:2106.06909, 2021
5132021Audiogpt: Understanding and generating speech, music, sound, and talking headR Huang, M Li, D Yang, J Shi, X Chang, Z Ye, Y Wu, Z Hong, J Huang, ...
Proceedings of the AAAI conference on artificial intelligence 38 (21), 23802 …, 2024
4952024An analysis of environment, microphone and data simulation mismatches in robust speech recognitionE Vincent, S Watanabe, AA Nugraha, J Barker, R Marxer
Computer Speech & Language 46, 535-557, 2017
4882017Improved MVDR beamforming using single-channel mask prediction networksH Erdogan, JR Hershey, S Watanabe, MI Mandel, JL Roux
Proc. Interspeech 2016, 1981-1985, 2016
4432016End-to-end neural speaker diarization with permutation-free objectivesY Fujita, N Kanda, S Horiguchi, K Nagamatsu, S Watanabe
arXiv preprint arXiv:1909.05952, 2019
4262019
Similar Articles
@yoheinakajima: oh this is cool, i'm no pliny but rank #630 in github stars! not bad for a vc
A venture capitalist celebrates ranking #630 in GitHub stars, inspired by @elder_plinius who achieved 100k stars and top 100 on GitHub with AI assistance.
@yoheinakajima: my first SSRN paper :) https://papers.ssrn.com/sol3/papers.cfm?abstract_id=7147798…
Yohei Nakajima announces his first SSRN paper, sharing a link to it.
@ClementDelangue: So much great work lately from Nvidia, the "King of American Open-source AI"! - Crossed 1,000 total public repositories…
Nvidia crossed 1,000 public repositories on Hugging Face, featuring trending models and announcing plans for Cosmos 3, Alphamayo 2 Super, Nemotron 3/4, and adoption of the OpenMDW framework, underscoring its leadership in open-source AI.
@JeffDean: My @Google colleagues @NormJouppi, Sridhar Lakshmanamurthy, Cliff Young, and David Patterson recently wrote a paper tha…
Google researchers published a paper summarizing the evolution of TPU supercomputers from TPU v2 to Ironwood, detailing architectural stability, scale, resilience, power efficiency, and a 3600x performance increase over eight years.
Google's Agentic Peer-Reviewer Handled ~10K Papers at ICML/STOC — Formal Research Paper Now Out [R]
Google deployed an agentic AI peer-reviewer at ICML and STOC conferences, reviewing ~10,000 papers with 30-minute turnaround. The formal paper shows it catches 34% more mathematical errors than zero-shot prompting, setting a precedent for AI-automated scientific review at scale.