Matthew Wiesner
Johns Hopkins University
48 Papers
133 Citations
Matthew Wiesner is an academic researcher from Johns Hopkins University. The author has contributed to research in topics: Computer science & Language model. The author has an hindex of 10, co-authored 32 publications. Previous affiliations of Matthew Wiesner include McGill University.
Chat about Author
Papers
ESPNet: End-to-end speech processing toolkit
Shinji Watanabe,Takaaki Hori,Shigeki Karita,Tomoki Hayashi,Jiro Nishitoba,Yuya Unno,Nelson Yalta,Jahn Heymann,Matthew Wiesner,Nanxin Chen,Adithya Renduchintala,Tsubasa Ochiai +11 more
- 30 Mar 2018
TL;DR: In this article, a new open source platform for end-to-end speech processing named ESPnet is introduced, which mainly focuses on automatic speech recognition (ASR), and adopts widely used dynamic neural network toolkits, Chainer and PyTorch, as a main deep learning engine.
1.3K
•Posted Content
ESPnet: End-to-End Speech Processing Toolkit
Shinji Watanabe,Takaaki Hori,Shigeki Karita,Tomoki Hayashi,Jiro Nishitoba,Yuya Unno,Nelson Yalta,Jahn Heymann,Matthew Wiesner,Nanxin Chen,Adithya Renduchintala,Tsubasa Ochiai +11 more
TL;DR: A major architecture of this software platform, several important functionalities, which differentiate ESPnet from other open source ASR toolkits, and experimental results with major ASR benchmarks are explained.
1K
Multilingual Sequence-to-Sequence Speech Recognition: Architecture, Transfer Learning, and Language Modeling
Jaejin Cho,Murali Karthick Baskar,Ruizhi Li,Matthew Wiesner,Sri Harish Mallidi,Nelson Yalta,Martin Karafiat,Shinji Watanabe,Takaaki Hori +8 more
- 04 Oct 2018
TL;DR: Experimental results show that the transfer learning approach from the multilingual model shows substantial gains over monolingual models across all 4 BABEL languages.
143
Massively multilingual adversarial speech recognition
Oliver Adams,Matthew Wiesner,Shinji Watanabe,David Yarowsky +3 more
- 01 Jun 2019
TL;DR: In this article, the relative importance of similarity between the target and pretraining languages along the dimensions of phonetics, phonology, language family, geographical location, and orthography is discussed.
87
The Multilingual TEDx Corpus for Speech Recognition and Translation
Elizabeth Salesky,Matthew Wiesner,Jacob Bremerman,Roldano Cattoni,Matteo Negri,Marco Turchi,Douglas W. Oard,Matt Post +7 more
- 02 Feb 2021
TL;DR: The Multilingual TEDx corpus as mentioned in this paper is built to support speech recognition (ASR) and speech translation (ST) research across many non-English source languages and includes audio recordings from TEDx talks in 8 source languages.
73