Fan Yu
5 أوراق في مجموعة PaperMetrix
أوراق هذا المؤلف
-
WeNet: Production oriented Streaming and Non-streaming End-to-End Speech Recognition Toolkit
2021 · arXiv (Cornell University)
In this paper, we propose an open source, production first, and production ready speech recognition toolkit called WeNet in which a new two-pass approach is implemented to unify streaming and non-streaming end-to-end (E2E) speech recognition …
-
A GPU-specialized Inference Parameter Server for Large-Scale Deep Recommendation Models
2022
Recommendation systems are of crucial importance for a variety of modern apps and web services, such as news feeds, social networks, e-commerce, search, etc. To achieve peak prediction accuracy, modern recommendation models combine deep learning …
-
A Comparative Study on Speaker-attributed Automatic Speech Recognition in Multi-party Meetings
2022 · arXiv (Cornell University)
In this paper, we conduct a comparative study on speaker-attributed automatic speech recognition (SA-ASR) in the multi-party meeting scenario, a topic with increasing attention in meeting rich transcription. Specifically, three approaches are evaluated in this …
-
LRTD: A Low-rank Transformer with Dynamic Depth and Width for Speech Recognition
2022 · 2022 International Joint Conference on Neural Networks (IJCNN)
Though Transformer-based models have achieved great success in the automatic speech recognition (ASR) field, they are generally resource-hungry and computation-intensive which makes them difficult to deploy in resource-restricted devices. In this paper, we propose LRTD, …
-
MindSpore Quantum: A User-Friendly, High-Performance, and AI-Compatible Quantum Computing Framework
2024 · arXiv (Cornell University)
We introduce MindSpore Quantum, a pioneering hybrid quantum-classical framework with a primary focus on the design and implementation of noisy intermediate-scale quantum (NISQ) algorithms. Leveraging the robust support of MindSpore, an advanced open-source deep learning …