ملف الباحث
Xiaoyu Qin
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
VoxInstruct: Expressive Human Instruction-to-Speech Generation with Unified Multilingual Codec Language Modelling
2024
Recent AIGC systems possess the capability to generate digital multimedia content based on human language instructions, such as text, image and video. However, when it comes to speech, existing methods related to human instruction-to-speech generation …