Researcher profile

Ying Shan

5 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. A Hierarchical Speaker Representation Framework for One-shot Singing Voice Conversion

    2022 · Interspeech 2022

    Typically, singing voice conversion (SVC) depends on an embedding vector, extracted from either a speaker lookup table (LUT) or a speaker recognition network (SRN), to model speaker identity.However, singing contains more expressive speaker characteristics than …

  2. GPT4Tools: Teaching Large Language Model to Use Tools via Self-instruction

    2023 · arXiv (Cornell University)

    This paper aims to efficiently enable Large Language Models (LLMs) to use multimodal tools. Advanced proprietary LLMs, such as ChatGPT and GPT-4, have shown great potential for tool usage through sophisticated prompt engineering. Nevertheless, these …

  3. Enhancing the vocal range of single-speaker singing voice synthesis with melody-unsupervised pre-training

    2023 · arXiv (Cornell University)

    The single-speaker singing voice synthesis (SVS) usually underperforms at pitch values that are out of the singer's vocal range or associated with limited training samples. Based on our previous work, this work proposes a melody-unsupervised …

  4. Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?

    2025 · arXiv (Cornell University)

    Recent advances in CoT reasoning and RL post-training have been reported to enhance video reasoning capabilities of MLLMs. This progress naturally raises a question: can these models perform complex video reasoning in a manner comparable …

  5. LoRA-Gen: Specializing Large Language Model via Online LoRA Generation

    2025 · arXiv (Cornell University)

    Recent advances have highlighted the benefits of scaling language models to enhance performance across a wide range of NLP tasks. However, these approaches still face limitations in effectiveness and efficiency when applied to domain-specific tasks, …