ملف الباحث
Yang Long
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Qualitative Measurements of Policy Discrepancy for Return-Based Deep Q-Network
2019 · IEEE Transactions on Neural Networks and Learning Systems
The deep Q-network (DQN) and return-based reinforcement learning are two promising algorithms proposed in recent years. The DQN brings advances to complex sequential decision problems, while return-based algorithms have advantages in making use of sample …
-
vMFCoOp: Towards Equilibrium on a Unified Hyperspherical Manifold for Prompting Biomedical VLMs
2025 · arXiv (Cornell University)
Recent advances in context optimization (CoOp) guided by large language model (LLM)-distilled medical semantic priors offer a scalable alternative to manual prompt engineering and full fine-tuning for adapting biomedical CLIP-based vision-language models (VLMs). However, prompt …