Yue Shi
3 papers in the PaperMetrix corpus
Papers by this author
-
An Off-Policy Trust Region Policy Optimization Method With Monotonic Improvement Guarantee for Deep Reinforcement Learning
2021 · IEEE Transactions on Neural Networks and Learning Systems
In deep reinforcement learning, off-policy data help reduce on-policy interaction with the environment, and the trust region policy optimization (TRPO) method is efficient to stabilize the policy optimization procedure. In this article, we propose an …
-
A Reference Implementation for a Quantum Message Passing Interface
2023
Practical applications of quantum computing are currently limited by the number of qubits that can be set with reasonable fidelity for each system. Therefore, a distributed quantum computing system with multiple quantum computers coherently connected …
-
TriStack: Efficient Stackelberg Game-Based Offloading for Cloud-Edge-Terminal Computing
2024
Cloud-edge-terminal computing has become a promising approach to improve resource utilization and service for communication and network applications. It is crucial to rationalize task offloading and resource allocation to better satisfy quality of experience (QoE). …