Nan Jiang
4 أوراق في مجموعة PaperMetrix
أوراق هذا المؤلف
-
Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
2019 · arXiv (Cornell University)
We offer an experimental benchmark and empirical study for off-policy policy evaluation (OPE) in reinforcement learning, which is a key problem in many safety critical applications. Given the increasing interest in deploying learning-based methods, there …
-
Progressive privacy-preserving batch retrieval of lung CT image sequences based on edge-cloud collaborative computation
2022 · PLoS ONE
BACKGROUND: A computer tomography image (CI) sequence can be regarded as a time-series data that is composed of a great deal of nearby and similar CIs. Since the computational and I/O costs of similarity measure, …
-
Word Embeddings Are Steers for Language Models
2023 · arXiv (Cornell University)
Language models (LMs) automatically learn word embeddings during pre-training on language corpora. Although word embeddings are usually interpreted as feature vectors for individual words, their roles in language model generation remain underexplored. In this work, …
-
vASP: Full VM Life-cycle Protection Based on Active Security Processor Architecture
2024
Cloud computing has been applied on a large scale due to its competitive advantages. However, the introduction of virtualization brings new risks, which can come from within the VM and the host. Due to the …