Hoang Le
3 أوراق في مجموعة PaperMetrix
أوراق هذا المؤلف
-
Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
2019 · arXiv (Cornell University)
We offer an experimental benchmark and empirical study for off-policy policy evaluation (OPE) in reinforcement learning, which is a key problem in many safety critical applications. Given the increasing interest in deploying learning-based methods, there …
-
Tracking of Surface Maneuvering Targets Based on Interactive Multi-model Algorithm
2023
Trajectory filtering is one of the important target tracking problems that has received attention in recent years. An interactive multi-model algorithm (IMM) is proposed for tracking maneuvering surface vessels. The simulation of the work of …
-
Cycle Training with Semi-Supervised Domain Adaptation: Bridging Accuracy and Efficiency for Real-Time Mobile Scene Detection
2025 · arXiv (Cornell University)
Nowadays, smartphones are ubiquitous, and almost everyone owns one. At the same time, the rapid development of AI has spurred extensive research on applying deep learning techniques to image classification. However, due to the limited …