Qi Qi
3 papers in the PaperMetrix corpus
Papers by this author
-
Variance-Reduced Off-Policy Memory-Efficient Policy Search
2020 · arXiv (Cornell University)
Off-policy policy optimization is a challenging problem in reinforcement learning (RL). The algorithms designed for this problem often suffer from high variance in their estimators, which results in poor sample efficiency, and have issues with …
-
A Hybrid Arithmetic Optimization and Golden Sine Algorithm for Solving Industrial Engineering Design Problems
2022 · Mathematics
Arithmetic Optimization Algorithm (AOA) is a physically inspired optimization algorithm that mimics arithmetic operators in mathematical calculation. Although the AOA has an acceptable exploration and exploitation ability, it also has some shortcomings such as low …
-
AndesVL Technical Report: An Efficient Mobile-side Multimodal Large Language Model
2025 · arXiv (Cornell University)
In recent years, while cloud-based MLLMs such as QwenVL, InternVL, GPT-4o, Gemini, and Claude Sonnet have demonstrated outstanding performance with enormous model sizes reaching hundreds of billions of parameters, they significantly surpass the limitations in …