Researcher profile

Qi Qi

3 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. Variance-Reduced Off-Policy Memory-Efficient Policy Search

    2020 · arXiv (Cornell University)

    Off-policy policy optimization is a challenging problem in reinforcement learning (RL). The algorithms designed for this problem often suffer from high variance in their estimators, which results in poor sample efficiency, and have issues with …

  2. A Hybrid Arithmetic Optimization and Golden Sine Algorithm for Solving Industrial Engineering Design Problems

    2022 · Mathematics

    Arithmetic Optimization Algorithm (AOA) is a physically inspired optimization algorithm that mimics arithmetic operators in mathematical calculation. Although the AOA has an acceptable exploration and exploitation ability, it also has some shortcomings such as low …

  3. AndesVL Technical Report: An Efficient Mobile-side Multimodal Large Language Model

    2025 · arXiv (Cornell University)

    In recent years, while cloud-based MLLMs such as QwenVL, InternVL, GPT-4o, Gemini, and Claude Sonnet have demonstrated outstanding performance with enormous model sizes reaching hundreds of billions of parameters, they significantly surpass the limitations in …