Researcher profile
Jian Hu
2 papers in the PaperMetrix corpus
Publications
Papers by this author
-
Aligning Language Models with Offline Learning from Human Feedback
2023 · arXiv (Cornell University)
Learning from human preferences is crucial for language models (LMs) to effectively cater to human needs and societal values. Previous research has made notable progress by leveraging human feedback to follow instructions. However, these approaches …
-
Policy Optimization and Multi-agent Reinforcement Learning for Mean-variance Team Stochastic Games
2025 · arXiv (Cornell University)
We study a long-run mean-variance team stochastic game (MV-TSG), where each agent shares a common mean-variance objective for the system and takes actions independently to maximize it. MV-TSG has two main challenges. First, the variance …