arxiv:2606.10479
Haoran Zhang
zzzhr97
AI & ML interests
Lange Language Models, Large Reasoning Models
Recent Activity
upvoted a paper 2 days ago
Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening upvoted a paper about 1 month ago
Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives upvoted a paper 2 months ago
xHC: Expanded Hyper-Connections