arxiv:2601.05167
Langlin Huang
shrango
AI & ML interests
LLM Reasoning, Machine Translation
Recent Activity
upvoted a paper 4 days ago
Hermes: Learning Contextual Reasoning Unlocks Test-Time Scaling upvoted a paper 4 days ago
MILO: Automated Harness Discovery via Orchestrated Multi-Agent Evolution upvoted a paper 4 days ago
On the Off-Policy Teacher in On-Policy Distillation