arxiv:2606.14885
Tom Lu
eigentom
AI & ML interests
MLLM, Reinforcement Learning, Agentic RL
Recent Activity
published a dataset about 2 hours ago
eigentom/qwen35-4b-dci-rl-rewardv2-step10-bcp100-eval updated a dataset about 3 hours ago
eigentom/qwen35-4b-dci-rl-rewardv2-step10-bcp100-eval upvoted a paper 14 days ago
Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models