Tom Lu
eigentom
AI & ML interests
MLLM, Reinforcement Learning, Agentic RL
Recent Activity
published a dataset about 5 hours ago
eigentom/qwen35-4b-dci-rl-rewardv2-step10-bcp100-eval updated a dataset about 5 hours ago
eigentom/qwen35-4b-dci-rl-rewardv2-step10-bcp100-eval upvoted a paper 14 days ago
Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models