Man-Lu
Lu-Man
AI & ML interests
None yet
Recent Activity
upvoted a paper 13 days ago
CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization upvoted a paper 10 months ago
A Theoretical Study on Bridging Internal Probability and
Self-Consistency for LLM Reasoning liked a Space 10 months ago
WNJXYK/RPCOrganizations
None yet