Open-Sourced model and data for ULTRAIF: Advancing Instruction Following from the Wild.
li sheng
bambisheng
AI & ML interests
None yet
Recent Activity
upvoted a paper 4 days ago
T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks liked a model 19 days ago
Qwen/Qwen3.8-Flash-Next liked a model about 1 month ago
Qwen/Qwen3.8-2.4T-A95B