yan
Nikoyan
ยท
AI & ML interests
LLM,RL,Agent
Recent Activity
upvoted a paper 2 days ago
TCAndon-Router: Adaptive Reasoning Router for Multi-Agent Collaboration authored a paper 8 days ago
SkillEvo: Self-Renewing Evolution Gradients from Multi-Turn Interaction Feedback commentedon a paper 13 days ago
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning