yan
Nikoyan
ยท
AI & ML interests
LLM,RL,Agent
Recent Activity
upvoted a paper about 10 hours ago
TCAndon-Router: Adaptive Reasoning Router for Multi-Agent Collaboration authored a paper 7 days ago
SkillEvo: Self-Renewing Evolution Gradients from Multi-Turn Interaction Feedback commentedon a paper 11 days ago
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning