arxiv:2602.14492
🤝 Open to Collab
Jiahao Yuan
Jhcircle
AI & ML interests
None yet
Recent Activity
upvoted a paper 3 days ago
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning liked a model 11 days ago
moonshotai/Kimi-K3 upvoted a paper 24 days ago
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement LearningOrganizations
None yet