arxiv:2608.27448
Aoshining
Aoshining
AI & ML interests
None yet
Recent Activity
upvoted a paper 2 days ago
RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning upvoted a paper 20 days ago
PaperGym: Rubric-Centered Evolution for Research-Plan Generation authored a paper 23 days ago
Code-A1: Adversarial Evolving of Code LLM and Test LLM via Reinforcement LearningOrganizations
None yet