arxiv:2607.28568
Nick Yang
RadioBlue
AI & ML interests
None yet
Recent Activity
upvoted a paper about 13 hours ago
Improving Test-Time Scaling with Adaptive Looped Transformers upvoted a paper 19 days ago
T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks upvoted a paper 26 days ago
Rethinking On-Policy Distillation of Large Language Models II: One Training Example