arxiv:2608.21500
Yibo Peng
pybbb
AI & ML interests
None yet
Recent Activity
upvoted a paper about 7 hours ago
Improving Test-Time Scaling with Adaptive Looped Transformers upvoted a paper 6 days ago
JEV-as-a-Judge: Accept When Confident, Escalate When Unsure upvoted a paper 7 days ago
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses