Demystifying Agent Skills: Why They Work-Until They Don't Paper • 2608.14036 • Published 7 days ago • 153
On-Policy Self-Distillation without Any Supervision Paper • 2608.06296 • Published 12 days ago • 215
DeepAgent: A General Reasoning Agent with Scalable Toolsets Paper • 2510.21618 • Published Oct 24, 2025 • 103
Harnessing Negative Signals: Reinforcement Distillation from Teacher Data for LLM Reasoning Paper • 2505.24850 • Published May 30, 2025 • 9