SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Paper • 2607.26784 • Published 2 days ago • 20
HumanCLAW: Can Vision-Language Models Act Through a Body? Paper • 2607.27180 • Published 2 days ago • 67
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM Paper • 2607.27205 • Published 2 days ago • 122
A New Role for Relevance: Guiding Corpus Interaction in Agentic Search Paper • 2607.24223 • Published 4 days ago • 88
Pass the Baton: Trajectory-Relayed On-Policy Distillation Paper • 2607.26057 • Published 3 days ago • 29
view article Article The Engineering Handbook for GRPO + LoRA with Verl: Training Qwen2.5 on Multi-GPU Weyaxi • Jan 2 • 24
Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation Paper • 2607.24731 • Published 4 days ago • 74
From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search Paper • 2607.24280 • Published 4 days ago • 81
Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems Paper • 2607.21503 • Published 8 days ago • 24
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Paper • 2607.21653 • Published 9 days ago • 30
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Paper • 2607.22529 • Published 7 days ago • 47
Running Featured 81 QED-Nano: Teaching a Tiny Model to Prove Hard Theorems 📝 81 Who needs 1T parameters? Olympiad proofs with a 4B model
WTF GENIUS PAPERS Collection Papers that made me appreciate my major and my life a little more. obs=Observation, innov=Innovation. Most papers are abt improving tiny models. • 237 items • Updated about 9 hours ago • 55
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 8 days ago • 149
SWE-Pruner Pro: The Coder LLM Already Knows What to Prune Paper • 2607.18213 • Published 11 days ago • 78