Running Reproduction: Adversarial Dual On-Policy Distillation from Expressive Flow-based Teacher 🎯 Explore and sync research logbook with AI agent
Running Reproduction: Self-Distillation Enables Continual Learning 🎯 Browse and sync research logbook with an AI coding agent
codemaivanngu/repro-adversarial-dual-on-policy-distillation-from-expressive-flow-based-teacher-metrics-bucket 1.1 MB
Running Reproduction: Adversarial Dual On-Policy Distillation from Expressive Flow-based Teacher 🎯 Explore and sync research logbook with AI agent
codemaivanngu/repro-adversarial-dual-on-policy-distillation-from-expressive-flow-based-teacher-artifacts 731 MB
codemaivanngu/repro-adversarial-dual-on-policy-distillation-from-expressive-flow-based-teacher-artifacts 731 MB
Running Reproduction: Self-Distillation Enables Continual Learning 🎯 Browse and sync research logbook with an AI coding agent
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Paper • 2607.21653 • Published 11 days ago • 31
Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering Paper • 2607.21848 • Published 10 days ago • 8
DataPrep-Bench: Benchmarking LLMs as Training Data Preparators Paper • 2607.20465 • Published May 19 • 54