deepdml/whisper-tiny-es-mix-norm Automatic Speech Recognition • 37.8M • Updated about 1 hour ago • 952 • 1
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU Paper • 2607.19191 • Published 6 days ago • 298
RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Model Paper • 2607.17977 • Published 8 days ago • 197
Discrete Diffusion Models: A Unified Framework from Tokenization to Generation Paper • 2607.13431 • Published 13 days ago • 20
X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras Paper • 2607.12993 • Published 14 days ago • 131
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 14 days ago • 226
Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation Paper • 2607.11886 • Published 15 days ago • 84
GnLOLot/MiniCPM5-1B-Claude-Opus-Fable5-Thinking-GGUF Text Generation • 1B • Updated 14 days ago • 291k • 305
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published 30 days ago • 170
Drop-Then-Recovery: How Redundant Are Vision-Language-Action Models? Paper • 2606.27755 • Published Jun 26 • 6
Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models Paper • 2606.03988 • Published Jun 3 • 126