deepdml/whisper-tiny-es-mix-norm Automatic Speech Recognition • 37.8M • Updated 23 minutes ago • 1.06k • 1
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU Paper • 2607.19191 • Published 8 days ago • 302
RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Model Paper • 2607.17977 • Published 9 days ago • 197
Discrete Diffusion Models: A Unified Framework from Tokenization to Generation Paper • 2607.13431 • Published 14 days ago • 20
X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras Paper • 2607.12993 • Published 15 days ago • 131
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 15 days ago • 227
Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation Paper • 2607.11886 • Published 16 days ago • 84
GnLOLot/MiniCPM5-1B-Claude-Opus-Fable5-Thinking-GGUF Text Generation • 1B • Updated 16 days ago • 304k • 306
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published Jun 28 • 170
Drop-Then-Recovery: How Redundant Are Vision-Language-Action Models? Paper • 2606.27755 • Published Jun 26 • 6
Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models Paper • 2606.03988 • Published Jun 3 • 126