Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper โข 2607.28227 โข Published 6 days ago โข 298
view post Post 7734 Frontier models use distillation as a step of their post-training pipelines. In 2026 it has three jobs: compress a big model into a small one, merge RL experts into a single model, and let a model teach itself.I wrote up which frontier models use each one and how: https://huggingface.co/blog/sergiopaniego/distillation-2026It pairs with Class 2 of the Training an Agent series Ben and I are doing, where we teach these techniques hands-on with TRL! See translation 3 replies ยท ๐ 14 14 ๐ฅ 7 7 โค๏ธ 3 3 + Reply
Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs Paper โข 2606.27378 โข Published May 7 โข 60
The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook Paper โข 2604.02029 โข Published Apr 2 โข 152
view article Article How ๐ค Accelerate runs very large models thanks to PyTorch sgugger โข Sep 27, 2022 โข 18