SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD Paper • 2607.20145 • Published 2 days ago • 45
SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD Paper • 2607.20145 • Published 2 days ago • 45
Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation Paper • 2607.12752 • Published 9 days ago • 19
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Paper • 2607.13124 • Published 10 days ago • 19
Registers Matter for Pixel-Space Diffusion Transformers Paper • 2605.16147 • Published 18 days ago • 26
Self-Improvements in Modern Agentic Systems: A Survey Paper • 2607.13104 • Published 10 days ago • 31
MetaView: Monocular Novel View Synthesis with Scale-Aware Implicit Geometry Priors Paper • 2607.12000 • Published 11 days ago • 38
GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch Paper • 2607.13960 • Published 9 days ago • 28
PolicyShiftGuard: Benchmarking and Improving Policy-Adaptive Image Guardrails Paper • 2607.05910 • Published 17 days ago • 38
Dynamic Long Context Reasoning over Compressed Memory via End-to-End Reinforcement Learning Paper • 2602.08382 • Published Feb 9 • 11
Dynamic Long Context Reasoning over Compressed Memory via End-to-End Reinforcement Learning Paper • 2602.08382 • Published Feb 9 • 11
LycheeDecode: Accelerating Long-Context LLM Inference via Hybrid-Head Sparse Decoding Paper • 2602.04541 • Published Feb 4 • 9
LycheeDecode: Accelerating Long-Context LLM Inference via Hybrid-Head Sparse Decoding Paper • 2602.04541 • Published Feb 4 • 9
KaLM-Embedding: Superior Training Data Brings A Stronger Embedding Model Paper • 2501.01028 • Published Jan 2, 2025 • 19
Stabilizing Long-term Multi-turn Reinforcement Learning with Gated Rewards Paper • 2508.10548 • Published Aug 14, 2025 • 1