Keep It InMind: Benchmarking the Implicit-Association Blind Spot in Agent Memory Paper • 2607.24368 • Published 7 days ago • 31
OmniDelta: Skill-Driven Budget Allocation for Token Compression in OmniLLMs Paper • 2607.25669 • Published 6 days ago • 8
MODUS: Decoder-Only Any-to-Any Modeling of Diverse Modalities Paper • 2607.25948 • Published 6 days ago • 15
Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model Paper • 2607.24904 • Published 7 days ago • 32
WorldDiT: A Unified Diffusion Architecture for World and Action Modeling Paper • 2607.23909 • Published 7 days ago • 7
A Vocabulary for Multi-Agent Automated Research Systems Paper • 2607.22682 • Published 21 days ago • 4
UltraViT: Latency-Optimized On-device Vision Encoder for Large Vision-Language Models Paper • 2607.23373 • Published 9 days ago • 6
ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Paper • 2607.24743 • Published 7 days ago • 11
OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation Paper • 2607.23855 • Published 8 days ago • 26
Sol-Attn: Accelerating Video Generation Inference via On-the-Fly Attention Sparsification Paper • 2607.24027 • Published 7 days ago • 35
StateAct: Program State, before Pixels, for Long-Horizon Computer-Use Agents Paper • 2607.22798 • Published 10 days ago • 60
Progress Reward Modeling for Robotic Learning: A Comprehensive Survey Paper • 2607.21655 • Published 12 days ago • 192
From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search Paper • 2607.24280 • Published 7 days ago • 81
JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Paper • 2607.23588 • Published 8 days ago • 122
Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making Paper • 2607.14277 • Published 19 days ago • 10
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Paper • 2607.22529 • Published 10 days ago • 47
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 11 days ago • 150
Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers Paper • 2607.21594 • Published 11 days ago • 16