-
Taming LLMs by Scaling Learning Rates with Gradient Grouping
Paper • 2506.01049 • Published • 40 -
Switch EMA: A Free Lunch for Better Flatness and Sharpness
Paper • 2402.09240 • Published • 5 -
Unveiling the Backbone-Optimizer Coupling Bias in Visual Representation Learning
Paper • 2410.06373 • Published • 36 -
OpenMixup: Open Mixup Toolbox and Benchmark for Visual Representation Learning
Paper • 2209.04851 • Published • 3
🤝 Open to Collab
Juanxi Tian
Juanxi
AI & ML interests
Efficient AI & Gen AI
Recent Activity
upvoted a paper about 9 hours ago
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes published a dataset 2 days ago
OpenEnvisionLab/WorldEngine authored a paper 6 days ago
VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System