Running Reproduction: Self-Distillation Enables Continual Learning 🎯 Browse and collaborate on project logbooks
codemaivanngu/repro-adversarial-dual-on-policy-distillation-from-expressive-flow-based-teacher-metrics-bucket 1.1 MB
Running Reproduction: Adversarial Dual On-Policy Distillation from Expressive Flow-based Teacher 🎯 Browse experiment logs and sync findings with an AI agent
Running Reproduction: Adversarial Dual On-Policy Distillation from Expressive Flow-based Teacher 🎯 Browse experiment logs and sync findings with an AI agent
codemaivanngu/repro-adversarial-dual-on-policy-distillation-from-expressive-flow-based-teacher-artifacts 731 MB
codemaivanngu/repro-adversarial-dual-on-policy-distillation-from-expressive-flow-based-teacher-artifacts 731 MB
Running Reproduction: Self-Distillation Enables Continual Learning 🎯 Browse and collaborate on project logbooks
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Paper • 2607.21653 • Published 7 days ago • 29
Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering Paper • 2607.21848 • Published 6 days ago • 5
DataPrep-Bench: Benchmarking LLMs as Training Data Preparators Paper • 2607.20465 • Published May 19 • 49
Scaling Laws for Hypernetwork-Based Knowledge Injection in Large Language Models Paper • 2607.19604 • Published 8 days ago • 17