Running Reproduction: Benchmarking the Scientific Mind: A Pathology-Derived Biomedical VQA Benchmark for Complex Scientific Reasoning 🎯 Collaborate with an AI agent through an interactive logbook
Running Reproduction: From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model 🎯 Collaborate with an AI agent to view and update a research logbook
Running Reproduction: Artificial Hippocampus Networks for Efficient Long-Context Modeling 🧠 Explore research logbook and sync findings with an AI agent
Running Reproduction: Olaf-World: Orienting Latent Actions for Video World Modeling 🎯 Browse experiment logs and collaborate with an AI agent
Running Reproduction: VideoLoom: A Video Large Language Model for Joint Spatial-Temporal Understanding 🎯 Explore video model insights with an interactive logbook
Running Reproduction: OmniSIFT: Modality-Asymmetric Token Compression for Efficient Omni-modal Large Language Models 🎯 Explore experiment logs and sync AI agent notes
Running Reproduction: Synergistic Intra- and Cross-Layer Regularization Losses for MoE Expert Specialization 🎯 Explore research logs and collaborate with an AI agent
Running Reproduction: DTop-p MoE: Sparsity-Controlled Dynamic Top-p MoE for Foundation Model Pre-training 🎯 Explore experiment logs and sync with your coding agent
Running Reproduction: FloorplanQA: A Benchmark for Spatial Reasoning in LLMs using Structured Representations 🎯 Explore FloorplanQA benchmark logbook and collaborate with an agent
Running Reproduction: From Prompts to Tokens: Internalizing Causal Supervision in Vision-Language Model for Multi-Image Causal Reasoning 🎯 Collaborate with an AI agent to view and update your logbook
Running 1 Reproduction: On the Epistemic Uncertainty of Overparametrized Neural Networks 🎯 Explore and sync experiment logbooks with an AI agent
Running Reproduction: Hermes: An Evidence-Driven Agentic Framework for Trustworthy and Explainable AI-Generated Video Detection 🎯 Browse research logbook and collaborate with a coding agent
Running Reproduction: Envisioning Beyond the Few: Disentangled Semantics and Primitives for Few-Shot Atypical Layout-to-Image Generation 🎯 Explore and collaborate on research logbooks online
Running Reproduction: ProtoVAR: Efficient Dataset Distillation via Prototype-Guided Visual Autoregressive Modeling 🎯 Explore and update a collaborative research logbook
Running Reproduction: Discrete Survival Knowledge Distillation for Competing Risks Analysis 🎯 Explore a research logbook and sync findings with an AI agent
Running Reproduction: The Crowded Embedding Space: A Mean-Field Mechanism for Emergent Marginalization in Retrieval-Augmented Agents 🎯 Collaborate with your coding agent using a shared logbook
Running Reproduction: Multilingual Safety Alignment Via Sparse Weight Editing 🎯 Explore a research logbook and sync with an AI agent
Running Reproduction: GePBench: Evaluating Fundamental Geometric Perception for Multimodal Large Language Models 🎯 Explore and sync experiment logs with an AI coding agent
Running Reproduction: Self-Refining Video Sampling 🎯 Collaborate with an AI agent to maintain a synced experiment logbook
Running Reproduction: DenseSteer: Steering Small Language Models towards Dense Math Reasoning 🎯 Collaborate on research logs with AI assistance
Running Reproduction: CVSearch: Empowering Multimodal LLMs with Cognitive Visual Search for High-Resolution Image Perception 🎯 Explore and share visual search logbook with AI agents