Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
zkj
JokerJan
11
3
Follow
Datawitch-Programmer's profile picture
jinzhuoran's profile picture
refkxh's profile picture
3 followers
·
0 following
GaryStack
AI & ML interests
None yet
Recent Activity
upvoted
a
paper
5 days ago
The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation
upvoted
a
paper
about 1 month ago
Why Multi-Step Tool-Use Reinforcement Learning Collapses and How Supervisory Signals Fix It
authored
a paper
about 2 months ago
Omni-Reward: Towards Generalist Omni-Modal Reward Modeling with Free-Form Preferences
View all activity
Organizations
None yet
JokerJan
's datasets
2
Sort: Recently updated
JokerJan/MMR-VBench
Viewer
•
Updated
Jul 1, 2025
•
1.26k
•
578
•
17
JokerJan/rl_think
Updated
Jun 3, 2025
•
13