Roger Yang
yangroger
AI & ML interests
{*}
Recent Activity
reacted to sergiopaniego's post with 🤗 about 20 hours ago
join us next Tuesday, July 28, for Class 3 of the Training Agents live series!
we'll dive into reinforcement learning for agent training, covering the intuition behind GRPO, how it works, and how to implement it in TRL with practical, e2e examples
see you there ðŸ¤
live: https://www.youtube.com/live/ztdTed5egrM
> in case you missed class 1:
https://x.com/SergioPaniego/status/2069382207618379813
> and in case you missed class 2: https://x.com/SergioPaniego/status/2075180665184686187 liked a Space 5 days ago
huggingface/number-tokenization-blog liked a model 5 days ago
google/gemma-4-12BOrganizations
None yet