Roger Yang
yangroger
AI & ML interests
{*}
Recent Activity
liked a model 2 days ago
sentence-transformers/all-MiniLM-L6-v2 upvoted a collection 2 days ago
NVIDIA Nemotron v3 reacted to sergiopaniego's post with 🤗 6 days ago
join us next Tuesday, July 28, for Class 3 of the Training Agents live series!
we'll dive into reinforcement learning for agent training, covering the intuition behind GRPO, how it works, and how to implement it in TRL with practical, e2e examples
see you there ðŸ¤
live: https://www.youtube.com/live/ztdTed5egrM
> in case you missed class 1:
https://x.com/SergioPaniego/status/2069382207618379813
> and in case you missed class 2: https://x.com/SergioPaniego/status/2075180665184686187Organizations
None yet