๐ฏ
grpo-rl-training
๐ฏSkillfrom orchestra-research/ai-research-skills
What it does
|A skill from the AI Research Engineering Skills library that teaches AI coding agents how to implement GRPO (Group Relative Policy Optimization) for reinforcement learning training of language models.
Same repository
orchestra-research/ai-research-skills(158 items)
grpo-rl-training
Installation
Vibe Index InstallInstalls to .claude/skills/
npx vibeindex add orchestra-research/ai-research-skills --skill grpo-rl-trainingskills.sh Installโ Installs to .agents/skills/
npx skills add orchestra-research/ai-research-skills --skill grpo-rl-trainingManual InstallCopy SKILL.md content and save to the path below
~/.claude/skills/grpo-rl-training/SKILL.mdSKILL.md
279Installs
10,466
-
Last UpdatedJun 16, 2026
More from this repository10
๐๐๐๐๐๐๐๐๐๐
distributed-training๐Plugin
Plugin
rag๐Plugin
Plugin
post-training๐Plugin
Plugin
mlops๐Plugin
Plugin
inference-serving๐Plugin
Plugin
tokenization๐Plugin
A tokenization skill from the AI Research Engineering Skills Library, which offers 83 skills across 20 categories covering model architecture, fine-tuning, inference, and other AI research areas.
ml-paper-writing๐Plugin
AI research skill for writing publication-ready ML papers for top conferences (NeurIPS, ICML, ICLR, ACL, AAAI, COLM) with LaTeX templates and citation verification.
optimization๐Plugin
Plugin
observability๐Plugin
Plugin
data-processing๐Plugin
Plugin