Skip to main content

grpo-finetuning

Implement GRPO (Group Relative Policy Optimization) fine-tuning for vision-language models on small datasets. Use when SFT underperforms or training data is limited (<1000 examples).

Jump to install

Source facts

Repository
aws-solutions-library-samples/guidance-for-claude-code-with-amazon-bedrock
Last source activity
January 27, 2026 at 18:47
Detected SKILL.md language
English
Stars
396
Forks
154

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.