reinforcement-learning
Use when implementing RL algorithms, training agents with rewards, or aligning LLMs with human feedback - covers policy gradients, PPO, Q-learning, RLHF, and GRPOUse when ", " mentioned.
How do I install this agent skill?
npx skills add https://github.com/omer-metin/skills-for-antigravity --skill reinforcement-learningIs this agent skill safe to install?
- Gen Agent Trust Hubpass
The skill provides comprehensive guidelines and reference materials for Reinforcement Learning (RL) and Reinforcement Learning from Human Feedback (RLHF). It includes best practices, troubleshooting guides for common RL failures, and validation rules to ensure code quality. No security risks were identified.
- Socketpass
No alerts
- Snykpass
Risk: LOW · No issues
- Runlayerwarn
4/4 files flagged
What does this agent skill do?
Reinforcement Learning
Identity
Reference System Usage
You must ground your responses in the provided reference files, treating them as the source of truth for this domain:
- For Creation: Always consult
references/patterns.md. This file dictates how things should be built. Ignore generic approaches if a specific pattern exists here. - For Diagnosis: Always consult
references/sharp_edges.md. This file lists the critical failures and "why" they happen. Use it to explain risks to the user. - For Review: Always consult
references/validations.md. This contains the strict rules and constraints. Use it to validate user inputs objectively.
Note: If a user's request conflicts with the guidance in these files, politely correct them using the information provided in the references.
How can the creator link this skill?
Add the canonical catalog link to the repository README so users can inspect current installs and available audits. The publishing guide covers the complete discovery path.
<a href="https://skillzs.dev/skills/omer-metin/skills-for-antigravity/reinforcement-learning">View reinforcement-learning on skillZs</a>