ovachiever/droid-tings

fine-tuning-with-trl

Fine-tune LLMs using reinforcement learning with TRL - SFT for instruction tuning, DPO for preference alignment, PPO/GRPO for reward optimization, and reward…

First seen Jan 20, 2026

Installation

$ npx skills add https://github.com/ovachiever/droid-tings

Similar popular skills

Related neighbors and high-traction skills in the same topics — useful to compare before installing.

Also in this package

Other skills from ovachiever/droid-tings · top by installs.

npx skills add https://github.com/ovachiever/droid-tings

Browse all from ovachiever/droid-tings

More details

Agent compatibility

Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.

Claude Code Not declared
Cursor Not declared
Codex Not declared
GitHub Copilot Not declared
Windsurf Not declared
Gemini CLI Not declared
Cline Not declared
OpenCode Not declared

Repository health

Stars 52
License LICENSE
Default branch master
Open issues 1
Status Active

History

  1. First seen on skills.sh
  2. First recorded snapshot · 50 installs