Preference Alignment skills
Free agent skills tagged preference alignment, ready to install into any SKILL.md-compatible agent.
3 skills
Fine-Tuning with TRL
davila7
Align language models with human preferences using reinforcement learning.
developmentintermediatePython · Shell30.2k repo
TRL Fine-Tuning
nousresearch
Align language models with human preferences using TRL.
developmentintermediatePython · Shell228.5k repo
Simple Preference Optimization
nousresearch
Optimize preferences without reference models.
developmentintermediatePython · Shell228.5k repo
