SkillSKL-5A7996D8
simpo-training
Simple Preference Optimization for LLM alignment. Reference-free alternative to DPO with better performance (+6.4 points on AlpacaEval 2.0). No reference model needed, more efficient than DPO. Use for preference alignment when want simpler, faster training than DPO/PPO.
Ecosystem Scores - What Happened to it
Powered by inVerus
Verification Signals
Updated Today
Stars
0No stars yetContributors
~286Approx. contributorsLast Push
3mo agoActive maintenanceForks
0No forks yetRepository Age
5mSince 2026Issue Health
22High activityPROTOCOL WARRANT
This score reflects origin + ecosystem signals. It is not a code audit.
Skill Lineage Map
Spatial graph · creator origin → derivative skills
Tap to expand · Click creator to view profile
