Clean, config-driven LLM Post-Training: SFT, DPO, GRPO, On-Policy Distillation — one YAML, any topology.

6 stars 5 forks 6 watchers Python Apache License 2.0
5 Open Issues Need Help Last updated: Sep 15, 2026

Open Issues Need Help

View All on GitHub

Clean, config-driven LLM Post-Training: SFT, DPO, GRPO, On-Policy Distillation — one YAML, any topology.

Python
enhancement good first issue

Clean, config-driven LLM Post-Training: SFT, DPO, GRPO, On-Policy Distillation — one YAML, any topology.

Python

Clean, config-driven LLM Post-Training: SFT, DPO, GRPO, On-Policy Distillation — one YAML, any topology.

Python

Clean, config-driven LLM Post-Training: SFT, DPO, GRPO, On-Policy Distillation — one YAML, any topology.

Python

Clean, config-driven LLM Post-Training: SFT, DPO, GRPO, On-Policy Distillation — one YAML, any topology.

Python