Open Issues Need Help
View All on GitHub enhancement good first issue
Clean, config-driven LLM Post-Training: SFT, DPO, GRPO, On-Policy Distillation — one YAML, any topology.
Python
enhancement good first issue
Clean, config-driven LLM Post-Training: SFT, DPO, GRPO, On-Policy Distillation — one YAML, any topology.
Python
enhancement good first issue
Clean, config-driven LLM Post-Training: SFT, DPO, GRPO, On-Policy Distillation — one YAML, any topology.
Python
enhancement good first issue
Clean, config-driven LLM Post-Training: SFT, DPO, GRPO, On-Policy Distillation — one YAML, any topology.
Python
enhancement good first issue
Clean, config-driven LLM Post-Training: SFT, DPO, GRPO, On-Policy Distillation — one YAML, any topology.
Python