🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

707 stars 216 forks 707 watchers Python Apache License 2.0
agent deepseek-v3-2 deepseek-v4 finetuning gemma3 gemma4 glm gpt-oss kimi-k2 llama llama3 llm minimax-m2 mistral openai qwen3 qwen3-6 qwen3-next vlm
38 Open Issues Need Help Last updated: Jul 13, 2026

Open Issues Need Help

View All on GitHub
bug good first issue MoE qa_rcca_done

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
Support Hy MT2 7 days ago
enhancement good first issue Feature

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
good first issue ckpt Performance ci automodel

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
good first issue community-request waiting-on-customer

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
Support MiMo-V2.5-Pro about 2 months ago
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue community-request

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
Support Falcon H1 3 months ago
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue community-request

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
YAML linter 3 months ago
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue community-request waiting-on-maintainers

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue Automation

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
S3 + DCP 4 months ago
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
bug good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
Support Skypilot 5 months ago
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue community-request external x-tencent t-algo

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
enhancement good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
good first issue Stale

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm
good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm

AI Summary: Implement deterministic job resume functionality for the NeMo AutoModel framework. This involves ensuring that resuming a training job from a saved state produces identical results to a continuous run, verifying the correct restoration of model, optimizer, dataloader, and data sampler states.

Complexity: 4/5
good first issue

🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support

Python
#agent#deepseek-v3-2#deepseek-v4#finetuning#gemma3#gemma4#glm#gpt-oss#kimi-k2#llama#llama3#llm#minimax-m2#mistral#openai#qwen3#qwen3-6#qwen3-next#vlm