Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

50 stars 6 forks 50 watchers Shell Apache License 2.0
cuda dgx dgx-spark docker docker-compose gb10 generative-ai inference llama local-llm mlops model-serving self-hosted
6 Open Issues Need Help Last updated: Jul 22, 2026

Open Issues Need Help

View All on GitHub
Dependency Dashboard about 2 months ago
help wanted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted
bug help wanted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted
enhancement help wanted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted