Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

50 stars 6 forks 50 watchers Shell Apache License 2.0
cuda dgx dgx-spark docker docker-compose gb10 generative-ai inference llama local-llm mlops model-serving self-hosted
6 Open Issues Need Help Last updated: Jul 22, 2026

Open Issues Need Help

View All on GitHub
help wanted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted
bug help wanted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted
enhancement help wanted

Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for now and single-spark. For the not-so-rich buddies. If you want latest/in-testing, look at the branches

Shell
#cuda#dgx#dgx-spark#docker#docker-compose#gb10#generative-ai#inference#llama#local-llm#mlops#model-serving#self-hosted