A Datacenter Scale Distributed Inference Serving Framework

diffusion disaggregated-serving kubernetes llm-inference omni routing-engine rust sglang tensorrt-llm vllm
42 Open Issues Need Help Last updated: Jul 28, 2026

Open Issues Need Help

View All on GitHub
enhancement good first issue language::rust dynamo-llm python frontend contribution-request

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue deployment::k8s

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue language::rust dynamo-llm

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue observability backend::trtllm approved-for-pr contribution-request

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
good first issue dynamo-runtime contribution-request External Contribution

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue backend::trtllm

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
bug good first issue dynamo-deploy deployment::k8s go deployment::nats DGDR

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue backend::sglang

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue Dynamo 0.9.0

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
bug good first issue language::go size/M deployment::k8s

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue backend::sglang

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
bug good first issue language::rust multimodal

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue backend::vllm

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
good first issue language::rust observability

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
bug good first issue language::rust frontend

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
good first issue language::rust observability frontend

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
bug good first issue frontend

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
bug good first issue language::rust frontend

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
bug enhancement good first issue backend::vllm frontend

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
good first issue

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
documentation good first issue language::python

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue language::rust dynamo-llm frontend

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
bug good first issue frontend

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

AI Summary: The `logging::tests::test_json_log_capture` Rust test is intermittently failing after the merge of PR #2406. The test passes when run in isolation but fails when executed alongside other tests, indicating a potential test isolation issue or log pollution. The failure manifests as the test observing error logs from other parts of the system instead of its expected output.

Complexity: 4/5
bug good first issue language::rust test

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm
enhancement good first issue language::rust

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

AI Summary: The task is to identify and resolve the issue of multiple versions of the `candle-core` dependency being installed within the Dynamo project. This involves standardizing on a single version across the entire workspace to improve build times. The solution likely involves updating dependency specifications and resolving any resulting compatibility conflicts.

Complexity: 4/5
enhancement good first issue

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

AI Summary: Debug a bug in the NVIDIA Dynamo distributed inference serving framework where models appear in the model list before they are ready for inference, resulting in 500 errors on inference requests. The task involves analyzing the provided reproduction steps, understanding the Dynamo architecture (as described in the project README), and identifying the source of the timing discrepancy between model registration and readiness.

Complexity: 4/5
bug good first issue

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

AI Summary: Modify the Dynamo framework's mockers to avoid downloading model weights during configuration. This involves decoupling the mockers from the ModelDeploymentCard's weight download mechanism, allowing configuration to proceed using only the tokenizer information.

Complexity: 4/5
enhancement good first issue

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

AI Summary: Improve the error handling of the Dynamo LLM serving framework's `/chat/completions` endpoint. The current error message for requests with an empty `messages` field is cryptic; it needs to be replaced with a clear and informative message indicating that the `messages` field cannot be empty.

Complexity: 3/5
bug good first issue

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm

AI Summary: Parallelize the tokenization process within the Dynamo framework's batch completion functionality using the Rayon library to improve performance. This involves identifying the appropriate sections of code for parallelization and integrating Rayon to handle the task efficiently.

Complexity: 4/5
enhancement good first issue

A Datacenter Scale Distributed Inference Serving Framework

Rust
#diffusion#disaggregated-serving#kubernetes#llm-inference#omni#routing-engine#rust#sglang#tensorrt-llm#vllm