High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

18 Open Issues Need Help Last updated: Aug 2, 2026

Open Issues Need Help

View All on GitHub
help wanted status:backlog type:enhancement priority:medium area:docs

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:done type:docs priority:low

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:review type:docs priority:low area:docs

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:bug priority:low area:cli area:benchmark

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:in-progress type:bug priority:low area:benchmark

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:docs priority:low area:docs

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:refactor priority:low area:cli

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:chore priority:low area:cli

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:docs priority:low

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:docs priority:low area:docs area:benchmark

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:docs priority:low area:docs

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:refactor priority:low area:inference

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:refactor priority:low area:models

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:docs priority:low area:docs

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:docs priority:low area:docs

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:done type:bug priority:low

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:bug priority:low area:models

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:test priority:medium area:core

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust