Open Issues Need Help
View All on GitHub Request: Publish helm charts as OCI artifacts 17 days ago
enhancement good first issue
AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.
Go
#ai#autoscaler#faster-whisper#inference-operator#k8s#kubernetes#llm#ollama#ollama-operator#openai-api#vllm#vllm-operator#whisper
good first issue
AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.
Go
#ai#autoscaler#faster-whisper#inference-operator#k8s#kubernetes#llm#ollama#ollama-operator#openai-api#vllm#vllm-operator#whisper
expose pvc size as parameter 6 months ago
good first issue
AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.
Go
#ai#autoscaler#faster-whisper#inference-operator#k8s#kubernetes#llm#ollama#ollama-operator#openai-api#vllm#vllm-operator#whisper
Vllm container version and oss models 10 months ago
good first issue
AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.
Go
#ai#autoscaler#faster-whisper#inference-operator#k8s#kubernetes#llm#ollama#ollama-operator#openai-api#vllm#vllm-operator#whisper