AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.

ai autoscaler faster-whisper inference-operator k8s kubernetes llm ollama ollama-operator openai-api vllm vllm-operator whisper
4 Open Issues Need Help Last updated: Jul 20, 2026

Open Issues Need Help

View All on GitHub
enhancement good first issue

AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.

Go
#ai#autoscaler#faster-whisper#inference-operator#k8s#kubernetes#llm#ollama#ollama-operator#openai-api#vllm#vllm-operator#whisper

AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.

Go
#ai#autoscaler#faster-whisper#inference-operator#k8s#kubernetes#llm#ollama#ollama-operator#openai-api#vllm#vllm-operator#whisper
good first issue

AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.

Go
#ai#autoscaler#faster-whisper#inference-operator#k8s#kubernetes#llm#ollama#ollama-operator#openai-api#vllm#vllm-operator#whisper
good first issue

AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.

Go
#ai#autoscaler#faster-whisper#inference-operator#k8s#kubernetes#llm#ollama#ollama-operator#openai-api#vllm#vllm-operator#whisper